Digna Legi Latin for “worth reading”. One reader’s scored index.211 pieces · 70 sources · updated Tue & Fri

AI40

Pieces about AI, gathered from every area. Each one also appears on its own subject page.

16 in the eighties24 in the seventies
The eightiesworth the time16 pieces
88

Beyond vibe checks: A PM’s complete guide to evals

Explains why evaluation skill is becoming essential for building reliable AI products.

substack.comfrom the editor’s archiveApr 2026
88

GenRec: Towards LLM-Native Recommendation at Netflix

Explains Netflix's shift from hand-crafted feature models to an LLM-native recommendation architecture.

Netflix Tech BlogJul 2026
88

Why the Legendary Erdős Problems Are Falling to AI

Reports that an AI model found a counterexample disproving a 1946 Erdős conjecture.

Quanta MagazineAug 2026
86

What's the best programming language for coding agents?

Scrutinizes the widely cited claim that dynamic languages are more token-efficient for coding agents.

Dan LuuAug 2026
85

How I Use Claude Code | Boris Tane

Describes a research-plan-implement workflow that withholds coding until a written plan is approved.

link.mail.beehiiv.comfrom the editor’s archiveApr 2026
85

Is AI Reasoning Right for the Wrong Reasons?

Asks whether AI 'reasoning' models reach correct answers through fundamentally flawed processes.

Quanta MagazineJul 2026
85

LLMs are (still) mostly powered by imitative learning, not RL

Argues that LLM capabilities stem mainly from imitative learning rather than reinforcement learning.

LessWrong (küratörlü)Aug 2026
84

Build a self-improving AI PM OS with Claude Code

Explains how to build a self-improving PM workflow with Claude Code's skills and CLAUDE.md router pattern.

tracking.tldrnewsletter.comfrom the editor’s archiveMay 2026
84

Stateless MCP has recaptured my interest (and inspired mcp-explorer and datasette-mcp)

Explains the design rationale behind MCP 2.0's move to a stateless protocol.

Simon WillisonAug 2026
82

Agentic Code Quality

Examines how AI coding agents strain code review practices built for human-scale output.

Addy OsmaniAug 2026
82

Beyond the Prompt: Claude Code

Walks through Claude Code's anatomy: the .claude directory, CLAUDE.md conventions, skills, and subagents.

arps18.github.iofrom the editor’s archiveJun 2026
80

How Cloudflare enforces engineering standards using AI

Reports measured results from an AI code reviewer that flagged violations and blocked merges over four months.

Cloudflare BlogAug 2026
80

LLM Security Basics: The Full Threat Model

Maps the full threat model for LLM security, using a 2025 Microsoft 365 Copilot data-exfiltration case.

ByteByteGoAug 2026
80

Stop Prompting AI. Start Directing It

Outlines four research-based ways analytical work can direct AI to generate surprising insights.

MIT Sloan Management ReviewAug 2026
80

Ten advances in mathematics and theoretical computer science

Compiles ten concrete results AI models have produced in mathematics and theoretical computer science.

Simon WillisonAug 2026
80

The Bruno Method — Engineering in the Compiler Age

Argues for making agent-driven engineering work falsifiable and resumable, via a compiler-age analogy.

softwareleadweekly.us6.list-manage.comfrom the editor’s archiveMay 2026
The seventiessolid, narrower24 pieces
79

How to Build a Compounding OS

Explains why teams that treat AI as shared infrastructure see bigger productivity gains than solo hackers.

Sachin RekhiAug 2026
79

Human judgment doesn't leave the software factory. It relocates.

Argues that human judgment doesn't disappear from automated software production, it moves elsewhere.

Addy OsmaniAug 2026
78

Now we have a timeline of the OpenAI accidental attack against Hugging Face

Lays out the full timeline of OpenAI's accidental agent-driven attack on Hugging Face.

Simon WillisonAug 2026
78

Why An LLM’s Memory Gets Expensive and How to Fix It

Breaks down why serving LLM memory gets costly at scale and how engineering teams can reduce it.

ByteByteGoAug 2026
77

Agentic test processes, LLM benchmarks, and other notes on agentic coding from Galapagos Island

Offers a skeptical, evidence-based account of using AI agents for coding tasks over several months.

Dan LuuJul 2026
77

Without a theory of intelligence

Argues that new tools, not just ideas, have driven history's biggest leaps in scientific understanding.

Kevin KellyAug 2026
76

Critical Thinking during the age of AI

Applies a who/what/where/when/why/how framework to critical thinking in the age of AI.

open.substack.comfrom the editor’s archiveJan 2026
76

Exercises in benchmarking and evals, part 7: DeepSWE, Senior SWE-Bench, napkin math, and winter tires

Works through benchmarking and evaluation puzzles, including SWE-Bench napkin math and DeepSWE.

Dan LuuJul 2026
76

How MCP Works

Explains in depth how the Model Context Protocol works under the hood.

newsletter.systemdesign.onefrom the editor’s archiveDec 2025
75

Open Questions On Open Weights

Weighs unresolved questions about the risks and benefits of releasing AI model weights openly.

Astral Codex TenAug 2026
75

Slop Readers

Argues that AI-generated writing with no communicative intent produces prose nobody truly wants to read.

PutanumonitMay 2026
75

Software Engineering fundamentals matter more

Argues that software engineering fundamentals matter more, not less, in the age of AI coding tools.

Hacker News (100+ puan)Aug 2026
74

Don’t Outsource Your Thinking

Argues for guarding independent judgment against the pull to offload thinking onto AI tools.

teltam.github.iofrom the editor’s archiveJan 2026
74

Vertical Small LLMs: A Deep Dive

Walks through designing a vertical small LLM system to classify and route support tickets end to end.

System Design NewsletterAug 2026
72

Agentic development basics

Introduces the fundamentals of agentic, AI-assisted software development practices.

link.mail.beehiiv.comfrom the editor’s archiveJan 2026
72

How Big Models Teach Small Models to Be Smart

Explains how large models transfer knowledge to smaller ones through the mechanism of distillation.

ByteByteGoAug 2026
72

I told Claude Code to build me an executive assistant. This is what my work as CTO looks like now

Recounts firsthand how building an AI executive assistant with Claude Code reshaped a CTO's daily work.

leadershipintech.comfrom the editor’s archiveMar 2026
72

Incident Report: unsanctioned agent behaviour during cyber testing

Reports on an AI agent that attacked unrelated companies during a government cyber evaluation.

Simon WillisonAug 2026
72

Patterns and problems in emerging multi-agent systems

Surveys recurring architectural patterns and failure modes in emerging multi-agent AI systems.

Hacker News (100+ puan)Aug 2026
72

Prediction: AI will make formal verification go mainstream

Argues that AI will push formal verification from a fringe pursuit into mainstream engineering.

Martin KleppmannDec 2025
72

Your File System Is Already a Graph Database

Argues the file system already behaves like a graph database for organizing personal knowledge.

producthabits.us14.list-manage.comfrom the editor’s archiveApr 2026
70

DAU, WAU and MAU Are the New Lighthouse Metric in B2B + AI. Harvey’s a Great Case Study.

Argues DAU, WAU and MAU are becoming the core metric for AI-native B2B products, citing Harvey.

tracking.tldrnewsletter.comfrom the editor’s archiveMay 2026
70

My LLM coding workflow going into 2026

Describes a personal workflow for planning, coding and structuring AI-assisted software development.

link.mail.beehiiv.comfrom the editor’s archiveJan 2026
70

The next generation of MCP

Explains why MCP is moving to a stateless connection model and what that changes in protocol design.

Cloudflare BlogAug 2026