This matrix compares the thirty people-and-publications profiled in this category, row by row, so deciding what to follow does not require reading thirty notes blind.
The axis that actually segments the field is what each voice gives you: hands-on tool practice, an evaluation method, the model-and-research layer, industry-and-org analysis, or a structured on-ramp, and the personalities span all five, with the hands-on band now the largest, split between the harness builders, the chroniclers, and the enterprise vantage.
Legend: each cell reads as a description; every cell traces to the linked member note and its references.
The matrix #
| Row | Addy Osmani | AI Jason | Andrej Karpathy | Andrew Ng | Armin Ronacher | Boris Cherny | Caleb Writes Code | Chip Huyen | Dax Raad | Dex Horthy | Geoffrey Huntley | Hamel Husain | Harper Reed | IndyDevDan | Jesse Vincent | Kent Beck | Latent Space | Lilian Weng | Mario Zechner | Matt Pocock | Nathan Lambert | Owain Lewis | Paul Gauthier | Peter Steinberger | Ray Amjad | Shreya Shankar | Simon Willison | Steve Yegge | The Pragmatic Engineer | Thorsten Ball |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Primary platform | Blog + books + repos | YouTube | Essays + talks | Newsletter + courses | Blog (lucumr.pocoo.org) + repos | X + Threads + rare blog posts and interviews | YouTube channel | Site + books | GitHub + HN threads + X (no blog) | GitHub methodology repo + weekly live show | ghuntley.com (Ghost blog) plus X and workshop talks | Blog + Substack | harper.blog (posts plus near-daily notes) | YouTube (@indydevdan) plus agenticengineer.com and GitHub | blog.fsck.com plus the obra/superpowers repo | Tidy First? Substack plus kentbeck.com | Substack + podcast + conferences | Blog (Lil’Log) | Long technical blog posts (mariozechner.at) | Course sites (Total TypeScript, AI Hero) + GitHub | Substack + site | YouTube (@owainlewis) plus GitHub repos and aiengineer.co | aider.chat docs and blog plus GitHub | GitHub tool fleet + blog (steipete.me) + X | YouTube (@RAmjad) plus agenticcoding.school and products | Personal site plus peer-reviewed papers and open-source tools | Daily blog + TIL | Blog (yegge.ai) + essays | Newsletter + podcast | Register Spill newsletter + ampcode.com docs and screencasts |
| Cadence | Periodic essays | Two to three videos a month | Low, periodic | Weekly newsletter | One to three essays a month (eleven, July 4 to September 29, 2026) | Sporadic (one blog post since mid-2024, latest 2026-09-19) | About twice a week | Books + periodic essays | Irregular and scattered (personal site dormant since October 2021) | Weekly show, irregular text output | Monthly or better through July 2026, slower since | Monthly-ish long posts | Long-form posts in monthly clusters, notes near daily | Strictly weekly, every Monday, 16 consecutive Mondays verified | Multiple posts per month, monthly Superpowers releases | Weekly, dated entries into 2026 | Weekly newsletter, daily AINews | Low, irregular | A few posts per year (three in 2026 through May), weekly pi releases | Irregular posts, active cohort and course cycles | High, multiple a week | About weekly, 16 uploads May 15 to September 28, 2026 | Historically prolific, stalled (no commits since May 2026) | Daily tool releases, blog cooled to one 2026 post (30-plus posts June to December 2025) | Two to three per month, 15 uploads April 9 to September 18, 2026 | Slow, a few posts a year, papers in conference batches | Daily, multiple posts | Sporadic + shipped code | Weekly issue | Weekly (issues 91 to 100, July to September 2026) |
| Focus | Enterprise agentic engineering, verification discipline, code quality | Agent workflows and context engineering | Conceptual vocabulary, frontier research | Education, agentic patterns, overview | Skeptical senior-engineer analysis of coding agents, tool-call internals, team coordination | Claude Code design philosophy, AI-first team process, TypeScript | Model releases and agentic-engineering explainers | AI systems design, production | Shipping and defending a top-tier open-source coding agent | Agent reliability principles, owned context, human checkpoints | Autonomous loops, context engineering, AI economics | Evaluation and data-driven improvement | End-to-end LLM codegen workflows, spec-first planning, harness experiments | Concept frameworks, core four, operating levels, software factories, swarms, model fusion | Agentic workflow methodology, skills engineering, TDD | Software process and design under agents, skill repricing, code deflation economics | Industry, labs, interviews, trends | Research surveys (agents, alignment) | Harness minimalism, MCP skepticism, observability, OSS governance | AI-coding skills, agent workflows, TypeScript depth | Models, post-training, open ecosystem | Agentic infrastructure, software factories, agent loops, coding agents built from scratch | Terminal AI pair programming, repo maps, git-first workflows, model benchmarking | Builder-evangelism, agent workflows, personal agents, developer tooling | Claude Code and Codex feature deep-dives, loop engineering, verification concepts | Data systems for LLMs, evals, semantic operators, data agents | Hands-on tools and agent practice | Agent-era thesis and builds | Org design, hiring, adoption data | Daily agent practice on Amp, systems programming, interpreters and compilers |
| Media format | Text + books | Video | Text + video talks | Newsletter + video courses | Long-form essays with published experiment numbers | Interviews, short posts, one O’Reilly book | Video | Text + books | Repos, Show HN, thread replies, podcast and YouTube interviews | Repo guide plus recorded episodes with code | First-person essays with embedded prompts, screenshots, and video | Text | Long-form posts with inline prompts, short notes | Weekly long-form videos with dense chapters, paired devlogs and repos | Point-in-time methodology writeups and release posts | Short essays, system-prompt case studies, book chapters | Newsletter + audio + events | Text | Long-form essays with benchmarks and code | Posts, exercises, installable skill files | Text + podcast | Long-form walkthroughs with engineering writeups and linked repos | Tool docs, benchmark leaderboards, release notes | Essays, small tools, meetups, talks (TED 2026) | Single-feature video essays with demos and timestamped chapters | Blog essays, papers, runnable benchmarks, course material | Text | Text | Newsletter + audio | Newsletter digest, self-published books, screencast |
| Depth vs breadth | Leader-plus-practitioner depth, enterprise scope | Broad, execution-level, shallow | Conceptual framing | Broad on-ramp, shallow | Deep on few things (pi, CPython, his own experiments), no news coverage | Deep on one product, no breadth | Broad, current, explainer depth | Broad systems survey | Deep on his own stack (serverless, auth, agents), nothing archived | One methodology, argued deep | Depth on one technique family, breadth on its consequences | Deep on evals, narrow scope | Deep on one workflow and its evolution, broad on adjacent experiments | Breadth across the whole conceptual stack, weekly | Deep on one disciplined workflow, wide harness coverage | Deep on process and economics, near zero on tools and models | Broad industry synthesis | Deep references, broad | Deep argument on one design position | Deep on working practice, narrow on models | Deep on post-training | Depth on one coherent factory thread, not release news | Deep on edit-and-commit ergonomics and context design | Extreme breadth of small tools over long analysis | Deep per release, deliberately narrow | Deep on data-systems measurement, deliberately narrow | Dense and deep on tools | Provocative theses | Org-level, broad | Deep on one product and one book topic |
| Builds tools | Yes (agent-skills repo) | Yes (production agents) | Yes (nanoGPT, AutoResearch) | No (platform) | Yes (Flask legacy; pi second-largest contributor) | Yes (Claude Code at Anthropic) | No | Micro tools | Yes (OpenCode at Anomaly, SST, OpenAuth, OpenNext) | Yes (12-factor agents repo, HumanLayer workspace) | Yes (Ralph technique, The Weaving Loom, workshop agent code) | Some (evals tooling) | Yes (breakaway-agent, observatory at 2389-research) | Yes (claude-code-hooks-mastery, pi-vs-claude-code, super-simple-software-factory) | Yes (Superpowers, Evener, superpowers-evals) | Yes, builds real projects (BPlusTree3) to test methodology claims | Indexes, does not build harnesses | No | Yes (pi at Earendil, libgdx) | Yes (skills repo, sandcastle, ts-reset) | Yes (OLMo, RL tools) | Yes, heavily (Machinist, Neo, Push, blueprint, Factory) | Yes (aider, largely written with aider) | Yes (OpenClaw, CodexBar, Peekaboo, mcporter, oracle) | Yes (AgentStack, Impello, HyperWhisper, workflow-creator) | Yes (DocETL, DocWrangler, Data Agent Bench) | Yes (LLM, Datasette) | Yes (beads; shut Gas Town down in September 2026) | No | Yes (Amp, co-founder) |
| Evaluation and verification | High, verification-first | Light, practical | High on his own terms | Skills map, moderate | Runs and publishes his own failure-prone experiments (35h software factory, $1,200 spend) | None of his own, the product faces measured third-party cost critique he does not engage | Light, explainer-level | Systems-level | Answers security critiques point by point with versions and defaults | Reasoned principles, no evals practice | Demonstration-based, ships artifacts like CURSED, no formal evals | Central, the whole method | Tests and TDD as anti-hallucination gates, checkable prompt plans | Benchmark skepticism and own-index method, repos carry gate checks | Public eval suite with published cost, speed, and bug deltas including unfavorable ones | TDD as the agent control loop, published prompts and time logs | Moderate, evangelistic on brand | Research rigor on limits | Terminal-Bench 2.0 submission plus token accounting for his claims | Teaches feedback loops and TDD, publishes no benchmarks | RewardBench, comparisons | States token costs and failure modes per build | Built the polyglot leaderboard, the field’s independent yardstick | First-person practice, not benchmarks, ecosystem ports as evidence | Verification-first thesis, prefer a verifier over an instruction | Core discipline, evals as systematic measurement with public traces | Security-critical, tests agents | Thematic, needs cross-check | Empirical data, surveys | Shows working proof in screencasts, no independent evaluation |
| Model and research layer | Light | Light | Central | Moderate | Occasional (reasoning-trace explainers, tool-schema forensics) | Minimal (stated via product choices) | Strong release coverage, light research depth | Moderate | Minimal (practical cache and provider work) | None, framework-agnostic by design | Informal model taxonomy from Amp work | Applied, light research | Pragmatic model choice per step | Strong, model stacking, fusion, and benchmark re-ranking | Practical per-model quirk adaptation across 16 harnesses | None, deliberately agnostic about models | Strong | Central | Minimal (models as interchangeable providers) | None | Central | Light, model choice framed as comparison and cost | None, model-agnostic BYOK tooling | Minimal (model-adjacent via his OpenAI role) | Moderate, tracks new models as they land in workflows | None, application and data layer | Tool and model churn, light research | Moderate | Moderate | Minimal |
| Commercial model | Free blog, paid books (one free online edition) | Free + paid community and sponsors | Free | Free + paid courses | Free, GitHub sponsors, vendor stake via Earendil | Anthropic salary, book royalties | Ad and sponsor-funded, plus Patreon | Free + paid books | Open source (MIT) at Anomaly, separate healthcare product | Free CC BY-SA methodology atop paid SaaS | Employed engineer (Sourcegraph/Amp), no sponsored content | Free + paid course and consulting | Free blog, revenue via 2389.ai and speaking | Free videos funneling to paid coding courses | Prime Radiant founder, commercial support around open-source Superpowers | Free posts with paid tier, consulting, speaking | Freemium + conference tickets | Free | Earendil salary and equity, planned MIT core plus paid tiers | Paid workshops and cohorts, free content as funnel | Free + reader-supported + book | Free videos funneling to aiengineer.co, Skool community, consultancy | Free open source, sponsor funded | OpenAI salary, foundation-stewarded open source | No sponsors by policy, cohort courses and training revenue | Free content, paid Maven course with Husain, advising | Free, sponsor-funded | Free | Freemium, paywalled core | Amp equity and product revenue, direct book sales |
| Reader slot | Enterprise lead and practitioner | Video practitioner | Vocabulary-setter | Educator, general AI reader | The monthly skeptic to check hype against, not a feed | Harness-builder insider | Release-tracking video viewer | Systems and architecture reader | The vendor defending his agent under fire, not a citable feed | Architecture principles for agent builders | The provocation-and-technique voice | Data-centric eval practitioner | The starting workflow to run and mutate | The weekly worldview slot that organizes everything else | The codified-practice voice | The process authority to consult after the trackers | Industry and community view | Research reference reader | Minimalist harness theorist | Practitioner education | Model and research reader | The clone-and-run slot for engineers who want the code | The reference archive for pre-agentic tool design | Energy and artifacts, not neutral analysis | The release-analysis slot for daily Claude Code users | The peer-reviewed measurement voice | Practitioner daily signal | Provocateur builder | Engineering leader | Working practitioner on the record |
| Enterprise vs frontier | Enterprise practice formed at Google, now Member of Technical Staff at Anthropic | Individual and startup | Frontier labs | General audience + enterprises | Practitioner and team level, vendor-aware | Frontier lab | Individual learners and practitioners | Production teams | BYOK practitioner tooling, no enterprise story | Production-bound product teams | Frontier, greenfield maximalism | Applied AI in companies | Individual and small-team practitioner | Frontier-chasing framing on toy-scale demos | Frontier methods with production discipline | Team and org process | Frontier labs and AI-native startups | Frontier labs | Frontier indie absorbed into a small lab | Individual practitioners | Frontier labs + open models | Solo-builder and consultancy scale | Individual developer and open source | Frontier and indie builders | Production practitioner framing with enterprise-oriented advice | Academic-applied, enterprise-data slant | Open, local, practitioner | Individual and teams | Established-company orgs | Frontier startup (Amp, spun out of Sourcegraph) |
| Skepticism | Moderate, with an evangelist’s vantage | Moderate, tool-hype risk | High on his own terms | Low, optimistic | High and quantified, the category’s best case-against voice | Low, everything he publishes doubles as product marketing | Moderate, sponsor-dense | Moderate, balanced | Practitioner-pragmatic, defensive about his own product’s record | Framework-skeptical, increasingly self-interested | Low on the industry, high on his own technique’s edges | High, evidence-based | Self-skeptical about shelf life, discloses AI-written drafts | Criticizes benchmarks and hype but trades in both | Moderate and self-applied, publishes costs and failures | Measured, admits extrapolating from personal experiments | Moderate, evangelistic on brand | High on her own terms | High of others’ tools, contested on his own YOLO security posture | Moderate, vendor-critical but sells his own courses | Sharp about hype in his domain | Self-skeptical descriptions, zero third-party critical coverage | Subject of the anti-non-agentic critique, autonomy refused as a feature | Low, evangelist with a hot self-narrative and documented pushback | Skeptical of agent output, unselfcritical about his own funnel | Engages anti-evals critics head-on | High, security-critical | Provocative, needs cross-check | High, empirical | Low self-skepticism, community critiques price and moat |
Reading the matrix #
I read this table by columns, matching a reader slot rather than a source. Nobody covers all five focus bands well, which is the argument for following several: pick a practitioner, an evaluation voice, a model-layer reader, an industry voice, and an educator.
The hands-on cluster now spans text, eval-discipline, a five-channel video band, and the enterprise vantage: Simon Willison is the reliable daily text chronicler, AI Jason builds complete workflows on camera, Caleb Writes Code explains each release and agentic concept at news speed, Hamel Husain supplies the data-driven method for deciding whether those workflows work, Steve Yegge remains the provocative thesis-builder, though his Gas Town build was shut down in September 2026 after he admitted it never worked for him, and Addy Osmani reports the same practice from 14 years inside a hyperscaler (he now works on Claude Code at Anthropic), holding agent output to a production quality bar.
The video band’s three new channels split by what they optimize, and all three sell around the free videos. IndyDevDan is the weekly worldview, organizing the whole conceptual stack into frameworks every Monday. Ray Amjad is the release analyst, one Claude Code or Codex feature at a time with a verification-first thesis and no sponsors. Owain Lewis is the clone-and-run builder, software-factory walkthroughs where every video ships a runnable repo. Take the frameworks, run the repos, and discount the funnels.
The harness-builder cluster is the biggest single addition, eight new columns of people who built the tools the rest of this table watches, and they need reading against their own incentives. Boris Cherny built Claude Code and publishes rarely, so his value is the maker’s own reasoning wherever it surfaces. Thorsten Ball co-founded Amp and screencasts his actual workday, the closest thing to on-the-record practice. Mario Zechner authors pi and argues its minimalism in benchmarked essays. Dax Raad created OpenCode and defends it point by point in public threads, the vendor under fire on the record. Armin Ronacher, pi’s second-largest contributor, is the cluster’s in-house skeptic, publishing invoice-backed doubts about long-horizon models. Peter Steinberger ships the OpenClaw tool fleet with evangelist energy and documented pushback. Dex Horthy supplies the design principles (12-Factor Agents) and Matt Pocock the installable curriculum, both selling around the free layer, so take them as structured entry points rather than neutral reviews. Paul Gauthier is the cluster’s pre-agentic root, the aider author whose repo maps and BYOK pair-programming set the template the wave built on, with development stalled since May 2026.
The workflow-and-methodology voices added this run cover how to run the loop rather than which harness to pick. Geoffrey Huntley reduces agentic coding to a persistence loop and forecasts what that does to the industry. Jesse Vincent documents his own agent practice and ships Superpowers, the disciplined-workflows skill suite with published eval deltas. Kent Beck reframes TDD for agents as augmented coding, the process authority adapting his own method. Harper Reed wrote the spec-first codegen workflow essay that discussion threads still treat as the reference. Shreya Shankar supplies the peer-reviewed measurement layer, making agent reliability and data quality testable with DocETL and public benchmark traces.
The model-and-research layer is now its own band: Lilian Weng writes the durable research references, Nathan Lambert tracks the current post-training and open-model state from inside the labs, and Andrej Karpathy sets the conceptual vocabulary they all operate inside.
The systems-and-education band fills the gap the seed explicitly named: Chip Huyen gives the production systems survey, Andrew Ng supplies the structured on-ramp and mainstream vocabulary, and together they serve the engineer who wants breadth before depth.
Latent Space and The Pragmatic Engineer remain the two industry synthesizers, split by audience: Latent Space points at the frontier labs and the AI-native startups, The Pragmatic Engineer points at established engineering organizations, so your employer’s profile picks your primary.
The enterprise-hands-on gap the seed named is now filled, with one boundary drawn: Addy Osmani is the sustained enterprise-hands-on voice, but he built that practice over 14 years inside Google and now works on Claude Code at Anthropic, two AI companies, with leader-level rather than terminal-level detail, so the still-empty scaffold is the low-drama, everyday terminal operator inside a large non-AI company.
Choosing from the matrix #
- Need daily, hands-on signal on tools and models: Simon Willison.
- Built the harness you run and want the maker’s own reasoning: Boris Cherny.
- Want a working practitioner’s weekly on-the-record account of agent-driven work: Thorsten Ball.
- Want harness minimalism argued with benchmarks: Mario Zechner.
- Want the skeptic’s invoice-backed case against agent hype: Armin Ronacher.
- Want the tool-builder’s energy and artifacts over analysis: Peter Steinberger.
- Want to watch a vendor defend his agent under public fire: Dax Raad.
- Want design principles for reliable agents before picking tools: Dex Horthy.
- Want agent workflows taught as an installable curriculum: Matt Pocock.
- Want the loop-and-persistence technique taken to its limit: Geoffrey Huntley.
- Want a disciplined personal workflow you can install: Jesse Vincent.
- Want process discipline adapted to agents by the person who wrote the book on it: Kent Beck.
- Want the starting codegen workflow to run and mutate: Harper Reed.
- Want peer-reviewed measurement behind agent-reliability claims: Shreya Shankar.
- Studying pre-agentic tool design and repo context: Paul Gauthier.
- Want a weekly worldview that organizes the agentic stack: IndyDevDan.
- Want Claude Code and Codex releases analyzed in depth: Ray Amjad.
- Want software-factory walkthroughs with runnable repos: Owain Lewis.
- Run agent adoption inside an established company and want enterprise-grounded practice: Addy Osmani.
- Ship an AI product and cannot tell if it works: Hamel Husain.
- Learn agent workflows best by watching: AI Jason.
- Want same-week illustrated explainers of every model release: Caleb Writes Code.
- Want a sharp, opinionated thesis and are willing to cross-check: Steve Yegge.
- Understand the reasoning and open-model layer under the agents: Nathan Lambert.
- Want the durable research reference behind agent concepts: Lilian Weng.
- Want the conceptual framing behind the churn: Andrej Karpathy.
- Want one structured overview of building LLM applications: Chip Huyen.
- Need the industry, lab, and community view plus events: Latent Space.
- Lead an engineering org adopting agents and want data: The Pragmatic Engineer.
- Are new to AI engineering and want a structured path: Andrew Ng.
Changes #
- 2026-08-29 - Created with five columns segmented on the focus axis, naming the empty enterprise-hands-on cell as scaffold for a future member.
- 2026-08-29 - Extended from five to eleven columns, re-segmenting the thesis onto five focus bands and filling all reader-slot rows.
- 2026-08-29 - Extended to twelve columns, adding the Caleb Writes Code column and a release-explainer choosing bullet.
- 2026-09-02 - Extended to thirteen columns, adding Addy Osmani and filling the enterprise-hands-on scaffold cell.
- 2026-09-13 - Corrected the Yegge primary-platform cell to his yegge.ai site (165 essays there, Substack carrying no published posts) and aligned the AI Jason cadence cell to two to three videos a month.
- 2026-09-18 - Updated the Osmani enterprise-vantage cell after his move from Google to Anthropic and the Yegge builds-tools cell after he shut Gas Town down; no membership change, columns stay at thirteen.
- 2026-09-24 - Extended from thirteen to twenty-one columns with the owner-commissioned harness-author wave (Boris Cherny, Thorsten Ball, Mario Zechner, Armin Ronacher, Peter Steinberger, Dax Raad, Dex Horthy, Matt Pocock), columns re-sorted alphabetically, thesis and reading rewritten for the harness-builder cluster, intro and choosing list extended, every new cell traced to its member note.
- 2026-09-24 - Extended from twenty-one to twenty-seven columns with the owner-commissioned second wave (Geoffrey Huntley, Jesse Vincent, Kent Beck, Harper Reed, Shreya Shankar, Paul Gauthier), columns re-sorted alphabetically, reading and choosing extended for the workflow-and-methodology cluster, every new cell traced to its member note.
- 2026-09-24 - Extended from twenty-seven to thirty columns with the owner-commissioned video-practitioner wave (IndyDevDan, Owain Lewis, Ray Amjad), columns re-sorted alphabetically, reading and choosing extended for the video band, every new cell traced to its member note.
- 2026-09-24 - Removed the verification preamble line per the no-preamble rule; verification history lives in this Changes list.
- 2026-09-24 - Removed the verification preamble line on owner request.
- 2026-09-27 - Reworded a cell off a banned-term compound; meaning unchanged.
- 2026-09-30 - Refreshed the Ronacher cadence cell to eleven essays between July 4 and September 29, 2026 and the IndyDevDan cadence cell to 16 consecutive Mondays; no membership change, columns stay at thirty.
- 2026-10-02 - Refreshed the Owain Lewis cadence cell to 16 uploads between May 15 and September 28, 2026; no membership change, columns stay at thirty.
- 2026-10-03 - Replaced the dead karpathy.bearblog.dev Sequoia Ascent reference (the Bear blog 404s as of 2026-10-03) with the Internet Archive snapshot; no cadence cell moved on re-verification, columns stay at thirty.
See also #
- Agentic Coding Tools Landscape - the tool landscape these voices report on and steer
- Simon Willison - the practitioner column in detail
- Hamel Husain - the evaluation method that verifies practitioner claims
- Nathan Lambert - the model and open-ecosystem layer in detail
- Managing Many Concurrent LLM Agent Sessions - the supervision subject several of these voices write about
References #
https://simonwillison.net/ - cadence, focus, and hands-on grounding for the Willison column
https://addyosmani.com/ - the enterprise-hands-on agentic-engineering focus grounding the Osmani column
https://hamel.dev/ - the evals focus grounding the Husain column
https://www.youtube.com/@AIJasonZ - platform and video focus for the AI Jason column
https://www.interconnects.ai/ - model and post-training focus for the Lambert column
https://lilianweng.github.io/ - the research-reference focus for the Weng column
https://web.archive.org/web/20260925165449/https://karpathy.bearblog.dev/sequoia-ascent-2026/ - the Software 3.0 essay grounding the Karpathy column (Internet Archive snapshot; the live Bear blog 404s as of 2026-10-03)
https://huyenchip.com/ - the systems-survey focus grounding the Huyen column
https://www.latent.space/ - platform, subscribers, and conference for the Latent Space column
https://newsletter.pragmaticengineer.com/ - cadence, paywall, and org focus for the Pragmatic Engineer column
https://www.deeplearning.ai/the-batch - the education and newsletter focus for the Ng column