wiki / queries / agentic-code-quality-cycle-2-entity-accounting

Agentic Code Quality Cycle 2 — Entity Accounting

high confidence updated 2026-08-27 evaluation · agents · retrospective

Agentic Code Quality Cycle 2 — Entity Accounting

This document records the accounting and disposition of authors and organizations across all sources ingested in Cycle 2.

Entity Dispositions

EntityDispositionTarget FileJustification / Reason
Simon WillisonUpdateentities/simon-willison.mdCataloged core agentic engineering patterns & proof-of-work doctrine.
Addy OsmaniUpdateentities/addy-osmani.mdAuthor of Loop Engineering, Outer Loop, and Autonomy Levels canon.
Kent BeckCreateentities/kent-beck.mdPioneer of XP/TDD; framed TDD as agent governor & test-deletion anti-pattern.
Martin FowlerCreateentities/martin-fowler.mdSoftware architecture authority; framed TDD as human comprehension anchor.
Gergely OroszCreateentities/gergely-orosz.mdAuthor of The Pragmatic Engineer; documented agent workflows with Beck/Cherny.
Boris ChernyCreateentities/boris-cherny.mdCreator of Claude Code; verification-first harness & attacks-become-evals pattern.
Armin RonacherUpdateentities/armin-ronacher.mdEvaluated agent SDK abstractions & identified testing/evals as hardest problem.
StripeCreateentities/stripe.mdBuilt Stripe Integration Benchmark; identified false-victory failure mode.
AirbnbCreateentities/airbnb.mdPioneered industrial Eval-Driven Development (EDD) & 3-layer eval funnel.
AnthropicUpdateentities/anthropic.mdPublished Demystifying Evals for AI Agents & Claude Code harness.
OpenAIUpdateentities/openai.mdPublished evaluation best practices & reference coding agent SDK harness.
GitHubCreateentities/github.mdPublished Spec-Driven Development toolkit for AI agents.
Hamel HusainUpdateentities/hamel-husain.mdAuthority on AI Evals FAQ & evals-skills framework.
Shreya ShankarUpdateentities/shreya-shankar.mdCo-author of evals-skills & evaluator alignment research.
Rohit GirmeSkipLead author on Airbnb EDD article; represented via airbnb organization entity.
Dan MillerSkipCo-author on Airbnb EDD article; represented via airbnb.
Carol LiangSkipAuthor on Stripe benchmark; represented via stripe.
Kevin HoSkipAuthor on Stripe benchmark; represented via stripe.
Zixiao ZhaoSkipLead author on arXiv:2604.16790; single citation below threshold.
Evidence — verified primary sources
simon-willison-agentic-engineering-patterns-2026 https://simonwillison.net/guides/agentic-engineering-patterns/
ingested 2026-08-27
sha256:f0fba6d5e10f…
simon-willison-code-proven-to-work-2025 https://simonwillison.net/2025/Dec/18/code-proven-to-work/
ingested 2026-08-27
sha256:7c2bf07e112f…
addy-osmani-practical-loop-engineering-2026 https://addyosmani.com/blog/practical-loop-engineering/
ingested 2026-08-27
sha256:293f5d50e3ff…
addy-osmani-loop-engineering-2026 https://addyosmani.com/blog/loop-engineering/
ingested 2026-08-27
sha256:e1dda92b345d…
addy-osmani-agentic-code-review-2026 https://addyosmani.com/blog/agentic-code-review/
ingested 2026-08-27
sha256:29c683772620…
addy-osmani-human-judgment-software-factory-2026 https://addyosmani.com/blog/human-judgment-doesnt-leave-the-software/
ingested 2026-08-27
sha256:3573d9ce5335…
addy-osmani-own-the-outer-loop-2026 https://addyosmani.com/blog/own-the-outer-loop/
ingested 2026-08-27
sha256:3f948a41d386…
addy-osmani-agentic-autonomy-levels-2026 https://addyosmani.com/blog/agentic-autonomy-levels/
ingested 2026-08-27
sha256:bf365ff64dfc…
kent-beck-gergely-orosz-tdd-ai-agents-2025 https://newsletter.pragmaticengineer.com/p/tdd-ai-agents-and-coding-with-kent
ingested 2026-08-27
sha256:15b8fadfccc7…
martin-fowler-fragments-2026-01-08 https://martinfowler.com/fragments/2026-01-08.html
ingested 2026-08-27
sha256:d8fd9cb6ce81…
boris-cherny-gergely-orosz-building-claude-code-2026 https://newsletter.pragmaticengineer.com/p/building-claude-code-with-boris-cherny
ingested 2026-08-27
sha256:569f43e4a99c…
boris-cherny-how-boris-uses-claude-code-2026 https://howborisusesclaudecode.com/
ingested 2026-08-27
sha256:e62a22b875a3…
armin-ronacher-agent-design-is-still-hard-2025 https://lucumr.pocoo.org/2025/11/21/agents-are-hard/
ingested 2026-08-27
sha256:736e31be887a…
stripe-can-ai-agents-build-real-stripe-integrations-2026 https://stripe.com/blog/can-ai-agents-build-real-stripe-integrations
ingested 2026-08-27
sha256:3d30b4fb0d9b…
stripe-you-cant-whisper-at-an-ai-agent-2026 https://stripe.dev/blog/ai-steering-experiments
ingested 2026-08-27
sha256:c6f4eaf70484…
airbnb-eval-driven-development-2026 https://medium.com/airbnb-engineering/eval-driven-development-lessons-from-evaluating-genai-at-scale-e817e5ae5788
ingested 2026-08-27
sha256:94414a24787c…
anthropic-demystifying-evals-for-ai-agents-2026 https://www.anthropic.com/engineering/demystifying-evals-for-ai-agents
ingested 2026-08-27
sha256:0be1b99962fe…
openai-evaluation-best-practices-2026 https://developers.openai.com/api/docs/guides/evaluation-best-practices
ingested 2026-08-27
sha256:5529f24f0c32…
openai-build-a-coding-agent-gpt-5-1-2025 https://developers.openai.com/cookbook/examples/build_a_coding_agent_with_gpt-5.1
ingested 2026-08-27
sha256:5405849943d3…
github-spec-driven-development-ai-2025 https://github.blog/ai-and-ml/generative-ai/spec-driven-development-with-ai-get-started-with-a-new-open-source-toolkit/
ingested 2026-08-27
sha256:efac995e17bb…
hamel-husain-ai-evals-faq-2026 https://hamel.dev/blog/posts/evals-faq/
ingested 2026-08-27
sha256:fe8dd2419e8d…
hamel-husain-shreya-shankar-evals-skills-2026 https://hamel.dev/blog/posts/evals-skills/
ingested 2026-08-27
sha256:c4fc84c125ac…
raw/papers/bias-in-the-loop-llm-judge-se-2026.md internal workspace doc
hn-practitioner-postmortems-agentic-code-quality-2026 https://news.ycombinator.com/item?id=47161209
ingested 2026-08-27
sha256:09a7378bca61…