Back Matter
Bibliography
Every work cited in this book, in one alphabetical list. Each chapter also carries its own numbered Works Cited; this is their union.
- Abdurrahman. “We Adopted Spec-Driven Development, Then the Specs Started Eating Our Context Window.” Medium, July 2026. https://abdurrahman5.medium.com/we-adopted-spec-driven-development-then-the-specs-started-eating-our-context-window-6cb9476dfedc.
- Scale AI. “SWE-Bench Pro: Can AI Agents Solve Long-Horizon Software Engineering Tasks?.” 2025. https://arxiv.org/abs/2509.16941.
- Ait, Adem, Gwendal Jouneaux, Javier Luis Cánovas Izquierdo, and Jordi Cabot. “Towards Automated Governance: A DSL for Human-Agent Collaboration in Software Projects.” In “Proceedings of the 40th IEEE/ACM International Conference on Automated Software Engineering (ASE), New Ideas and Emerging Results Track.” Special issue, Proceedings of the 40th IEEE/ACM International Conference on Automated Software Engineering (ASE), New Ideas and Emerging Results Track, 2025. https://arxiv.org/abs/2510.14465.
- Alur, Rajeev, and David L. Dill. “A Theory of Timed Automata.” Theoretical Computer Science 126, no. 2 (1994): 183–235.
- Anthropic. “How Anthropic Runs Large-Scale Code Migrations with Claude Code.” Anthropic, July 16, 2026. https://claude.com/blog/ai-code-migration.
- Argote, Linda, and Paul Ingram. “Knowledge Transfer: A Basis for Competitive Advantage in Firms.” Organizational Behavior and Human Decision Processes 82, no. 1 (2000): 150–69.
- Baier, Christel, and Joost-Pieter Katoen. Principles of Model Checking. MIT Press, 2008.
- Bass, Len, Paul Clements, and Rick Kazman. Software Architecture in Practice. 3rd ed. Addison-Wesley, 2012.
- Beach, Derek, and Rasmus Brun Pedersen. Process-Tracing Methods: Foundations and Guidelines. 2nd ed. University of Michigan Press, 2019.
- Beck, Kent, Mike Beedle, Arie van Bennekum, et al. “Manifesto for Agile Software Development.” 2001. https://agilemanifesto.org/.
- Becker, Joel, Nate Rush, Elizabeth Barnes, and David Rein. “Measuring the Impact of Early-2025 AI on Experienced Open-Source Developer Productivity.” 2025. https://arxiv.org/abs/2507.09089.
- Bollinger, Terry B., and Clement L. McGowan. “A Critical Look at Software Capability Evaluations.” IEEE Software 8, no. 4 (1991): 25–41. https://doi.org/10.1109/52.300034.
- Booch, Grady, James Rumbaugh, and Ivar Jacobson. The Unified Modeling Language User Guide. 2nd ed. Addison-Wesley, 2005.
- Brambilla, Marco, Jordi Cabot, and Manuel Wimmer. Model-Driven Software Engineering in Practice. 2nd ed. Morgan & Claypool, 2017.
- Brooks, Frederick P. “No Silver Bullet: Essence and Accidents of Software Engineering.” Computer 20, no. 4 (1987): 10–19.
- Brooks, Frederick P. The Mythical Man-Month: Essays on Software Engineering. Anniversary. Addison-Wesley, 1995.
- Carlini, Nicholas. “Building a C Compiler with a Team of Parallel Claudes.” Anthropic, February 5, 2026. https://www.anthropic.com/engineering/building-c-compiler.
- Cui, Zheyuan (Kevin), Mert Demirer, Sonia Jaffe, Leon Musolff, Sida Peng, and Tobias Salz. “The Effects of Generative AI on High-Skilled Work: Evidence from Three Field Experiments with Software Developers.” Management Science, ahead of print, 2025. https://doi.org/10.1287/mnsc.2025.00535.
- Czarnecki, Krzysztof, and Simon Helsen. “Feature-Based Survey of Model Transformation Approaches.” IBM Systems Journal 45, no. 3 (2006): 621–45. https://doi.org/10.1147/sj.453.0621.
- Davis, James C., Paschal C. Amusuo, Tanmay Singla, Berk Çakar, and Kirsten A. Davis. “Cheap Code, Costly Judgment: A Case Study on Governable Agentic Software Engineering.” 2026. https://arxiv.org/abs/2607.01087.
- Dhanorkar, Shipi, Samir Passi, and Mihaela Vorvoreanu. “Human Oversight of Agentic Systems in Practice: Examining the Oversight Work, Challenges, And Heuristics of Developers Using Software Agents.” 2026. https://arxiv.org/abs/2606.05391.
- “Do Llms Generate Test Oracles That Capture the Actual or the Expected Program Behaviour?.” 2024. https://arxiv.org/abs/2410.21136.
- Faulkner, Steve. “How We Rebuilt Next.js with AI in One Week.” Cloudflare, February 24, 2026. https://blog.cloudflare.com/vinext/.
- Foster, J. Nathan, Michael B. Greenwald, Jonathan T. Moore, Benjamin C. Pierce, and Alan Schmitt. “Combinators for Bidirectional Tree Transformations: A Linguistic Approach to the View-Update Problem.” ACM Transactions on Programming Languages and Systems 29, no. 3 (2007): 17:1–17:65.
- Friedenthal, Sanford, Alan Moore, and Rick Steiner. A Practical Guide to Sysml: The Systems Modeling Language. 3rd ed. Morgan Kaufmann, 2014.
- Gamma, Erich, Richard Helm, Ralph Johnson, and John Vlissides. Design Patterns: Elements of Reusable Object-Oriented Software. Addison-Wesley, 1994.
- Gill, Gurbinder. “The Rise of the Agent OS: Comparing Harnesses, Runtimes, And Orchestration Layers for LLM Agents.” LinkedIn, July 15, 2026. https://www.linkedin.com/pulse/rise-agent-os-comparing-harnesses-runtimes-layers-llm-gurbinder-gill-72c9c/.
- DocAble Research Group. “Deterministic Workflows with Bounded Model Delegation for Software Verification.” 2026. https://arxiv.org/abs/2605.10712.
- Hora, Andre, and Romain Robbes. “Are Coding Agents Generating over-Mocked Tests? An Empirical Study.” 2026. https://arxiv.org/abs/2602.00409.
- Hou, Xinyi, Yanjie Zhao, Yue Liu, et al. “Large Language Models for Software Engineering: A Systematic Literature Review.” ACM Transactions on Software Engineering and Methodology 33, no. 8 (2024). https://doi.org/10.1145/3695988.
- Hutchins, Edwin. Cognition in the Wild. MIT Press, 1995.
- Jackson, Victoria, Susannah Liu, and André van der Hoek. “Using Generative AI in Software Design Education: An Experience Report.” 2025. https://arxiv.org/abs/2506.21703.
- Kruchten, Philippe. “The 4+1 View Model of Architecture.” IEEE Software 12, no. 6 (1995): 42–50.
- Lamport, Leslie. Specifying Systems: The TLA+ Language and Tools for Hardware and Software Engineers. Addison-Wesley, 2002.
- Lee, Han. “Hidden Technical Debt of AI Systems: Agent Harness.” May 8, 2026. https://leehanchung.github.io/blogs/2026/05/08/hidden-technical-debt-agent-harness/.
- Lee, Yoonho, Roshen Nair, Qizheng Zhang, Kangwook Lee, Omar Khattab, and Chelsea Finn. “Meta-Harness: End-to-End Optimization of Model Harnesses.” 2026. https://arxiv.org/abs/2603.28052.
- Leveson, Nancy G. Engineering a Safer World: Systems Thinking Applied to Safety. MIT Press, 2011.
- Li, Kenneth, Aspen K. Hopkins, David Bau, Fernanda Viégas, Hanspeter Pfister, and Martin Wattenberg. “Emergent World Representations: Exploring a Sequence Model Trained on a Synthetic Task.” 2022. https://arxiv.org/abs/2210.13382.
- Liang, Shanchao, Spandan Garg, and Roshanak Zilouchian Moghaddam. “The SWE-Bench Illusion: When State-of-the-Art Llms Remember Instead of Reason.” In “Proceedings of the IEEE/ACM 48th International Conference on Software Engineering, Software Engineering in Practice (ICSE-Seip).” Special issue, Proceedings of the IEEE/ACM 48th International Conference on Software Engineering, Software Engineering in Practice (ICSE-SEIP), 2026. https://doi.org/10.1145/3786583.3786882.
- Lin, Jiahang, Shichun Liu, Chengjun Pan, et al. “Agentic Harness Engineering: Observability-Driven Automatic Evolution of Coding-Agent Harnesses.” 2026. https://arxiv.org/abs/2604.25850.
- Liu, Xiangyan, Bo Lan, Zhiyuan Hu, et al. “Codexgraph: Bridging Large Language Models and Code Repositories via Code Graph Databases.” 2024. https://arxiv.org/abs/2408.03910.
- March, James G. “Exploration and Exploitation in Organizational Learning.” Organization Science 2, no. 1 (1991): 71–87.
- Meadows, Donella H. Thinking in Systems: A Primer. Chelsea Green Publishing, 2008.
- Meyer, Bertrand. “From Probable to Provable.” Communications of the ACM, ahead of print, 2025. https://doi.org/10.1145/3773295.
- Murphy, Gail C., David Notkin, and Kevin Sullivan. “Software Reflexion Models: Bridging the Gap between Source and High-Level Models.” In “Proceedings of the 3rd ACM SIGSOFT Symposium on the Foundations of Software Engineering (FSE-3).” Special issue, Proceedings of the 3rd ACM SIGSOFT Symposium on the Foundations of Software Engineering (FSE-3), 1995, 18–28.
- OpenAI. “Why SWE-Bench Verified No Longer Measures Frontier Coding Capabilities.” OpenAI, February 2026. https://openai.com/index/why-we-no-longer-evaluate-swe-bench-verified/.
- Ousterhout, John. A Philosophy of Software Design. Yaknyam Press, 2018.
- Ouyang, Siru, Wenhao Yu, Kaixin Ma, et al. “Repograph: Enhancing AI Software Engineering with Repository-Level Code Graph.” 2024. https://arxiv.org/abs/2410.14684.
- Peng, Sida, Eirini Kalliamvakou, Peter Cihon, and Mert Demirer. “The Impact of AI on Developer Productivity: Evidence from Github Copilot.” 2023. https://arxiv.org/abs/2302.06590.
- Reganti, Aishwarya Naresh. “The AI Agent Stack in 2026.” The Nuanced Perspective, April 29, 2026. https://thenuancedperspective.substack.com/p/the-ai-agent-stack-in-2026.
- E. Hollnagel, D. D. Woods, N. Leveson. Resilience Engineering: Concepts and Precepts. Ashgate, 2006.
- Ringer, Talia, Karl Palmskog, Ilya Sergey, Milos Gligoric, and Zachary Tatlock. “QED at Large: A Survey of Engineering of Formally Verified Software.” Foundations and Trends in Programming Languages 5, nos. 2–3 (2019): 102–281. https://arxiv.org/abs/2003.06458.
- Runeson, Per, and Martin Höst. “Guidelines for Conducting and Reporting Case Study Research in Software Engineering.” Empirical Software Engineering 14, no. 2 (2009): 131–64. https://doi.org/10.1007/s10664-008-9102-8.
- Shah, Pratik, Rajat Ghosh, Aryan Singhal, and Debojyoti Dutta. “RANGER: Repository-Level Agent for Graph-Enhanced Retrieval.” 2025. https://arxiv.org/abs/2509.25257.
- Shingo, Shigeo. Zero Quality Control: Source Inspection and the Poka-Yoke System. Translated by Andrew P. Dillon. Productivity Press, 1986.
- B. Beyer, C. Jones, J. Petoff, N. R. Murphy. Site Reliability Engineering: How Google Runs Production Systems. O'Reilly Media, 2016.
- Smith, Edward K., Earl T. Barr, Claire Le Goues, and Yuriy Brun. “Is the Cure Worse Than the Disease? Overfitting in Automated Program Repair.” In “Proceedings of the 2015 10th Joint Meeting on Foundations of Software Engineering (Esec/fse).” Special issue, Proceedings of the 2015 10th Joint Meeting on Foundations of Software Engineering (ESEC/FSE), 2015, 532–43. https://doi.org/10.1145/2786805.2786825.
- Stevens, Perdita. “Maintaining Consistency in Networks of Models: Bidirectional Transformations in the Large.” Software and Systems Modeling, ahead of print, 2020. https://doi.org/10.1007/s10270-019-00736-x.
- “SWE-Rebench: An Automated Pipeline for Task Collection and Decontaminated Evaluation of Software Engineering Agents.” 2025. https://arxiv.org/abs/2505.20411.
- Tan, Garry. “Thin Harness, Fat Skills.” April 9, 2026. https://github.com/garrytan/gbrain/blob/master/docs/ethos/THIN_HARNESS_FAT_SKILLS.md.
- Tu, Haoxin, Huan Zhao, Yahui Song, Mehtab Zafar, Ruijie Meng, and Abhik Roychoudhury. “Agentic Verification of Software Systems.” 2025. https://arxiv.org/abs/2511.17330.
- Vaswani, Ashish, Noam Shazeer, Niki Parmar, et al. “Attention Is All You Need.” In “Advances in Neural Information Processing Systems 30 (NIPS 2017).” Special issue, Advances in Neural Information Processing Systems 30 (NIPS 2017), 2017, 5998–6008.
- Vella, Annie, and Kelly Blincoe. “The Impact of AI Coding Assistants on Software Engineering: A Longitudinal Study.” 2026. https://arxiv.org/abs/2605.23135.
- Vogel, Martin, Falk Meyer-Eschenbach, Severin Kohler, Elias Grünewald, and Felix Balzer. “Codebase-Memory: Tree-Sitter-Based Knowledge Graphs for LLM Code Exploration via Mcp.” 2026. https://arxiv.org/abs/2603.27277.
- Wikipedia. “Computer-Aided Software Engineering.” 2026. https://en.wikipedia.org/wiki/Computer-aided_software_engineering.
- Wikipedia. “Hybrid System.” 2026. https://en.wikipedia.org/wiki/Hybrid_system.
- Wikipedia. “Model-Driven Architecture.” 2026. https://en.wikipedia.org/wiki/Model-driven_architecture.
- Wikipedia. “Systems Modeling Language.” 2026. https://en.wikipedia.org/wiki/Systems_modeling_language.
- Wikipedia. “Timed Automaton.” 2026. https://en.wikipedia.org/wiki/Timed_automaton.
- Winters, Titus, Tom Manshreck, and Hyrum Wright. Software Engineering at Google: Lessons Learned from Programming over Time. O'Reilly Media, 2020.
- Young, Justin. “Effective Harnesses for Long-Running Agents.” Anthropic, November 26, 2025. https://www.anthropic.com/engineering/effective-harnesses-for-long-running-agents.
- Zaharia, Matei, Kasey Uhlenhuth, and Corey Zumar. “Introducing Omnigent: A Meta-Harness to Combine, Control and Share Your Agents.” Databricks, June 13, 2026. https://www.databricks.com/blog/introducing-omnigent-meta-harness-combine-control-and-share-your-agents.
- Zhang, Linghao, Shilin He, Chaoyun Zhang, et al. “SWE-Bench Goes Live!.” 2025. https://arxiv.org/abs/2505.23419.
- Zhong, Hailin, and Shengxin Zhu. “AI Harness Engineering: A Runtime Substrate for Foundation-Model Software Agents.” 2026. https://arxiv.org/abs/2605.13357.
© James C. Davis, 2026–present