WE WILL FIX EVERYTHING A manifesto for the agent era · Andrius Bartulis CODING IS DEAD. Long live software engineering. Humans own product and technical design. Give agents the implementation and build the software factory. Product & technical design Goals · Scope · User journeys Architecture · Data · APIs Build and review Separate agents build, test and challenge the result. Try it. Refine the design. Review the working product. Request changes, then rebuild. Monitoring and recovery Production monitoring. Adversarial staging, 24/7. Iterate on product and technical design as you learn. Keep agents building and checking each revision. Where to focus How should we build software in the age of AI? We can build far more software if we stop writing it all ourselves. 01 Abstraction Move up the abstraction stack. Use agents to build more useful software than you could implement yourself. Developers should take responsibility for the product and technical design: what it needs to do, how the parts fit together and where failure would be costly. We’ve been through this before. Some assembly programmers resisted higher-level languages because they could write more efficient code by hand. But faster computers made it affordable to trade some efficiency for shorter development time. Judge agents by the total cost of delivery and maintenance, and by how much useful software they help you ship. 02 Product & technical design Design with agents, then let them build. Work out the product and technical design with agents, then give them the full implementation. Humans own Product goals, architecture, API contracts, data models, authentication boundaries and performance budgets. Agents deliver Implementation, integration, migrations, documentation and tests, plus architecture and security audits. Define the product. Agree on who it serves, the problem to solve, the smallest useful journey and how you’ll judge success. Design the system. Work out the technical design with agents. Compare approaches, challenge assumptions and record the trade-offs. Let agents build it. Have agents deliver a complete, working slice, including integration and tests. Separate agents check it against the requirements. Review the working product. Humans try the user journey, check the product and technical design, and review evidence from independent agents. Refine the design and repeat. Request product and technical design changes as gaps become clear. Have agents implement and recheck each revision. One week from idea to MVP. Plan for one week to a usable MVP; two weeks is the maximum. If a project needs longer, cut the first version to a complete journey you can deliver within that limit. Keep the release checks. Delivery targets Work — Target — Devs Small changes — One day — 1 Small features — A few days — 1 Large features — One–two weeks — 1–2 max Keep the team small. With agents doing the implementation, adding developers can slow delivery through extra coordination. Product and the relevant teams try the MVP, then decide whether to improve it before release or launch a beta and iterate. You remain responsible for the design and for what reaches users. 03 The factory Optimise the software factory. Put your effort into automated testing and review so routine changes can ship without you reading every diff. The agent that built a feature can miss the same mistake when it writes the tests. Have other agents check behaviour against the requirements, review architecture and audit security. Use separate review contexts so agents don’t simply inherit the builder’s assumptions. Before release Check real user journeys and failure recovery. Test permissions, data integrity and API contracts. Measure performance against agreed budgets. Audit security with independent agents. Prepare observability and a tested recovery path. Most PRs shouldn’t wait for human review. Keep PRs as a record of the change, decisions and checks. Human approval should be optional for most changes, and developers can request a review whenever they want. Have a categorisation agent approve low-risk changes for merge when they fit the agreed design, are easy to reverse and pass independent checks. Route changes to the agreed architecture or technical design, unusually complex work and changes with serious consequences to the right staff member. Point to the areas of risk and explain what needs their judgment. Escalate uncertain classifications too. Keep work moving through the whole system. Track time from idea to a usable result, including waits for decisions, environments, reviews and rework. Fix the largest constraint, then measure again. Limit work in progress so unfinished tasks don’t pile up. Borrow the focus on flow from Toyota’s production system and the attention to constraints from The Goal. When a defect escapes, add a check that would have caught it. 04 Cloud workflow Move the work to the cloud. Developers shouldn’t need a local IDE or write routine code by hand. Run agents in parallel cloud workspaces where they can implement, test and review without depending on your laptop. Give each task a reproducible environment and only the access it needs. Be selective about reading code. Read API contracts, database changes, authentication logic and mission-critical code when a decision needs your judgment. Let automated checks and agent reviews cover routine implementation. Use system maps, data models, performance traces and review reports to assess the design. 05 Continuous assurance Keep trying to break it. Have agents attack staging 24 hours a day, seven days a week. Let them explore user journeys, fuzz inputs, test permissions, interrupt dependencies and probe performance limits. Use isolated test data and cap resource use. Keep watching after deployment. Agents should watch deployments, logs, errors and latency, trace regressions to changes, and propose fixes. Test those fixes in staging and check that recovery worked. Agents can act Let agents handle known repairs within agreed permissions, using the same release checks as new features. Humans decide Bring uncertain root causes, design changes, sensitive data decisions and actions beyond agreed authority to a human. START HERE / ONE REAL SYSTEM Build your first factory. Choose one troublesome user journey. Agree on the product and technical design, then have agents build and independently test an MVP within a week. Let product try it and decide what comes next. Track delivery time, escaped defects and how often someone has to intervene. WE CAN FIX EVERYTHING. AND WE WILL.