Sign InOpen Brain
Back
AI EngineerVideoSource LinkedagentNew

Harness Engineering is not Enough: Why Software Factories Fail — Dex Horthy, HumanLayer

Agree on design before implementation and build vertical slices so generated changes remain readable and testable.

AI EngineerAI EngineerAI EngineerJul 23, 20262 min
Open SourceOpen MarkdownOpen JSON
Source Summary

Coding-agent loops can raise throughput without preserving maintainability. Keep human ownership of code, and use upfront alignment to make review affordable instead of trying to automate it away.

Practical Implication

Keep humans responsible for the resulting code. Use model-assisted planning, agree on design before implementation, and build in vertical slices so every generated change remains practical to read and test.

Agent-Ready Context
The talk argues that coding models are rewarded mainly when **code runs and tests pass**, not when architecture remains easy to change. Review agents and extra loops can raise the floor, but cannot supply a missing maintainability signal.

Keep humans responsible for the resulting code. Use **model-assisted planning**, agree on design before implementation, and build in vertical slices so every generated change remains practical to read and test.

There is no established benchmark here that proves how well current models preserve codebase quality. Longer-task evaluations such as **SWE Marathon**, DeepSuite, and FrontierCode may help, but model-based quality judges have their own ceiling.
Context Map
agentcoding#harness-engineering#agent-reliability#agent-evalsGeneric AgentPrepare Coding Session
Rate This Item
Personal Note

No note yet. Notes are included in exported bundles.

Related — Every Edge Explained

No approved edges yet for this post.