What is agentic engineering?
TL;DR Highlight
Simon Willison coins the term 'Agentic Engineering' for software development with coding agents, explaining how it's different from plain 'vibe coding' and what the developer's role looks like in this new paradigm.
Who Should Read
Developers looking to adopt coding agents like Claude Code, OpenAI Codex, or Gemini CLI in real work, and software engineers thinking through how their role changes when LLMs are generating the code.
Core Mechanics
- Agentic Engineering is defined as software development where the developer maintains full understanding of the system while delegating implementation to AI agents — in contrast to 'vibe coding' where you accept output without deep comprehension.
- The developer's role shifts from 'writing code' to 'architecture, review, and guidance.' You specify what to build, validate what the agent outputs, and course-correct when it goes wrong.
- Willison emphasizes that Agentic Engineering only works if the developer can read and understand the code the agent generates. The ability to review AI output is a core skill.
- The guide argues that context management — effectively communicating constraints, existing architecture, and requirements to the agent — is now a critical competency.
- He also warns that agents often take unexpected shortcuts or introduce technical debt, so systematic review processes and test automation are even more important in this paradigm.
Evidence
- Commenters largely agreed with Willison's framing, with many sharing their own experiences where using AI agents felt very different from vibe coding — because they still needed to understand the full system to guide the agent effectively.
- Several developers noted that 'agentic engineering' is really just 'engineering with better tools' — the fundamentals haven't changed, but the leverage has increased dramatically.
- Some pushed back, arguing the distinction between vibe coding and agentic engineering is blurry in practice. Where exactly is the line between 'fully understanding the system' and 'understanding enough'?
- The observation that junior developers doing vibe coding often create unmaintainable messes while experienced engineers using agents dramatically accelerate output resonated with many commenters.
How to Apply
- Before starting an agent session, take time to document the current system architecture and key constraints. This 'context document' becomes your primary tool for guiding the agent.
- Treat every piece of code the agent generates as code you wrote yourself — you're responsible for understanding and maintaining it. Don't merge anything you don't understand.
- Build a review checklist: security, performance, test coverage, consistency with existing patterns. Apply it systematically to agent output.
- When the agent goes in a wrong direction, don't just retry — figure out why it went wrong and revise your prompt or context document before the next attempt.
Terminology
Related Papers
Migrating a production AI agent to GPT-5.6: 2.2x faster, 27% cheaper
마케팅 웹사이트를 자동 생성하는 프로덕션 AI 에이전트를 Claude Opus 4.8에서 GPT-5.6 Sol로 전환한 실전 경험담으로, 단순 모델 교체가 아니라 eval 하네스, 툴 스키마, 캐싱, 추론 리플레이까지 손봐야 했던 과정을 구체적인 수치와 함께 정리했다.
What xAI's Grok build CLI sends to xAI: A wire-level analysis
xAI의 공식 코딩 CLI 도구 Grok Build가 사용자 동의 없이 전체 Git 저장소와 .env 시크릿 파일을 xAI 서버로 업로드한다는 사실이 네트워크 트래픽 분석으로 밝혀졌다.
Remember When It Matters: Proactive Memory Agent for Long-Horizon Agents
LLM 에이전트가 긴 작업 중 중요한 정보를 잊어버리는 문제를 별도의 메모리 에이전트가 '적절한 타이밍에' 끼어들어 해결하는 방법
WebSwarm: Recursive Multi-Agent Orchestration for Deep-and-Wide Web Search
복잡한 웹 검색을 재귀적으로 분해하고 각 노드에 적합한 검색 모드를 동적으로 할당하는 멀티에이전트 프레임워크
Show HN: Reverse-engineering web apps into agent tools
로그인된 웹 앱의 API 호출을 브라우저에서 감시해 자동으로 MCP 도구로 변환하는 에이전트를 만들었다. 소스 코드나 공식 API 문서 없이도 Jira, Spotify 같은 서비스에 AI 어시스턴트를 붙일 수 있다.
Show HN: FableCut – A browser video editor AI agents can drive (zero deps)
타임라인 전체를 JSON 파일 하나로 표현하고 MCP/REST로 AI 에이전트가 직접 편집할 수 있는 브라우저 비디오 에디터로, Claude 같은 AI가 프롬프트 하나로 영상을 자동 컷편집하고 결과를 실시간으로 UI에 반영해준다.