← Back to Thinking
AI Thinking
A Zero-to-One AI Coding SOP: From One Sentence to Verifiable Delivery
Don't just tell AI to write code — clarify, propose, plan, execute in stages, and prove correctness with tests
How to Use Benchmarks When Building Agents
A benchmark is not an exam — it is an answer bank. Read leaderboards to choose, copy scoring methods to build your own eval set, slice to diagnose
From Loop to Loop Engineering
From prompt engineering to loop engineering — as AI's continuous autonomous work span grows, the engineering focus shifts