Prompts Aren't Real
Summary
The article argues that prompts are not the core driver of LLM behavior and advocates for building interlocking evaluation and optimization pipelines. It describes a workflow using pass^k tests, automated prompt optimization, holdout testing, and production monitoring to create reliable agent-based systems, with emphasis on measurement over manual prompting.