Towards Data Science

Changing One Prompt Can Affect 50 Others — I Built a Prompt Dependency Graph to Find What Needs Retesting

The article describes how the author constructed a prompt dependency graph to identify which prompts are affected when a single prompt changes. By separating all reachable components from the smaller subset that truly requires evaluation, the graph helps focus retesting efforts. This approach streamlines testing by pinpointing only the prompts that need targeted evaluation.

Towards Data Science
Jul 29

Prompt Engineering Is Solved—Prompt Management Isn’t

Prompt engineering helps you write better prompts—but it doesn’t help you change them safely. This article explores a common production failure where a simple variable rename breaks every live call, and introduces a lightweight static analysis tool that treats prompts like contracts, catching breaking changes before they ship.

By Emmimal P Alexander
Simon Willison
Sep 22

llm 0.36

The release of llm 0.36 introduces new OpenAI models gpt-6-sol and gpt-6-luna, and adds support for model plugins to declare that they do not support conversations via supports_conversation = False. When such models receive assistant or tool history, llm raises a ConversationNotSupported error and the chat interface rejects them before starting a session. Additional changes include wrapping reasoning traces in Markdown output with <details> tags and bug fixes from five contributors.