
Long-Context Prompting: Working with Million-Token Windows
Million-token context windows change everything. Here's how to prompt effectively when your entire codebase, library, or conversation fits in a single request.
Posts tagged prompt caching collect Will Spurlock's writing on this topic. The three newest excerpts: Million-token context windows change everything. Here's how to prompt effectively when your entire codebase, library, or conversation fits in a single request. OpenAI DevDay 2024 introduces the Realtime API for voice apps, automatic Prompt Caching with 50% discounts, and Model Distillation for creating efficient small models from frontier outputs. Anthropic launches Prompt Caching in public beta, cutting repeat-context costs by up to 90%. Here's how I structure prompts to trigger the API and drive production cost savings.
3 posts as of 2025-07-17, counted from content/blog frontmatter
Discover insights and strategies for leveraging technology in business
Posts tagged prompt caching collect Will Spurlock's writing on this topic. The three newest excerpts: Million-token context windows change everything. Here's how to prompt effectively when your entire codebase, library, or conversation fits in a single request. OpenAI DevDay 2024 introduces the Realtime API for voice apps, automatic Prompt Caching with 50% discounts, and Model Distillation for creating efficient small models from frontier outputs. Anthropic launches Prompt Caching in public beta, cutting repeat-context costs by up to 90%. Here's how I structure prompts to trigger the API and drive production cost savings.
3 posts as of 2025-07-17, counted from content/blog frontmatter.
Start with "Long-Context Prompting: Working with Million-Token Windows". Million-token context windows change everything. Here's how to prompt effectively when your entire codebase, library, or conversation fits in a single request.