Cut LLM API Token Costs by 90% — 5 Tactics from Prompt Caching to Model Tiering

Answer Capsule You can cut your monthly LLM API bill to less than half with a combination of five tactics: Prompt Caching (up to 90% off), separating the system prompt, early termination of streaming, Model Tiering (routing to Haiku/Flash), and RAG-based context management. All figures are based on official price sheets. Run an LLM API … Read more

Claude Code Max Subscription vs API Key — Which Costs Less in 2026?

This post is part of the Coupang Partners program, and I earn a commission from qualifying purchases. Answer Capsule — If you use Claude Code every day, a Max subscription ($100 or $200) is the clear win. If you only reach for it occasionally, or if you wire it into team automation, per-token API key … Read more