llm-cost-optimization
sickn33/agentic-awesome-skills
This skill helps reduce LLM API and infrastructure costs by 50-90% using strategies like model selection, prompt caching, batching, quantization, and self-hosting. It covers cost tracking, right-sizing models for tasks, provider-side caching, and semantic caching to avoid redundant calls. Suitable for developers managing AI spend and optimizing cloud infrastructure.