What does this skill do?
Cost Optimization for LLM Skill provides a composable architecture to control and reduce spending on LLM APIs without sacrificing response quality. It combines intelligent model routing based on task complexity, an immutable cost tracker, selective retries in response to transient errors, and a prompt cache to minimize latency and cost per token in repeated iterations.