← All Articles
Tag

Model quantization

Articles tagged "Model quantization".

1 article since Aug 2026

LLM Cost Control in .NET: Debugging Billing Surprises in Production

Learn how to slash Azure OpenAI spend in .NET services with proven caching, model‑shrinking, and routing patterns—real‑world code, metrics, and a 60% cost‑cut case study.

Aug 29, 2026 7 min read
LLM cost optimization .NET Azure OpenAI Caching Model quantization Routing