AIToday
Large Language ModelsOpen-Source AITop Companies' AI MovesAI Business & IndustryTop Companies AIPublished: Aug 24, 2026, 06:32 JST2 min read

AT&T slashes AI costs 56% with model routers

AT&T slashes AI costs 56% with model routers

Key takeaway

  • AT&T cut AI costs by up to 56% using model routers that send simpler tasks to cheaper models.

  • Quality dropped only 2%.

  • The company plans to increase open-source model usage from 40% to 60-70% of employee queries.

3 Key Points

  1. What happened

    AT&T cut the costs of coding and some other advanced AI tasks by as much as 56% by using LiteLLM model routers, which route employee queries to cheaper AI models when appropriate. The quality of the AI's performance declined by only 2%, according to an interview with AT&T vice president Mark Austin.

  2. Why it matters

    The tools decide how complex a task is and send simpler ones to cheaper models, an approach that lets the telecom keep spending on Anthropic and OpenAI models flat while expanding use of open-source models like Nvidia's Nemotron, Meta's Llama, and Google's Gemma. AT&T aims to raise the share of employee queries powered by open-source models from 40% to between 60% and 70% in the coming years.

  3. What to watch

    AT&T says open-source model capabilities generally lag frontier models by six to 10 months, but the gap is narrowing and they are "just as good or better" than older Anthropic and OpenAI models. The company is evaluating the potential risks of using open-source models from Chinese firms DeepSeek and Moonshot but is not using them yet.

Ask the AI about this article →

Context & Analysis

AT&T's reported cost savings arrive as businesses grapple with rising AI expenses, a trend highlighted in June reports about companies seeking better cost management. The shift from chatbots to more compute-intensive agents, along with AI labs moving from flat subscriptions to token-based billing, has driven costs upward. This context helps explain why a large enterprise would invest in routing technology rather than simply cutting AI usage.

The company's strategy pairs cost-cutting routers with a deliberate push toward open-source models, whose capabilities generally trail frontier models by six to 10 months. Austin's observation that this gap is narrowing suggests AT&T sees open-source as an increasingly viable alternative to premium Anthropic and OpenAI models for many internal tasks. The evaluation of DeepSeek and Moonshot models indicates interest in expanding options further, though risk assessment remains a hurdle.

The reported end of "tokenmaxxing" — pushing employees toward the biggest models and heaviest usage — aligns with AT&T's approach. By matching task complexity to model capability, the company appears to be institutionalizing a more selective, cost-conscious AI strategy that other firms may study.

FAQ

How did AT&T achieve the cost reduction?
AT&T used LiteLLM model routers, which assess task complexity and decide whether a cheaper AI model can handle it. This approach cut costs by up to 56% with only a 2% decline in performance quality.
Which open-source models is AT&T using?
AT&T is using open-source or open-weight models such as Nvidia's Nemotron, Meta's Llama, and Google's Gemma. It is not using models from Chinese firms DeepSeek and Moonshot, but is evaluating potential risks.
Top Companies AIRead Original Article

Get the latest Large Language Models news every morning

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytime

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleMerck acquires AI vaccine stake in Evaxion