Switch language한국어
Back to the list

Gartner Predicts 90%+ Inference Cost Cut for 1-Trillion-Parameter LLMs by 2030

TL;DR AI

Key summary

2 min read
  1. Gartner predicts that inference costs for 1-trillion-parameter LLMs will drop by over 90% by 2030 versus 2025.

  2. The firm says the reduction will come from combined improvements in semiconductors, infrastructure, model design, chip utilization, and inference silicon.

  3. Gartner presented two scenarios—Frontier (leading-edge chips) and Legacy Blend (existing semiconductor benchmarks)—with Legacy Blend showing higher absolute costs.

  4. Gartner warned cost-per-inference cuts may not lower enterprise AI spending because AI agents could process 5–30x more tokens.

  5. The firm advised using small domain-specific models for frequent tasks and large LLMs only for complex processing.

Read the original