Switch language한국어
Back to the list

Gartner predicts that by 2030, inference costs for LLMs with 1 trillion parameters will be reduced by more than 90%

TL;DR AI

Key summary

2 min read
  1. Gartner says inference costs for trillion-parameter LLMs could fall by more than 90% by 2030 versus 2025, driven by advances in semiconductors and model design.

  2. The firm warns that wider use of AI agents could increase token consumption, offsetting enterprise-wide cost savings.

  3. Gartner recommends using smaller models where possible and reserving large models for tasks that truly need them.

Read the original