• News In Brief
  • Awards nights
  • AI
  • Education
  • Pro AV
  • Case Study
  • Interview
No Result
View All Result
SUBSCRIBE
Smart Solutions World
  • News In Brief
  • Awards nights
  • AI
  • Education
  • Pro AV
  • Case Study
  • Interview
No Result
View All Result
No Result
View All Result
Home AI

Gartner Predicts That by 2030, Performing Inference on an LLM With 1 Trillion Parameters Will Cost GenAI Providers Over 90% Less Than in 2025

SmartSolutionUser1 by SmartSolutionUser1
March 30, 2026
in AI
0
Gartner Predicts That by 2030, Performing Inference on an LLM With 1 Trillion Parameters Will Cost GenAI Providers Over 90% Less Than in 2025
76
SHARES
1.3k
VIEWS
Share on FacebookShare on Twitter

By 2030, performing inference on a large language model (LLM) with one trillion parameters will cost GenAI providers over 90% less than it did in 2025, according to Gartner, Inc. a business and technology insights company.

You might also like

Vigyanlabs Launches FEMTO, a Micro Data Centre Built for the Next Generation of AI Infrastructure

AI Doesn’t Need More Intelligence, It Needs Protection – Coditation

JFrog Partners with Wiz to Close the Gap on AI-Era Threats, Keeping Global Businesses Secure

AI tokens are the units of data that GenAI models process. For the purposes of this analysis a token is 3.5 bytes of data, or approximately 4 characters.

Mr. Will Sommer, Sr. Director Analyst at Gartner.
Mr. Will Sommer, Sr. Director Analyst at Gartner.

“These cost improvements will be driven by a combination of semiconductor and infrastructure efficiency improvements, model design innovations, higher chip utilization, increased use of inference-specialized silicon, and application of edge devices for specific use cases,” said Mr. Will Sommer, Sr. Director Analyst at Gartner.

As a result of these trends, Gartner forecasts LLMs in 2030 will be up to 100 times more cost-efficient than the earliest models of similar size developed in 2022.

The forecasted model results are split between two sets of semiconductor scenarios:

  • Frontier scenarios: Model processing is based on a representation of cutting edge chips.
  • Legacy blend scenarios: Model processing is based on a representative blend of available semiconductors benchmarked to Gartner forecasts.

Modeled costs in the “blend” forecast scenarios are considerably higher than in the “frontier” scenarios, given lower computational power (see Figure 1).

Figure 1: Gartner GenAI Inference Cost Scenario Forecasts

Source: Gartner (March 2026)

Falling Token Costs will not Democratize Frontier Intelligence

However, falling GenAI provider token costs will not be fully passed on to enterprise customers. Moreover, frontier intelligence will demand significantly more tokens than current mainstream applications. Agentic models, for example, require between 5-30 times more tokens per task than a standard GenAI chatbot, and can perform many more tasks than a human using GenAI.

While lower token unit costs will enable more advanced GenAI capabilities, these advancements will drive disproportionately higher token demand. As token consumption rises faster than token costs fall, overall inference costs are expected to increase.

“Chief Product Officers (CPOs) should not confuse the deflation of commodity tokens with the democratization of frontier reasoning,” said Sommer. “As commoditized intelligence trends toward near-zero cost, the compute and systems needed to support advanced reasoning remain scarce. CPOs who mask architectural inefficiencies with cheap tokens today will find agentic scale elusive tomorrow.”

Value will accrue to platforms that can orchestrate workloads across a diverse portfolio of models. Routine, high-frequency tasks must be routed to more efficient small and domain-specific language models, which perform better than generic solutions at a fraction of the cost when aligned to specialized workflows. Expensive inference of frontier-level models must be heavily gated and reserved exclusively for high-margin, complex reasoning tasks.

If you have an interesting Article / Report/case study to share, please get in touch with us at editors@roymediative.com roy@roymediative.com, 9811346846/9625243429.

Tags: Gartner Predicts That by 2030Performing Inference on an LLM With 1 Trillion Parameters Will Cost GenAI Providers Over 90% Less Than in 2025smart solutions world
Share30Tweet19
SmartSolutionUser1

SmartSolutionUser1

Recommended For You

Vigyanlabs Launches FEMTO, a Micro Data Centre Built for the Next Generation of AI Infrastructure

by SmartSolutionUser1
September 8, 2026
0
Vigyanlabs Launches FEMTO, a Micro Data Centre Built for the Next Generation of AI Infrastructure

Vigyanlabs Innovations Pvt. Ltd., a Mysuru-based deep-tech company, launched FEMTO, its new Sovereign AI-in-a-Box platform, marking an important step in the company's journey to build advanced AI infrastructure...

Read moreDetails

AI Doesn’t Need More Intelligence, It Needs Protection – Coditation

by SmartSolutionUser1
September 7, 2026
0
AI Doesn’t Need More Intelligence, It Needs Protection – Coditation

By Mr. Chetan Saundankar, Founder & CEO, Coditation Systems and Plant360.AI There's an old, simple idea behind any real promise of protection: someone has your back even when...

Read moreDetails

JFrog Partners with Wiz to Close the Gap on AI-Era Threats, Keeping Global Businesses Secure

by SmartSolutionUser1
September 7, 2026
0
JFrog Partners with Wiz to Close the Gap on AI-Era Threats, Keeping Global Businesses Secure

JFrog Ltd, creators of the JFrog Software Supply Chain Platform, the system of record for trusted software artifacts, binaries, and AI assets, announced a new integration with Wiz,...

Read moreDetails

OpenAI launches GPT-6 Astra for the next-generation in intelligence for work 

by SmartSolutionUser1
September 7, 2026
0
OpenAI launches GPT-6 Astra for the next-generation in intelligence for work 

OpenAI has launched GPT-6 Astra, the most intelligent and aligned model OpenAI has ever released.  Astra represents a generational leap in capability, delivering state-of-the-art performance on computer use,...

Read moreDetails

Mindsprint launches conversational AI PR Agent on Procuresprint

by SmartSolutionUser1
September 7, 2026
0
Mindsprint launches conversational AI PR Agent on Procuresprint

Mindsprint, a technology firm offering purpose-built AI-led solutions to modernize enterprise operations, and a Wipro company, announced the launch of a new Purchase Requisition (PR) Agent within Procuresprint®,...

Read moreDetails
Next Post
Tech Mahindra Inks MoU with IIT Bombay to Build 3D Digital Twin to Enable Smart Infrastructure

Tech Mahindra Inks MoU with IIT Bombay to Build 3D Digital Twin to Enable Smart Infrastructure

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Browse by Category

Browse by Category

Smart Solutions World

We bring you the best Premium news, magazine, personal blog, etc. Check our landing page for details.

  • News In Brief
  • Awards nights
  • AI
  • Education
  • Pro AV
  • Case Study
  • Interview

BROWSE BY TAG

Agentic AI AI AI-powered Akamai AMD CloudKeeper Coforge CrowdStrike Cybersecurity Databricks Fortinet Gartner Google Cloud HCLTech Honeywell IBM India Infosys Kaspersky Keysight Kramer Microsoft New Relic Nvidia OpenAI Palo Alto Networks PPDS Qlik Qualcomm Seqrite ServiceNow SiMa.ai smart solutions world smartsolutionsworld smart solutions world latest news Snowflake Software Solutions Sophos Tata Communications Tech Mahindra Technology Tenable UiPath Vertiv

© 2024 NCN - Premium news & magazine by NCN.

No Result
View All Result
  • News In Brief
  • Awards nights
  • AI
  • Education
  • Pro AV
  • Case Study
  • Interview

© 2024 NCN - Premium news & magazine by NCN.

Not enough quota to unlock this post
Unlock left : 0
Are you sure want to cancel subscription?