← Back to topics page

Articles about "inference"

Nvidia’s Vera Chip Targets $200 Billion AI Inference Market

Nvidia has unveiled its new Vera central processor, aiming to capture a $200 billion market focused on AI inference workloads. CEO Jensen Huang highlighted the Vera chip as a critical component in Nvidia’s strategy to maintain its AI hardware leadership as major cloud providers and chip competitors ramp up their own silicon offerings. Despite record earnings, the company faces supply constraints and intensified competition in the AI hardware sector.

dataconomy.com

Samsung Begins Mass Production of Nvidia Groq 3 LPU on 4nm Process

Samsung Electronics has started large-scale manufacturing of Nvidia's Groq 3 Language Processing Unit using its 4-nanometer process, marking a key development for both companies in the artificial intelligence hardware sector. The move underscores Nvidia’s ongoing strategy to broaden its AI hardware beyond traditional GPUs, while positioning Samsung as a major foundry partner in the global AI chip race.

dataconomy.com

NVIDIA Blackwell Ultra Offers Major Efficiency Gains for Agentic AI

NVIDIA’s Blackwell Ultra GPUs deliver up to 50 times higher efficiency for agentic AI and coding assistant workloads, according to new benchmark results. The platform significantly reduces operational costs for cloud providers and enables wider deployment of advanced language models. Cloud infrastructure firms including CoreWeave, Microsoft, and Oracle have begun large-scale deployments using this technology.

The Essential Weekly Update

Stay informed with curated insights delivered weekly to your inbox.