AMD and Marvell have both performed well over the past year.
"Disaggregated Inference," promises better utilization, lower costs, and faster AI responses. Major players like Nvidia/Groq and AMD/Cerebras will vie for the prize.
11hon MSN
OpenAI’s Jalapeño AI chip brings new 'threat' to Nvidia margins as custom silicon gains ground
OpenAI’s Jalapeño chip beat Nvidia Blackwell systems on key inference-efficiency tests as custom AI silicon gains ground among major tech companies.
The pilot stage is the best time to consider the implications of architecture, security, and operations for running AI ...
The better AI stock to own for the next five years may not be the one investors expect.
Samsung LPDDR5X-PIM, presented at Hot Chips 2026 on August 25, delivers 3.01x AI token throughput versus standard LPDDR5X, ...
OpenAI built the chips in record time by using its own AI models to speed up the design and verification phase of development.
As agents reason, replan, call other agents, and work continuously in the background, Gartner predicts inference costs per workflow will rise more than fivefold through 2028.
The organizations that treat infrastructure as a supporting function may watch their AI ambitions outgrow the platforms built ...
Nvidia CEO Jensen Huang unveils a high-speed AI inference system using Groq technology, targeting growing demand.
Nvidia's dedicated inference accelerator Groq 3 LPX enters full production to supercharge AI agents - SiliconANGLE ...
The investment seeks to track the total return performance, before fees and expenses, of the BITA AI Inference Chip Select Index (the “Index”). Under normal circumstances, at least 80% of the fund’s ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results