CONNECT WITH US
NEWS TAGGED AI INFERENCE
Wednesday 16 September 2026
Samsung backs Dutch AI chip startup Euclyd in US$230M funding round
Samsung Electronics is investing in Dutch AI chip startup Euclyd as part of a more than EUR200 million (about US$230.9 million) Series A funding round, highlighting growing efforts...
Tuesday 15 September 2026
Astera Labs targets the KV cache bottleneck as agentic AI outgrows GPU memory
Astera Labs has added three products to its Leo memory controller line, aimed at a problem that arises as AI inference shifts from single queries to agents that run continuous loops:...
Friday 11 September 2026
d-Matrix brings Raptor chips to Nvidia racks as falling token prices pressure GPU economics
AI chip startup d-Matrix is adopting Nvidia's NVLink Fusion technology to bring its next-generation Raptor inference processors into Nvidia MGX racks, giving the company a path into...
Thursday 10 September 2026
Qualcomm's AWS deal raises stakes for MediaTek and Alchip
Qualcomm this week announced a multi-generation collaboration with Amazon to develop custom silicon for AI inference and high-performance optical connectivity for Amazon Web Services...
Wednesday 9 September 2026
Qualcomm to build custom AI inference chips for AWS
Qualcomm's multi-generation agreement to build customized AI inference silicon and optical interconnects for Amazon gives the company something its data center roadmap has lacked since...
Friday 4 September 2026
Analysis: Physical AI upgrades autonomous driving and lifts Taiwan supply chains
As end-to-end (E2E) autonomous driving architectures gradually become mainstream, the inference demands of physical AI are driving radical transformations in automotive system-on-chip...
Thursday 3 September 2026
SanDisk pushes HBF ecosystem as AI inference reshapes token economics
The rapid rise of AI inference, multimodal inputs, and AI agents is giving NAND flash a new strategic role.
Wednesday 2 September 2026
Chinese MRAM chipmaker ICY Tech rethinks HBM for rack-scale AI inference
China-based ICY Tech has introduced its uHBM and uLPU AI inference architectures, extending persistent MRAM computing from validation silicon into a product roadmap spanning compute-memory...
Friday 21 August 2026
Google, Marvell long-term deal keeps Broadcom at bay, limits MediaTek impact

Marvell filed a Form 8-K with the US Securities and Exchange Commission on August 19, 2026, revealing that it signed a commercial agreement...

Thursday 20 August 2026
Analysis: Google keeps Sunfish at Broadcom, Zebrafish at MediaTek — Marvell's US$12.2bn warrant buys a category nobody was defending

One word in Marvell Technology's August 19 filing decides how the Google agreement should be read. The custom silicon programs, it says,...

Wednesday 19 August 2026
Google and AMD target post-GPU era with Frozen v2 and Taalas
Advanced Micro Devices (AMD) recently announced its acquisition of Canadian AI chip startup Taalas to sharpen its competitive edge in AI inference. Taalas' core technology bakes model...
Friday 7 August 2026
AMD to buy Taalas in push for faster AI inference chips

AMD said on August 6 that it has agreed to buy Taalas, a Toronto-based startup focused on specialized AI inference silicon, in a move...

Tuesday 4 August 2026
Sandisk and SK Hynix release HBF spec to advance AI memory standard
Sandisk and SK Hynix have published a technical specification for high-bandwidth flash through the Open Compute Project, a move that could shape future AI infrastructure worldwide...
Monday 27 July 2026
Rebellions readies Rebel100 shipments as AI inference demand grows
South Korean AI chip designer Rebellions plans to begin shipping its next-generation Rebel100 accelerator in the second half of 2026, betting that wider use of AI agents and commercial...