DIGITIMES Intelligence observes that AI applications are moving through a clear progression: from early classification AI focused on feature recognition, to generative AI capable of...
Cerebras Systems is expanding its push into AI inference infrastructure through a collaboration with Gimlet Labs, as demand grows for computing systems capable of delivering faster...
Calls to slow frontier AI model training have sparked debate across the AI semiconductor industry, but executives at the AI Infra Summit in Santa Clara, California, said the real buildout...
Market sources say Apple is developing an enterprise AI server that could use 2-4 unreleased M8 Ultra processors, aimed at AI developers, large enterprises and government agencies...
Apple is exploring a return to the enterprise server market with an AI inference system built around future M8 Ultra chips, potentially using Nvidia's NVLink Fusion interconnect te...
Samsung Electronics is investing in Dutch AI chip startup Euclyd as part of a more than EUR200 million (about US$230.9 million) Series A funding round, highlighting growing efforts...
Astera Labs has added three products to its Leo memory controller line, aimed at a problem that arises as AI inference shifts from single queries to agents that run continuous loops:...
AI chip startup d-Matrix is adopting Nvidia's NVLink Fusion technology to bring its next-generation Raptor inference processors into Nvidia MGX racks, giving the company a path into...
Qualcomm this week announced a multi-generation collaboration with Amazon to develop custom silicon for AI inference and high-performance optical connectivity for Amazon Web Services...
Qualcomm's multi-generation agreement to build customized AI inference silicon and optical interconnects for Amazon gives the company something its data center roadmap has lacked since...
As end-to-end (E2E) autonomous driving architectures gradually become mainstream, the inference demands of physical AI are driving radical transformations in automotive system-on-chip...
China-based ICY Tech has introduced its uHBM and uLPU AI inference architectures, extending persistent MRAM computing from validation silicon into a product roadmap spanning compute-memory...