Google's 'Frozen v2' chip targets 6-10x efficiency gains for Gemini inference
Alphabet is developing a custom AI accelerator designed to reduce per-token power consumption and reduce dependence on Nvidia's dominance.
Alphabet is developing a custom AI accelerator designed to reduce per-token power consumption and reduce dependence on Nvidia's dominance.
Anthropic is exploring a collaboration with Samsung on a custom processor to reduce dependence on Nvidia, following OpenAI's recent inference chip announcement.
The inference-focused chipmaker has raised $800M total and secured $1B in system orders as competition in custom AI silicon intensifies.
OpenAI's custom inference chip, built with Broadcom, joins a wave of Big Tech companies reducing single-supplier risk in AI hardware.
OpenAI unveiled Jalapeño, an ASIC chip co-developed with Broadcom to handle AI inference workloads and reduce dependence on Nvidia GPUs.
OpenAI and Broadcom announced Jalapeño, a purpose-built AI chip for language model inference, designed to improve efficiency and reduce costs at scale.
The Korean chip startup raises Series B at $570M valuation, targeting the data-movement bottleneck that GPUs can't solve alone.
Snowflake's multi-billion dollar investment in Amazon's custom processors signals accelerating enterprise demand for alternatives to Nvidia in AI infrastructure.