Tag Archives: Accelerators

NVIDIA and Mediatek Ink $3.5B Investment Deal, Accelerate NVLink Fusion Adoption

Post Syndicated from Ryan Smith original https://www.servethehome.com/nvidia-and-mediatek-ink-3-5b-investment-deal-accelerate-nvlink-fusion-adoption/

NVIDIA and MediaTek have inked a new deal this week that more closely ties together the two companies financially and technologically. With NVIDIA investing $3.5B into the Taiwanese fabless chip designer, MediaTek will now offer the NVLink Fusion platform to customers designing custom XPUs at MediaTek

The post NVIDIA and Mediatek Ink $3.5B Investment Deal, Accelerate NVLink Fusion Adoption appeared first on ServeTheHome.

Google’s TPUv8s for Training and Inference at Hot Chips 2026

Post Syndicated from Ryan Smith original https://www.servethehome.com/googles-tpuv8s-for-training-and-inference-at-hot-chips-2026/

Hot Chips 2026 sees Google discussing its new eighth-generation TPU family for the technical crowd. One of the only hyperscalers to develop its own training hardware, the company has developed the TPU 8t for training, as well as the TPU 8i for inference

The post Google’s TPUv8s for Training and Inference at Hot Chips 2026 appeared first on ServeTheHome.

NVIDIA’s Groq 3 LPU Accelerators for Heterogeneous AI Compute at Hot Chips 2026

Post Syndicated from Ryan Smith original https://www.servethehome.com/nvidias-groq-3-lpu-accelerators-for-heterogeneous-ai-compute-at-hot-chips-2026/

The newest member of NVIDIA’s AI hardware family, at Hot Chips 2026 NVIDIA is diving into the use of LPUs as part of Vera Rubin clusters. The specialized chips from acquihire Groq are being tapped to offer significantly lower latency in the decode phase of inference

The post NVIDIA’s Groq 3 LPU Accelerators for Heterogeneous AI Compute at Hot Chips 2026 appeared first on ServeTheHome.

Intel Crescent Island 160GB to 480GB LPDDR5X AI GPU at Hot Chips 2026

Post Syndicated from Patrick Kennedy original https://www.servethehome.com/intel-crescent-island-160gb-to-480gb-lpddr5x-ai-gpu-at-hot-chips-2026/

We learned more about the new 160GB to 480GB LPDDR5X Intel Crescent Island GPU focused on memory capacity at Hot Chips 2026

The post Intel Crescent Island 160GB to 480GB LPDDR5X AI GPU at Hot Chips 2026 appeared first on ServeTheHome.

d-Matrix Raptor 3D-DRAM Accelerator for Generative Inference at Hot Chips 2026

Post Syndicated from Patrick Kennedy original https://www.servethehome.com/d-matrix-raptor-3d-dram-accelerator-for-generative-inference-at-hot-chips-2026/

At Hot Chips 2026, d-Matrix showed off its Raptor 3D-DRAM accelerator for AI breaking free of using HBM for memory by stacking DRAM and logic

The post d-Matrix Raptor 3D-DRAM Accelerator for Generative Inference at Hot Chips 2026 appeared first on ServeTheHome.

Cerebras Intros Faster WSE-3 Turbo Processor and First Rack-Scale CS-4 System

Post Syndicated from Ryan Smith original https://www.servethehome.com/cerebras-intros-faster-wse-3-turbo-processor-and-first-rack-scale-cs-4-system/

Cerebras this week has introduced a major upgrade to its hardware ecosystem. The company is launching their first rack-scale AI inference system, the CS-4, which is powered by the upgraded WSE-3 Turbo processor

The post Cerebras Intros Faster WSE-3 Turbo Processor and First Rack-Scale CS-4 System appeared first on ServeTheHome.

AMD Instinct MI455X Deep Dive: CDNA 5 Marks The Next Era of Instinct

Post Syndicated from Ryan Smith original https://www.servethehome.com/amd-instinct-mi455x-deep-dive-cdna-5-marks-the-next-era-of-instinct/

We are taking a deep dive look into AMD’s Instinct MI455X accelerator and its CDNA 5 architecture, the backbone of AMD’s next-gen AI server offerings and their massive Helios rackscale system

The post AMD Instinct MI455X Deep Dive: CDNA 5 Marks The Next Era of Instinct appeared first on ServeTheHome.

AMD Helios Architecture Deep Dive: The Power of AMD’s Hardware Combined

Post Syndicated from Ryan Smith original https://www.servethehome.com/amd-helios-architecture-deep-dive-amd-broadcom-hardware-combined/

The Helios rackscale system is the culmination of AMD’s server hardware, as well as their AI datacenter ambitions. For Advancing AI 2026, the company dove into the architecture of their first rackscale systems, outlining how they have scaled up 72 Instinct MI455X accelerators to act as a single system

The post AMD Helios Architecture Deep Dive: The Power of AMD’s Hardware Combined appeared first on ServeTheHome.

AMD Advancing AI 2026 Keynote Live Coverage

Post Syndicated from Ryan Smith original https://www.servethehome.com/amd-advancing-ai-2026-keynote-live-coverage/

It’s time for AMD’s 2026 Advancing AI conference. Join ServeTheHome for our live blog coverage of the AAI keynote, where AMD will announce their latest and greatest hardware and software technologies

The post AMD Advancing AI 2026 Keynote Live Coverage appeared first on ServeTheHome.

The AMD Instinct MI350P is a HBM PCIe AI Accelerator That Has Been All Over

Post Syndicated from Patrick Kennedy original https://www.servethehome.com/the-amd-instinct-mi350p-is-a-hbm-pcie-accelerator-that-has-been-all-over/

We have been seeing the AMD Instinct MI350P 144GB HBM3E PCIe accelerator everywhere over the past few weeks as this appears to be a popular GPU

The post The AMD Instinct MI350P is a HBM PCIe AI Accelerator That Has Been All Over appeared first on ServeTheHome.

Qualcomm Announces Dragonfly Brand for Data Center Products, More Info to Come June 24th

Post Syndicated from Ryan Smith original https://www.servethehome.com/qualcomm-announces-dragonfly-brand-for-data-center-products/

Qualcomm this week has introduced its Dragonfly brand for its upcoming data center products. The brief teaser promised more details to come on June 24th, during Qualcomm’s 2026 Investor’s Day

The post Qualcomm Announces Dragonfly Brand for Data Center Products, More Info to Come June 24th appeared first on ServeTheHome.

NVIDIA Computex 2026 News Bytes: Vera Rubin Now In Production, DGX Station Gets Windows

Post Syndicated from Ryan Smith original https://www.servethehome.com/nvidia-computex-2026-news-bytes-vera-rubin-now-in-production-dgx-station-gets-windows/

At Computex 2026, NVIDIA announced that its next-gen Vera Rubin platform is now in full production. The company is also bringing Windows to its high-end DGX Station systems, which will be available in Q4

The post NVIDIA Computex 2026 News Bytes: Vera Rubin Now In Production, DGX Station Gets Windows appeared first on ServeTheHome.

AMD Intros Instinct MI350P Accelerator: CDNA 4 Comes to PCIe Cards

Post Syndicated from Ryan Smith original https://www.servethehome.com/amd-intros-instinct-mi350p-accelerator-cdna-4-comes-to-pcie-cards/

AMD has released a PCIe version of its flagship MI350 accelerators, the MI350P. Half of a MI350X, the card is aimed at customers who need to fit an modern AI accelerator into a traditional PCIe server

The post AMD Intros Instinct MI350P Accelerator: CDNA 4 Comes to PCIe Cards appeared first on ServeTheHome.

Intel Announces Arc Pro B70 and B65 Video Cards: Big Battlemage Brings Big Memory for AI Workstations

Post Syndicated from Ryan Smith original https://www.servethehome.com/intel-announces-arc-pro-b70-and-b65-video-cards-big-battlemage-brings-big-memory-for-ai-workstations/

Today Intel is expanding their Arc B-series video card lineup in a big way, with the launch of a pair of new Arc Pro graphics cards: the Arc Pro B70 and the Arc Pro B65. Joining Intel’s existing Arc Pro B-series video cards, the latest cards out of Intel are also the company’s most powerful […]

The post Intel Announces Arc Pro B70 and B65 Video Cards: Big Battlemage Brings Big Memory for AI Workstations appeared first on ServeTheHome.

Decoding the Future of Inference At NVIDIA: Groq LPUs Join Vera Rubin Platform For Low-Latency Inference

Post Syndicated from Ryan Smith original https://www.servethehome.com/decoding-the-future-of-inference-at-nvidia-groq-lpus-join-vera-rubin-platform-for-low-latency-inference/

With its upcoming Vera Rubin rackscale architecture, NVIDIA is going to be integrating LPUs from acquihire Groq, marking a major expansion beyond using GPUs alone for AI inference

The post Decoding the Future of Inference At NVIDIA: Groq LPUs Join Vera Rubin Platform For Low-Latency Inference appeared first on ServeTheHome.