d-Matrix
  • Technology
  • Product
  • Ecosystem
  • Blog
  • About
  • Careers

d-Matrix Blog

Featured

Scaling AI the Right Way: Introducing Our Rack-Level Inference Solution  

Scaling AI the Right Way: Introducing Our Rack-Level Inference Solution  

October 14, 2025
What is AI Inference and why it matters in the age of Generative AI

What is AI Inference and why it matters in the age of Generative AI

June 4, 2025
The Complete Recipe to Unlock AI Reasoning at Enterprise Scale

The Complete Recipe to Unlock AI Reasoning at Enterprise Scale

February 13, 2025
How disaggregated AI inference pipelines generate less heat — and smaller bills 

How disaggregated AI inference pipelines generate less heat — and smaller bills 

Modern GPUs, whether they’re just sitting there idling or running hot during inference, cost enormous sums to keep them from melting into a puddle of lost throughput potential.  Those GPUs — drawing more than a kilowatt… Read More
August 12, 2026
Fast token generation emerges as the key differentiator as heterogeneous inference takes hold

Fast token generation emerges as the key differentiator as heterogeneous inference takes hold

d-Matrix Acquires Wallaroo.ai to Speed up Deployment of Heterogeneous AI Inference Workloads

d-Matrix Acquires Wallaroo.ai to Speed up Deployment of Heterogeneous AI Inference Workloads

The new perf/TCO math: maximizing GPU utilization with disaggregated pipelines

The new perf/TCO math: maximizing GPU utilization with disaggregated pipelines

Advancing Full-Model AI Inference on Corsair with Infinity

Advancing Full-Model AI Inference on Corsair with Infinity

IA for AI: The New Reader That Doesn’t Use the Nav Bar 

IA for AI: The New Reader That Doesn’t Use the Nav Bar 

Parasail to Combine NVIDIA AI Infrastructure with d-Matrix Accelerators to Achieve 10x Faster Token Generation 

Parasail to Combine NVIDIA AI Infrastructure with d-Matrix Accelerators to Achieve 10x Faster Token Generation 

View All Posts >

Trending

Blazing the Trail Toward More Scalable, Affordable AI with 3DIMC 

How to Bridge Speed and Scale: Redefining AI Inference with Ultra-Low Latency Batched Throughput

Going Vertical: Why we created a 3D DRAM solution to advance low latency AI inference

Transforming AI: d-Matrix’s Pivotal Moments in Pursuit of Gen AI Inference At Scale

What is AI Inference and why it matters in the age of Generative AI

Featured Video

The DeepSeek Moment

In this short talk, d-Matrix CTO Sudeep Bhoja discusses the release of the Deep Seek R1 model, highlighting its impact on inference compute. He discusses the evolution of reasoning models and the significance of inference time compute in enhancing model performance.

Learn more about d-Matrix

From the Media

The Cube

Fast token generation emerges as the key differentiator as heterogeneous inference takes hold

d-Matrix | Wallaroo logos

d-Matrix Acquires Wallaroo.ai to Speed up Deployment of Heterogeneous AI Inference Workloads

Parasail to Combine NVIDIA AI Infrastructure with d-Matrix Accelerators to Achieve 10x Faster Token Generation 

LinkedIn AI Breakthrough award

d-Matrix Corsair Inference Accelerator Wins 2026 AI Breakthrough Award

Upstart chipmakers keep challenging Nvidia. This time it’s Microsoft-backed D-Matrix

d-Matrix founders with Corsair

d-Matrix Corsair AI Inference Platform Enters Full Production to Meet Customer Demand

View all media articles > For all press inquiries, please email pr@d-matrix.ai>
Transforming AI from
unsustainable to attainable.
  • Technology
  • Product
  • Ecosystem
  • About
  • Careers
  • Blog
  • Newsletter
  • Media Kit
  • Contact
  • Privacy Policy
  • Terms of Use
© d-Matrix, Inc. 2026
X Twitter Logo Streamline Icon: https://streamlinehq.com