Making the math behind frontier-level models work
Frontier AI is pushing today’s infrastructure to its limits. Inference demands a new approach as models grow larger and agentic AI becomes more complex.
This month, we’re showcasing how d-Matrix and our ecosystem partners are accelerating AI with heterogeneous, disaggregated infrastructure that delivers lower latency, higher efficiency, and better economics at scale. From strategic partnerships to software innovations, here’s what’s new.
Accelerating disaggregated infrastructure with d-Matrix + Parasail
d-Matrix has partnered with Parasail to deploy Corsair alongside NVIDIA Hopper and Blackwell GPUs, delivering up to 10x faster, more cost-efficient AI inference. See how heterogeneous infrastructure helps customers scale frontier AI workloads with lower latency and greater efficiency.
Advancing performant full-model inference with Infinity
The d-Matrix and Infinity teams have brought Qwen3 from initial tensor-parallel operations to complete, stateful inference. See how purpose-built AI infrastructure and software optimization can accelerate getting next-generation models online and running inference faster.
Interested in more from d-Matrix?
Get the latest d-Matrix updates to your inbox. Sign up below:
d-Matrix Spotlight: Corsair wins 2026 AI Breakthrough Award
Corsair has been recognized with the 2026 AI Breakthrough Award for AI Processor Innovation, highlighting d-Matrix's leadership in purpose-built AI inference.
Using attention-FFN disaggregation to accelerate AI pipelines
See how attention-FFN disaggregation works and how heterogeneous hardware deployment in disaggregated inference infrastructure makes pipelines powered by even massive frontier models more efficient and significantly faster.
Technology | Product | Ecosystem | About
d-Matrix, 5201 Great America Pkwy, Ste 300, Santa Clara, CA 95054, United States


