Aviator software: The only way to pilot your model
Corsair’s architecture was designed with software in mind. Our Aviator software stack is performant and easy to use, offering a seamless user experience.
Scale your inference-time compute fearlessly to unlock new levels of autonomy and intelligence with Corsair and JetStream!
Meet skyrocketing AI demands with JetStream™, purpose-built I/O accelerator for AI inference
Power AI operations with millions of requests using our custom NIC built for accelerator-to-accelerator communications.
Deploy AI with blazing fast multi-node scale-up
PCIe form factor for easy plug-and-play with existing datacenter infrastructure
Unique Digital-In Memory Compute Architecture (DIMC™) Purpose-Built for AI Inference Acceleration
Integrated Performance Memory for Blazing Fast Interactivity
Off-chip Capacity Memory for Offline Batched Inference
Block Floating Point Numerical Formats for Inference Efficiency
Aviator software: The only way to pilot your model
Corsair’s architecture was designed with software in mind. Our Aviator software stack is performant and easy to use, offering a seamless user experience.
Aviator Software
Integrates with broadly adopted open frameworks such as PyTorch and Triton DSL to allow developers to easily deploy their model. Built with open-source software such as MLIR, PyTorch, OpenBMC.
How it works
Hover over icons for more information
Tools
A comprehensive suite of tools to deploy and monitor d-Matrix cards, and analyze workload performance
Model Factory
A template of PyTorch models that enables distributed inference across multiple d-Matrix cards
Compressor
A model compression toolkit to enable efficient block floating point numerical formats supported by d-Matrix hardware
Compiler
A state-of-the-art compiler that leverages MLIR ecosystem and produces optimized code targeted for d-Matrix hardware
Inference Engine
A distributed inference execution engine that deploys inference requests efficiently to a cluster of d-Matrix hardware
Host Runtime
A runtime that manages host interaction with the d-Matrix cards
Chip Runtime
On-chip firmware that launches jobs asynchronously from the host and enables high utilization of the d-Matrix devices
Tools – A comprehensive suite of tools to deploy and monitor d-Matrix cards, and analyze workload performance
Model Factory – A template of PyTorch models that enables distributed inference across multiple d-Matrix cards
Compressor– A model compression toolkit to enable efficient block floating point numerical formats supported by d-Matrix hardware
Compiler – A state-of-the-art compiler that leverages MLIR ecosystem and produces optimized code targeted for d-Matrix hardware
Inference Engine– A distributed inference execution engine that deploys inference requests efficiently to a cluster of d-Matrix hardware
Host Runtime – A runtime that manages host interaction with the d-Matrix cards
Chip Runtime – On-chip firmware that launches jobs asynchronously from the host and enables high utilization of the d-Matrix devices
Scalable
System software easily integrates into inference servers for process spawning, pipeline management, and scale-out communication.
Frictionless
Compiler and runtime produces highly optimal placements and schedules for graph execution without manual intervention.