ComputeLabs Research
d-Matrix presented its Raptor generative-inference accelerator, stacking three-dimensional DRAM and logic instead of relying on HBM.
· ComputeLabs Research · from the August 23, 2026 edition
d-Matrix presented its Raptor accelerator at Hot Chips 2026. The device is designed for generative-AI inference, distinguishing its target workload from general-purpose central processing units or training-focused accelerator systems.
Raptor uses a three-dimensional architecture that stacks Dynamic Random-Access Memory (DRAM) and logic. ServeTheHome characterized the design as moving away from reliance on High Bandwidth Memory (HBM).
The supplied source did not provide performance, power, memory-capacity, pricing or shipment-volume figures. It also did not identify a server configuration, cloud deployment or named customer associated with Raptor.
Sources
- HBM

