d-Matrix builds ultra-low-latency, energy-efficient AI inference infrastructure for generative AI in the datacenter. Its memory-centric compute architecture, next-generation I/O, and stacked-DRAM approach are delivered through the Corsair inference accelerator (chiplet-based, supporting models up to 100B parameters), the JetStream… - View it on GitHub
Star
0
Rank
14385819