Memory-centric AI inference silicon and systems
d-Matrix is a Santa Clara semiconductor company building purpose-built hardware for AI inference. Its public materials describe moving compute into memory to reduce latency and improve efficiency versus traditional architectures that keep compute farther from memory. The company has announced collaboration with NVIDIA around NVLink Fusion rack-scale integration for its next-generation Raptor XPUs.