Skip to main content
AI Products

d-Matrix to integrate inference XPUs with Nvidia GPUs via NVLink Fusion

d-Matrix CEO Sid Sheth said the startup has formed a partnership with Nvidia that will integrate its inference-focused XPUs with Nvidia GPUs using NVLink Fusion, according to a Bloomberg Technology interview. The arrangement places d-Matrix's inference accelerators inside Nvidia's interconnect ecosystem as demand shifts toward specialized systems built for inference rather than training. Nvidia is embracing the rise of custom AI chips, and Sheth described inference as emerging as the next major battleground in AI compute. The segment did not disclose commercial terms, deployment volumes, or a timeline for the integration.
Sources
In this story
Published by Tech & Business, a media brand covering technology and business. This story was sourced from Bloomberg Technology and reviewed by the T&B editorial agent team.
Back to Newswire
Keep reading
Full wire
AI Products
AI Products

Slack adds AI-built interactive reports and dashboards inside chats

Slack said a new feature, Slackforce Surfaces, will let users describe a report, dashboard, poll, presentation or microsite to Slackbot, which uses AI to assemble it from relevant conversations and connected apps such as Google Dr...

AI Products
AI Products

Amazon Quick desktop AI assistant reaches general availability

Amazon says its Quick desktop application is now generally available on macOS and Windows, following a preview used by customers in manufacturing, healthcare and sports. The company claims Quick runs on AWS infrastructure custome...

AI
AI

AWS introduces Ray Serve container for model inference

AWS has introduced a Ray Serve Deep Learning Container for inference workloads, positioning it as a migration option for teams using unmaintained TorchServe. AWS says the image bundles PyTorch, Ray Serve, FastAPI, Uvicorn and GPU-...