Inference accelerator for modern workloads
RNGD is presented as FuriosaAI’s flagship accelerator for enterprise and cloud inference, with support for LLM and multimodal deployment.
FuriosaAI provides AI accelerators and server hardware for enterprise inference, optimized for LLMs, multimodal models, and data-center workloads.

FuriosaAI designs AI accelerators for data-center inference, with RNGD as its flagship product and the NXT RNGD Server as a packaged deployment option. The company frames the hardware around high-performance, power-efficient execution for computer vision, generative AI, LLMs, and agentic workloads.
The product family centers on Tensor Contraction Processor architecture, which FuriosaAI describes as a hardware-software approach optimized for tensor contraction. The site emphasizes deployment in enterprise and cloud settings, including air-cooled data centers, on-premises installations, managed environments, and colocation facilities, supported by a software stack for compilation, optimization, and production rollout.
RNGD is presented as FuriosaAI’s flagship accelerator for enterprise and cloud inference, with support for LLM and multimodal deployment.
The platform uses Tensor Contraction Processor architecture, which FuriosaAI says is built around tensor contraction rather than fixed matmul primitives.
RNGD is described with a 180W power profile and the NXT RNGD Server with a 3 kW power consumption target for air-cooled data centers.
Furiosa Software provides compilation, optimization, and production deployment workflows for LLM inference and agentic workloads.
The source mentions PyTorch 2.x integration, along with containerization, SR-IOV, Kubernetes, and other cloud-native components.
RNGD includes PCIe P2P support for LLMs, BF16/FP8/INT8/INT4 support, multiple-instance and virtualization features, and secure boot with model encryption.
Run high-throughput LLM inference in data centers where power density and rack utilization matter, using RNGD or the NXT RNGD Server as the deployment target.
Deploy multimodal models in cloud or enterprise environments where the site highlights low-latency execution and production-ready software support.
Evaluate, integrate, and qualify accelerators through the Furiosa Access Program before moving to production deployment.
Package inference into an air-cooled appliance for on-premises, managed, or colocation environments using the NXT RNGD Server.
Use the Furiosa software stack to compile, optimize, and ship models with PyTorch 2.x and cloud-native tooling.
FuriosaAI positions RNGD as an AI accelerator for enterprise and cloud inference. The source describes it as supporting high-performance LLM and multimodal deployment capabilities, and the NXT RNGD Server as a 3 kW inference appliance for agentic systems.
The source states that Furiosa Software provides a toolchain for LLM inference and agentic workloads, from compilation and optimization to production deployment. It also mentions PyTorch 2.x integration and containerization, SR-IOV, and Kubernetes support.
The Furiosa Access Program is described as a structured path for customers and partners to evaluate, integrate, qualify, and deploy Furiosa accelerators. It is available worldwide through online and offline access.
The source does not show public pricing. The pricing URL returns a not found page, so commercial packaging appears to require direct contact or another sales path rather than self-serve pricing on the site.