Image of Furisoa RNGD
Image of Furisoa RNGD

Furiosa RNGD accelerates Ai transformation

FuriosaAI and Samsung SDS launch Korea’s first domestic NPUaaS to expand enterprise AI access

News
​
FuriosaAI and Samsung SDS launch Korea’s first domestic NPUaaS to expand enterprise AI access

FuriosaAI expands European AI infrastructure with RNGD deployment at Equinix’s Lisbon data center

News
​
FuriosaAI expands European AI infrastructure with RNGD deployment at Equinix’s Lisbon data center

Furiosa SDK 2026.3: A new kernel framework, and the models it unlocks

Technical Updates
​
Furiosa SDK 2026.3: A new kernel framework, and the models it unlocks

FuriosaAI and Broadcom partner on next-gen inference for Agentic AI

News
​
FuriosaAI and Broadcom partner on next-gen inference for Agentic AI
Run 30B+ frontier models with the latest release of SDK 2026.4
Close-up of a Furiosa AI acceleration server with multiple drive bays, status LEDs, and a power/reset button on the left handle.

Furiosa NXT RNGD Server

Enterprise-ready 3kW inference appliance for agentic systems
​
NXT RNGD Server
Learn more

Furiosa RNGD PCIe Card

The most performant and efficient PCIe inference card
​
RNGD PCIe Card
Learn more

Drop-in GPU replacement. Migrate your entire client code today.

Furiosa-llm delivers exceptional inference performance for advanced multimodal and agentic workloads.
​
Go to Furiosa-LLM Docs
​
Explore our software

shell

pip install furiosa-llm
furiosa-llm serve [model]
# Serving on :8000 (OpenAI-compatible)
from openai import OpenAI

client = OpenAI(
    base_url="http://localhost:8000/v1"
)
client.chat.completions.create(model="[model]", ...)
from furiosa_llm import LLM

llm = LLM(
    model="[model]",
    dtype="[precision]",
)
out = llm.generate(prompts)
Samsung logo
LG logo
Kakao logo
NAVER logo
Logo of Upstage
Logo of Aramco
lablup logo
Daum logo
Nota AI logo
Okestro logo
sea logo
Logo of LG U+
A narrow aisle in a data center with tall racks of servers and network cables on both sides, bathed in red lighting.

Start Your RNGD Evaluation

Validate serving performance and token rates against your target production SLOs. Choose between direct bare-metal access for deep hardware benchmarking or instant cloud API endpoints to test model performance with zero setup. Available worldwide.
​
Request access
Three horizontal sliders or toggle controls stacked vertically on a square gray panel with rounded corners, each slider has a black outline with two black lines inside representing the slider positions.

Bare metal

Benchmark compiler efficiency, memory bandwidth, and sustained serving concurrency.
Gray square button with a black outlined cloud icon in the center on a slightly textured background.

API endpoint

Measure token generation rates and streaming latency across your target models.
PARTNERS
Logo of Equinix
Logo of DCP AI Cloud Center
Logo of MEGAZONE Cloud
Logo of GUC The Advanced ASIC Leader
​
Request Access

Blog

​
See all

FuriosaAI establishes Singapore hub to drive APAC expansion and accelerate global RNGD deployment

News
​
FuriosaAI establishes Singapore hub to drive APAC expansion and accelerate global RNGD deployment

FuriosaAI at ICML 2026: Advancing full-stack software efficiency

News
​
FuriosaAI at ICML 2026: Advancing full-stack software efficiency

White Paper: Benchmarking RNGD on Backend.AI for 1.3–1.5x greater efficiency in enterprise AI inference

Technical Updates
​
White Paper: Benchmarking RNGD on Backend.AI for 1.3–1.5x greater efficiency in enterprise AI inference