Dual custom Ampere compute
Two modified Nvidia Ampere GPUs tuned as one compact local inference appliance.
Available for pre-order
A compact local AI appliance for builders who want dual custom-modded Ampere compute, private data, and high-VRAM workflows without assembling a workstation.
Finished hardware, not a hobby rig
Duet AI is built for local model work that repeats all day: coding agents, document workflows, longer context experiments, and private inference. The enclosure keeps its custom-modded Ampere hardware controlled, cooled, and accessible through clean rear power, USB-C, and a direct development environment.
Custom Ampere architecture
Two modified Nvidia Ampere GPUs tuned as one compact local inference appliance.
More working memory for larger quantized models, longer contexts, and repeatable local runs.
Custom cooling and power tuning keep the appliance desk-friendly under active local workloads.
Average score by model
Duet AI average versus RTX 3090 average.
Across the listed model runs.
Gemma and Qwen benchmark averages.
| Model | Duet AI average | RTX 3090 average | Difference |
|---|---|---|---|
| Gemma 4 26B Q4/Q5 | 79 t/s | 88 t/s | 9 t/s lower |
| Qwen 3.6 27B Q4/Q5 | 72 t/s | 80 t/s | 8 t/s lower |
| Qwen 3.6 35B Q4/Q5 | 58 t/s | 64 t/s | 6 t/s lower |
Pre-order specification
Connection
Plug Duet AI into a PC and open the local terminal workflow.
After setup, connect it to the local network for cable-free access.
Duet AI runs an x86 Linux-based environment for local inference stacks.
Built and shipped
Assembly, setup validation, packaging, and production-slot allocation are handled directly by the team behind the product.
Benchmark videos
https://youtu.be/WHRsyEE-FRQ
https://youtu.be/rTVFknP98Cg
Technical questions
Duet AI includes an internal Intel N100 x86 system running Linux. You can update the Linux image or replace the operating system.
Yes. The GPU interface is not structurally changed, so Duet AI uses the native Nvidia driver stack and remains 100% compatible with standard drivers.
Duet AI exposes only USB-C. Internally, the 2.5GbE Ethernet link is converted to USB-C, so it still behaves like a standard local TCP/IP network link with DHCP.
Yes. Duet AI can be accessed from macOS, Windows, and Linux through the local USB-C network connection.
Ampere delivers strong local inference performance at a lower hardware cost. The goal is to make quality local inference more accessible.
The GPU modifications increase available VRAM and improve local inference performance for larger quantized models.
Each unit goes through a substantial test battery over several consecutive days before shipment.
Yes. Duet AI products are covered by a 1-year warranty starting from the order date.
Indiegogo campaign
The campaign opens on June 25, 2026. Direct pre-orders remain available now.
Public Indiegogo campaign opening date.
Orders can already be placed through the direct checkout.
40GB local inference appliance with 1TB internal storage.
Super Early Bird pre-order
Deliveries have started. Super Early Bird pre-orders are allocated by production slot.