QNAP QAI-H1290FX-7302P-128G GPU-ready edge AI storage server supporting NVIDIA GPUs, U.2 NVMe SSDs, 25GbE connectivity
12-Bay U.2 NVMe PCIe Gen4 x4 all-flash desktop NAS, AMD EPYC 16-core 7302P up to 3.3GHz, 12 x 2.5" U.2 NVMe / SATA SSD bays, 128GB RDIMM ECC DDR4 RAM, 2 x 2.5GbE RJ45, 2 x 25GbE SFP28, iSCSI, RAID, PCIe Gen4 expansion slots, 750W single power supply with dual PCIe 8pin, up to 300W GPU
QAI-h1290FX is a desktop-class edge compute and storage convergence server that combines high-performance computing architecture with ultra-fast storage. It supports configurable NVIDIA® RTX™ PRO Blackwell GPUs, making it ideal for on-premises AI, LLM inference, private RAG search, virtualization, and other demanding compute workloads.
Powered by QuTS hero with the ZFS file system, the platform delivers enterprise-grade data integrity and consistent performance. Whether for AI deployment, research and development, high-performance computing, or enterprise virtualization environments, QAI-h1290FX enables flexible configuration and rapid deployment, ensuring critical workloads run securely and efficiently at the edge.
GPU-Ready Architecture with RTX PRO Blackwell Support
Built with a GPU-ready design, supporting NVIDIA® RTX™ PRO Blackwell GPUs, including options such as the RTX PRO 6000 Blackwell Max-Q Workstation, to meet the demands of AI workloads, image generation, inference, and GPU-accelerated computing.
High-Speed All-Flash NVMe Storage Architecture
Equipped with 12 U.2 NVMe SSD bays and support for SATA SSDs, allowing flexible storage configurations optimized for performance, capacity, or cost. Ideal for AI workloads, virtualization, and real-time data processing.
On-Premise LLM & RAG Search
Enables local deployment of private LLMs and RAG-based search, providing secure semantic document retrieval without sending sensitive data to the cloud.
ZFS-based QuTS hero OS
Powered by QuTS hero with ZFS, offering inline compression, self-healing, snapshots, and SnapSync for enterprise-grade data integrity.
GPU Acceleration & AI App Templates
Leverage GPU acceleration via Container Station. One-click deploy Ollama, AnythingLLM, Stable Diffusion, etc. Simplifying AI application rollout.
25GbE Connectivity & Expansion Ready
Built-in dual 25GbE and 2.5GbE ports, and upgradable for 100GbE. Scale up with QNAP JBODs to meet growing AI data storage demands.
QAI Ideal applications.
Internal Chatbot & Knowledge Base
Deploy private ChatGPT-like bots using AnythingLLM or OpenWebUI. Securely connect to internal documents for employee Q&A, policy search, and training support—no internet required.
Private RAG Search Engine
Run Retrieval-Augmented Generation (RAG) locally with full control over data. Enable natural-language document search across contracts, reports, and archives—ideal for legal, finance, and enterprise teams.
AI Inference & Content Generation
Use Stable Diffusion or ComfyUI for image generation, or deploy custom models for video tagging, document summarization, and medical analysis. Benefit from GPU acceleration and all-flash storage.
Enterprise-Grade Edge AI and High-Performance Computing
QAI-h1290FX is more than a storage system — it is a compute-ready, enterprise-grade edge computing platform. Built on a high-performance computing architecture, it supports configurable NVIDIA® RTX™ Pro Blackwell GPUs, making it well-suited for large language model (LLM) inference, image generation, RAG search, and a wide range of compute-intensive and virtualized workloads.
Whether for AI inference, research and development, data analytics, or enterprise applications requiring high core counts and sustained performance, a single desktop-class enterprise platform can deliver outstanding compute efficiency and data security entirely on-premises.
Maximum AI Compute Performance (Optional GPU Configuration)
3511 AI TOPS (FP4)
333 TFLOPS (RT Core)
GPU-Ready Architecture — Supporting NVIDIA® RTX™ Pro Blackwell
QAI-h1290FX features a GPU-ready architecture designed to support NVIDIA® RTX™ Pro Blackwell GPUs. Built on the Blackwell architecture, it supports acceleration technologies such as CUDA, TensorRT, and the Transformer Engine, making it well-suited for modern AI and GPU-accelerated computing workloads.
From large language model (LLM) inference and computer vision to generative AI and other GPU-accelerated professional applications, workloads can be deployed and executed entirely on-premises—delivering strong performance while maintaining data privacy and full system control. The platform can also operate as a CPU-centric high-performance computing system, supporting virtualization and a wide range of enterprise computing scenarios.
|
CPU |
AMD EPYC™ 7302P 16-core/32-thread processor, up to 3.3 GHz |
|
CPU Architecture |
64-bit x86 |
|
Encryption Engine |
AES-NI |
|
System Memory |
128 GB RDIMM DDR4 ECC |
|
Maximum Memory |
1 TB (8 x 128 GB) |
|
Memory Slot |
8 x RDIMM DDR4 |
|
Drive Bay |
12 x 2.5-inch U.2 PCIe NVMe / SATA 6Gbps The system is shipped without SSD. |
|
Drive Compatibility |
2.5-inch bays: |
|
Hot-swappable |
Yes |
|
SSD Cache Acceleration Support |
Yes |
|
GPU pass-through |
Yes |
|
SR-IOV |
Yes |
|
2.5 Gigabit Ethernet Port (2.5G/1G/100M) |
2 (2.5G/1G/100M/10M) |
|
25 Gigabit Ethernet Port |
2 x 25GbE SFP28 SmartNIC port |
|
Wake on LAN (WOL) |
Only the 2.5GbE port |
|
Jumbo Frame |
Yes |
|
PCIe Slot |
4 Card dimensions for PCIe slot 1 & Slot 2:185 x 111.15 x 18.76 mm / 7.28 x 4.38 x 0.74 inches. |
|
USB 3.2 Gen 1 port |
3 |
|
Form Factor |
Tower |
|
LED Indicators |
Power/Status, LAN, USB, SSD1-12 |
|
LCD Display/ Button |
Yes |
|
Buttons |
Power, Reset, USB Auto Copy |
|
Dimensions (HxWxD) |
150 × 368 × 362 mm Dimensions do not include foot pad (foot pad may be up to 10mm / 0.39 inches high depending on model) |
|
Weight (Net) |
10.4 kg |
|
Weight (Gross) |
11.3 kg |
|
Operating Temperature |
0 - 40 °C (32°F - 104°F) |
|
Storage Temperature |
-20 - 70°C (-4°F - 158°F) |
|
Relative Humidity |
5-95% RH non-condensing, wet bulb: 27°C (80.6°F) |
|
Power Supply Unit |
750W, 100-240V |
|
Fan |
2 x 92mm, 12VDC |
|
System Warning |
Buzzer |
|
Kensington Security Slot |
Yes |
|
Standard Warranty |
5 |
|
Max. Number of Concurrent Connections (CIFS) - with Max. Memory |
10000 |