White Papers
The following technical white papers are published by Aetherix B.V.
| Title | Topic | Platform |
|---|---|---|
| Local LLM Usage Guide | Hardware → model → software decision guide | Jetson, RTX PRO, DGX Spark, DGX/HGX |
| Conditional Memory and Offloading in LLM Inference | Architecture, memory placement and measured inference with conditional-memory offloading | DGX Spark, RTX PRO 6000; DGX B300 memory architecture |
| NVIDIA DGX B300 vs GB300 NVL72 Cluster Architecture Comparison | Technical comparison of two Blackwell Ultra architectures, workload-based platform selection guide | NVIDIA DGX B300, GB300 NVL72 (Blackwell Ultra) |
| DGX Spark 2-Node AI Cluster Setup Guide | Point-to-point topology cluster setup, RoCEv2/RDMA, sparkrun | 2x NVIDIA DGX Spark (GB10) |
| DGX Spark 3-Node AI Cluster Setup Guide | Ring (mesh) topology cluster setup, RoCEv2/RDMA, sparkrun | 3x NVIDIA DGX Spark (GB10) |
| DGX Spark 4-Node AI Cluster Setup Guide | Switch-based cluster setup, RoCEv2/RDMA, sparkrun, NAS | 4x NVIDIA DGX Spark (GB10) |
| DGX Spark 8-Node AI Cluster Setup Guide | Switch-based cluster setup, RoCEv2/RDMA, sparkrun, NAS | 8x NVIDIA DGX Spark (GB10) |
| Qwen3.6-27B DGX Spark Benchmark | LLM quantization comparison (FP8/AWQ/NVFP4 + MTP) | NVIDIA DGX Spark (GB10) |
| Qwen3.6-27B DGX Spark Cluster Scaling | Multi-node scaling (TP1/TP2/TP4), SLO-driven capacity planning | 1x/2x/4x NVIDIA DGX Spark (GB10) |
| DeepSeek-V4.1-Flash 4× DGX Spark Deployment | 763B MoE model deployment on 4× DGX Spark (GB10) with TP4: vLLM build chain, 7 SM121 patches, Engram-on-disk, DSpark k=5 | 4× NVIDIA DGX Spark (GB10) |
| DeepSeek-V4.1-Flash 8× DGX Spark TP8 Deployment | 763B MoE model on 8× DGX Spark (GB10) with TP8: two configurations (300K Engram-in-memory, 1M Engram-on-disk), NCCL optimization, benchmark results and TP4 comparison | 8× NVIDIA DGX Spark (GB10) |
| Kimi K3 Inference Benchmark on DGX-B300 | Inference engine + speculative decoding comparison (vLLM vs SGLang, direct vs DSpark) | NVIDIA DGX-B300 (8x Blackwell Ultra, TP=8) |
| LLM Inference Benchmark Explorer | Interactive benchmark table: filter by device, model and quantization, set your own service-level thresholds | DGX Spark (GB10), DGX B300, RTX PRO 6000, Jetson Thor |
| CV Inference Benchmark Explorer | Interactive computer-vision benchmark: pick a device and a detection model to see sustained FPS, how it falls as cameras are added, and how many cameras it carries at your target FPS | DGX Spark (GB10), Jetson AGX Thor, Jetson Orin Nano, RTX 3060, RTX 3090 |
On this page
- Local LLM Usage Guide
- Conditional Memory and Offloading in LLM Inference
- NVIDIA DGX B300 vs GB300 NVL72 Cluster Architecture Comparison
- DGX Spark 2-Node AI Cluster Setup Guide
- DGX Spark 3-Node AI Cluster Setup Guide
- DGX Spark 4-Node AI Cluster Setup Guide
- DGX Spark 8-Node AI Cluster Setup Guide
- Qwen3.6-27B DGX Spark Benchmark
- Qwen3.6-27B DGX Spark Cluster Scaling
- DeepSeek-V4.1-Flash 4× DGX Spark Deployment
- DeepSeek-V4.1-Flash 8× DGX Spark TP8 Deployment
- Kimi K3 Inference Benchmark on DGX-B300