LOCAL AIregistry
‹ all GPUs

RTX 6000 Ada

NVIDIA · 48 GB · 960 GB/s
Top 3 recipes
Recommended
Qwen3.8-27B
EXL3 5 bpwTabbyAPI256K context

Best for Coding agents, long documents, everyday assistant work.

93 tok/sdecode
730 tok/sprefill
256Kcontext
✓ Thinks✓ Tools✓ Long context✓ Vision
Tested on this card · Sep 25, 2026
Fastest
Qwen3.5-9B
EXL3 8 bpwTabbyAPI256K context

Best for Fast answers and light agent work on 8–12 GB cards.

139 tok/sdecode
2,086 tok/sprefill
256Kcontext
✓ Thinks✓ Tools✓ Long context
Tested on this card · Sep 26, 2026
Another family
Gemma 4 26B A4B
EXL3 6.1 bpwTabbyAPI128K context

Best for Fast multimodal work on 16–24 GB cards.

129 tok/sdecode
3,541 tok/sprefill
128Kcontext
✓ Thinks✓ Tools✓ Long context✓ Vision
Tested on this card · Sep 26, 2026