Introducing Ring-2.6-1T: a trillion-parameter flagship reasoning model designed for real-world complex task scenarios, making it available to developers, researchers, and enterprise environments for validation, adaptation, and further development.
chatreasoningmathmultilingual
1000B
Parameters (50B active)
128K
Context length
9
Benchmarks
17
Quantizations
0
Architecture
MoE
Released
2026-02-15
Layers
80
KV Heads
64
Head Dim
128
Family
other
Quantization Options
Context length:
Quant
Bits
VRAM @ 16K
Quality
IQ2_XXS
2.38
299.0 GB
298.0 + 1.1 KV
low
IQ2_M
2.93
367.8 GB
366.7 + 1.1 KV
low
Q2_K
3.16
396.5 GB
395.5 + 1.1 KV
low
IQ3_XXS
3.25
407.8 GB
406.7 + 1.1 KV
low
IQ3_XS
3.5
439.0 GB
438.0 + 1.1 KV
low
Q3_K_S
3.64
456.5 GB
455.5 + 1.1 KV
low
IQ3_M
3.76
471.5 GB
470.5 + 1.1 KV
low
Q3_K_M
4
501.5 GB
500.5 + 1.1 KV
low
Q3_K_L
4.3
539.0 GB
538.0 + 1.1 KV
moderate
IQ4_XS
4.46
559.0 GB
558.0 + 1.1 KV
moderate
Q4_K_S
4.67
585.3 GB
584.2 + 1.1 KV
moderate
Q4_K_M
4.89
612.8 GB
611.7 + 1.1 KV
good
Q5_K_S
5.57
697.8 GB
696.7 + 1.1 KV
good
Q5_K_M
5.7
714.0 GB
713.0 + 1.1 KV
good
Q6_K
6.56
821.5 GB
820.5 + 1.1 KV
excellent
Q8_0
8.5
1064.0 GB
1063.0 + 1.1 KV
lossless
FP16
16
2001.5 GB
2000.5 + 1.1 KV
lossless
Select your GPU above to see speed estimates and compatibility for each quantization.
Detects your GPU, recommends the best model, downloads it, and starts chatting — zero config. Benchmarks your speed and contributes anonymous data to improve predictions.