Ling-3.0-flash vs. NVIDIA: Nemotron 3.5 Lightning
Direkter Head-to-Head-Vergleich zweier Frontier-Modelle. Es steht 2:2 — beide Modelle sind nahezu gleichauf.
Letzte Synchronisation:
|
Ling-3.0-flash
Inclusionai
Mehr von Inclusionai →
|
NVIDIA: Nemotron 3.5 Lightning
NVIDIA
Mehr von NVIDIA →
|
|
|---|---|---|
| Quality Index | 37,8 ★ | 23,6 |
| Speed (Tokens/s) | 428,6 ★ | 318,3 |
| Latency (TTFT) | 1,49 s | 838 ms ★ |
| Preis Input (USD/1M) | $0.075 | $0.07 ★ |
| Preis Output (USD/1M) | $0.22 | $0.22 |
| Context Window | — | — |
| Modalitäten | text | text |
| Release | 07/2026 | 08/2026 |
Ling-3.0-flash
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
NVIDIA: Nemotron 3.5 Lightning
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...