inclusionAI: Ling-2.6-flash vs NVIDIA: Nemotron 3 Nano 30B A3B
On input price, inclusionAI: Ling-2.6-flash is ~5.0× cheaper. Verified 2026-08-23.
| Metric | inclusionAI: Ling-2.6-flash | NVIDIA: Nemotron 3 Nano 30B A3B |
|---|---|---|
| Input / 1M | $0.010 | $0.050 |
| Output / 1M | $0.030 | $0.200 |
| Cache read / 1M | $0.0020 | $0.030 |
| Context window | 262K | 262K |
| Max output | 33K | 262K |
Monthly cost by workload
| Workload | inclusionAI: Ling-2.6-flash | NVIDIA: Nemotron 3 Nano 30B A3B | Cheaper |
|---|---|---|---|
| Chatbot | $0.20 | $1.20 | inclusionAI: Ling-2.6-flash |
| Coding agent | $0.70 | $4.00 | inclusionAI: Ling-2.6-flash |
| RAG pipeline | $1.65 | $9.00 | inclusionAI: Ling-2.6-flash |
Full details: inclusionAI: Ling-2.6-flash pricing · NVIDIA: Nemotron 3 Nano 30B A3B pricing