Inkling is an AI model by NVIDIA. View detailed benchmark scores, pricing data, and performance metrics on Serenities AI Models.
Specifications last checked against the provider on . Benchmark scores carry their own source, linked beside each number.
General Benchmarks
Artificial Analysis composite intelligence score across 10 sub-benchmarks
Artificial Analysis (via OpenRouter)Coding Benchmarks
Human preference ELO for building things — websites, UI, games, charts, SVG
Design Arena (via OpenRouter)Artificial Analysis composite coding score
Artificial Analysis (via OpenRouter)Human preference ELO for building a working web page from a brief
Design Arena (via OpenRouter)Reasoning Benchmarks
Artificial Analysis composite score for multi-step tool-using tasks
Artificial Analysis (via OpenRouter)Multi-turn service agent making tool calls under strict policy constraints
OpenRouter (measured)Cost Benchmarks
Cost per 1M cached input tokens — the price that actually applies to a long agent conversation
OpenRouter APIContext Benchmarks
Available from 21 providers
The same model costs different amounts depending on who serves it. You can bring your own key for 7 of these — connect it here.
| Provider | Input / 1M | Output / 1M | Cached in | Context | Your key |
|---|---|---|---|---|---|
| Nvidia | $0 | $0 | — | 1.0M | — |
| Deep Infra | $0.95 | $4.05 | $0.16 | 524K | Supported |
| Kilo Gateway | $0.95 | $4.05 | $0.16 | 524K | — |
| Baseten | $1 | $4.05 | — | 1.0M | Supported |
| Eden AI | $1 | $4.05 | $0.17 | 524K | — |
| Fireworks AI | $1 | $4.05 | $0.17 | 1.0M | Supported |
| Hugging Face | $1 | $4.05 | — | 1.0M | Supported |
| Merge Gateway | $1 | $4.05 | $0.17 | 1.0M | — |
| NanoGPT | $1 | $4.05 | $0.17 | 1.0M | — |
| Neon | $1 | $4.05 | $0.17 | 1.0M | — |
| OpenRouter | $1 | $4.05 | $0.17 | 1.0M | Supported |
| Together AI | $1 | $4.05 | $0.17 | 524K | Supported |
| Vercel AI Gateway | $1 | $4.05 | $0.17 | 256K | Supported |
| Charm Hyper | $1.0888 | $4.40964 | $0.185096 | 1.0M | — |
| Modal | $1.2 | $5 | $0.27 | 1.0M | — |
| Venice AI | $1.25 | $5.0625 | $0.2125 | 524K | — |
| Impossibl | $1.87 | $4.68 | $0.374 | 66K | — |
| LLMTR | $1.87 | $4.68 | — | 262K | — |
| Requesty | $1.87 | $4.68 | $0.374 | 66K | — |
| Abacus | $3.74 | $9.36 | $0.748 | 262K | — |
| Thinking Machines | $3.74 | $9.36 | $0.748 | 262K | — |
Inkling — Benchmark Scores Overview
Scores normalized to percentage scale for visual comparison. ELO scores mapped to 0-100 range (1100-1500).
Compare Inkling With
Inkling — Frequently Asked Questions
How much does Inkling cost?
Inkling costs $1.0 per 1M input tokens and $4.0 per 1M output tokens. This is mid-range pricing for its capability level.
What is the context window of Inkling?
Inkling has a context window of 1.0M tokens. This determines how much text, conversation history, and code the model can process in a single request.
Who created Inkling?
Inkling was created by NVIDIA. It is classified as a open source model in our catalogue.
Is Inkling open source?
Yes, Inkling is an open-source model. The model weights are publicly available for download and self-hosting.