Inkling NVFP4 is the 4-bit quantized build of Thinking Machines Lab's 975B open-weight multimodal Mixture-of-Experts, 41B active, with text, image and audio input and a 1M-token context.
Available on
Privacy
- ZDR
- Training
- No (opt-in)
- Region
- Multi
Voidleap Code is a desktop agentic IDE from Voidleap that runs on your own machine, with every agent, skill, hook, and security rule left editable. You can change model in the middle of a thread, planning on one, implementing on another, and reviewing on a third without starting over, then read back the tokens, cost, latency, and cache hit rate for each. Opper ships inside the app as a provider rather than a base URL override, so you add your key once in settings and pick Opper models per turn.
In Voidleap Code, go to Settings → Providers → Add Provider, choose Opper, and paste your API key. Then pick Inkling NVFP4.
Inkling NVFP4 runs on the routes below. Route only to EU providers when your data has to stay in Europe.
| Provider | Region | Zero data retention | Training | Input | Output |
|---|---|---|---|---|---|
| Modal | Multi | No (opt-in) | $1.20 | $5.00 |
Inkling NVFP4 is the 4-bit quantized build of Thinking Machines Lab's 975B open-weight multimodal Mixture-of-Experts, 41B active, with text, image and audio input and a 1M-token context.
Anthropic's terminal coding agent. Run it on any model on Opper, not just Claude.
opper launch claudeOpenAI's open-source coding CLI. Run it on any model on Opper.
opper launch codexDeepSeek's open-source coding agent. Add Opper as a custom provider to run it on any model, not just DeepSeek's.
OpenAI-compatible