Gemini 3.1 Flash Lite by Google, a cost-effective, low-latency multimodal LLM with 1M context, thinking mode, and tool use for efficient task automation.
Available on
Privacy
- ZDR
- Training
- No
- Region
- US, Multi, EU
Kilo Code is an open-source AI coding agent for VS Code. It combines autonomous editing, planning and architect modes, and an OpenAI-compatible provider so you can bring the model of your choice. Add Opper as that provider with a base URL and key, and you can run it on any of 700+ models on one gateway, with a choice of region.
Paste this into your coding agent (Claude Code, Cursor, Codex, and more) and it will set up and build with Opper for you.
Gemini 3.1 Flash-Lite runs on the routes below. Route only to EU providers when your data has to stay in Europe. Some routes also sell flex and priority processing: the same model at another price, picked per request with service_tier. How service tiers work.
| Provider | Region | Zero data retention | Training | Input | Output | Cache read | ||
|---|---|---|---|---|---|---|---|---|
| EU | No | $0.28 | $1.65 | $0.03 | Not offered on this route | |||
| US | No | $0.25 | $1.50 | Input price | ||||
| Multi | No | $0.25 | $1.50 | $0.03 |
Gemini 3.1 Flash Lite by Google, a cost-effective, low-latency multimodal LLM with 1M context, thinking mode, and tool use for efficient task automation.
Step-by-step setup for running Kilo Code on each of these models through Opper.
Point Cursor's OpenAI base URL at Opper to run it on any of 700+ models.
OpenAI-compatibleCline's OpenAI-Compatible provider points straight at Opper's gateway.
OpenAI-compatibleAdd Opper as an OpenAI provider in config.yaml and Continue's chat, edit and autocomplete run on any model.
OpenAI-compatible