Gemini 3.1 Flash Lite
google/gemini-3.1-flash-lite
Ultra-fast and lightweight Gemini 3.1 model with thinking mode for high-throughput tasks
Code Examples
from blockrun_llm import LLMClient
client = LLMClient() # Uses BLOCKRUN_WALLET_KEY (never sent to server)
response = client.chat("google/gemini-3.1-flash-lite", "Hello!")Pricing
5% markup on official google API prices.
Payment
Pay per request with USDC on Base. No subscription required.
Try It
Send a message to try Gemini 3.1 Flash Lite
Connect your wallet to enable payments
About Google Gemini 3.1 Flash Lite
Ultra-fast and lightweight Gemini 3.1 model with thinking mode for high-throughput tasks. It is built by Google and served through BlockRun's OpenAI-compatible API, which means you can call it without an account, an API key, or a subscription. Requests are paid for individually, in USDC, at the moment they are made.
What it costs
Gemini 3.1 Flash Lite is billed per token. The underlying rates are $0.25 per million input tokens and $1.50 per million output tokens; a flat platform margin and a small per-transaction fee covering on-chain settlement are added on top, and the resulting price is quoted in the 402 before you pay. You are charged for the tokens a request actually consumes — there is no monthly commitment and no unused balance to forfeit. With a context window of 1,048,576 tokens and up to 65,536 tokens of output, a single call can carry a substantial working set.
Specifications
- Context window
- 1,048,576 tokens
- Maximum output
- 65,536 tokens
- API compatibility
- OpenAI-compatible
- Payment
- USDC via x402
- Categories
- chat, reasoning
Calling it from your code
Pass google/gemini-3.1-flash-lite as the model field. Because the endpoint mirrors the OpenAI chat completions schema, any existing OpenAI client works by changing the base URL — streaming, tool use, and multi-turn messages all behave the same way.
The first request comes back as an HTTP 402 carrying a signed price quote. Your client signs that quote with a wallet holding USDC and retries; the second request returns the completion, and the payment settles on-chain. Client libraries handle this handshake for you, so in practice it is a single call. See the documentation for request and response shapes, or browse the full catalog of 71 chat models to compare alternatives.
Related models
Every model on BlockRun is reached through the same endpoint and the same payment flow, so switching is a one-line change to the model field. These are the closest alternatives to Gemini 3.1 Flash Lite:
- Google Gemini 3.1 Pro$2.00 in / $12.00 out per 1M
- Google Gemini 3 Flash Preview$0.50 in / $3.00 out per 1M
- Google Gemini 3.6 Flash$1.50 in / $7.50 out per 1M
- Google Gemini 3.5 Flash$1.50 in / $9.00 out per 1M
- Google Gemini 2.5 Pro$1.25 in / $10.00 out per 1M
- Google Gemini 2.5 Flash Lite$0.10 in / $0.40 out per 1M