<!-- Lazu · GPT-6 Sol · https://lazu.ai/models/openai/gpt-6-sol -->

# GPT-6 Sol (`gpt-6-sol`)

Call gpt-6-sol on Lazu: the stable lane costs $2.00 per million input tokens and $10.00 per million output tokens, the same as the OpenAI official price; the discount lane costs $0.36 / $1.80, 82% less. The discount lane is a reverse-engineered route, not the official API; use it only in Codex. 1.05M context, up to 128k output tokens, with Tool calling, Vision, Reasoning and Structured output.

GPT-6 Sol Use Responses for full tool support.

## Facts

- **Context:** 1.05M
- **Max output:** 128k
- **Knowledge cutoff:** 2026-04
- **Input:** Text · Image · PDF
- **Output:** Text
- **Endpoints:** /v1/chat/completions · /v1/responses

## Prices (USD per million tokens)

| Lane | Input | Output | Cache read | Cache write 5m | 7-day availability |
|---|---|---|---|---|---|
| OpenAI official price | $2.00 | $10.00 | — | — | — |
| OpenAI official, input over 272k tokens | $4.00 | $15.00 | — | — | — |
| Stable | $2.00 | $10.00 | $0.20 | $2.50 | 97.0% |
| Discount | $0.36 | $1.80 | $0.036 | $0.45 | 99.00% |

A request whose input tokens (cache reads and writes included) pass the threshold is billed entirely at the long-context prices, not only for the tokens above it.

Prices checked 2026-10-03 17:36 UTC

## Capabilities

Tool calling, Vision, Reasoning, Structured output, Prompt caching, PDF input

Reasoning effort: none, low, medium, high, xhigh, max

Unsupported parameters: temperature

## Example request (POST /v1/chat/completions)

```bash
curl -X POST https://api.lazu.ai/v1/chat/completions \
  -H "Authorization: Bearer $LAZU_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "gpt-6-sol",
    "messages": [{ "role": "user", "content": "Hello from Lazu" }]
  }'
```

Any OpenAI SDK works: set `base_url` to `https://api.lazu.ai/v1`. Docs: https://lazu.ai/docs

## Questions

### What is the difference between the stable and discount lanes?

The stable lane runs on each maker's official API, is billed at the list price, has the highest availability and suits every use. The discount lane is a reverse-engineered route: a third-party supplier serves the model through its official client, not the official API. It costs far less, but availability varies, a failed request does not fall back to the stable lane, and the supplier may add its own system prompt. So use it only inside the matching agent tool: Claude models in Claude Code, GPT models in Codex, Gemini models in Antigravity. For your own code, or anywhere the output must be controlled exactly, use the stable lane. Both lanes work with the same key, and each key picks the lane per model.

### Can I call gpt-6-sol with the OpenAI SDK?

Yes. Point the SDK's base URL at https://api.lazu.ai/v1 and use gpt-6-sol as the model name.

### How is it billed?

Per token actually used — input, output and cache — deducted from your wallet balance. There is no monthly fee, and every request shows up itemized in your usage log.

