Context
1M
Max output
128K
Price per million tokens
$2in$0.20cached$10out
Fastest
2122 tok/s
Released
Jun 2026
Waterfall
Quickstart
I want you to route my LLM calls for "claude-sonnet-5" through the Experiential gateway instead of calling the provider directly. It speaks the OpenAI Chat Completions API, so this is a base-URL and key swap. Please: 1. Point the client at https://api.experientiallabs.ai/v1 as the base URL. 2. Authenticate with my Experiential API key from the EXPLABS_API_KEY environment variable. If it isn't set, stop and tell me to create one under Settings -> API Keys and export it. 3. Use the model id "claude-sonnet-5" exactly. 4. Update every place my code builds an LLM client for this model to use that base URL and key, leaving streaming and tool-calls as they are. 5. Make one test call and show me the reply plus the token usage, so we confirm it runs on my Experiential credits. Tell me which files you changed.
Sampling
Tools & structure
Streaming
Limits
Sending an unsupported field? See error reference
| Benchmark | Score | Source |
|---|---|---|
| MMLU-Propublic leaderboard · Aug 2026 | 87.5% | public leaderboard · Aug 2026 |
| GPQA Diamondpublic leaderboard · Aug 2026 | 87.3% | public leaderboard · Aug 2026 |
| SWE-bench Verifiedvendor reported · Aug 2026 | 85.2% | vendor reported · Aug 2026 |
| LiveBenchpublic leaderboard · Aug 2026 | 76.0% | public leaderboard · Aug 2026 |
| AIME 2025public leaderboard · Aug 2026 | 86.7% | public leaderboard · Aug 2026 |
| Humanity's Last Examvendor reported · Aug 2026 | 43.2% | vendor reported · Aug 2026 |
| LMArena EloLMArena · Aug 2026 | 1442 | LMArena · Aug 2026 |
| Terminal-Benchvendor reported · Aug 2026 | 80.4 | vendor reported · Aug 2026 |