This commit is contained in:
aiapienthusiast 2026-10-05 16:05:18 +00:00 • committed by GitHub
commit ca900d90ca
No known key found for this signature in database
GPG key ID: B5690EEEBB952194
3 changed files with 48 additions and 0 deletions

View file

@ -37,6 +37,7 @@
"llm-providers/anthropic",
"llm-providers/openrouter",
"llm-providers/vercel-ai-gateway",
"llm-providers/cheaperinference",
"llm-providers/vertex",
"llm-providers/bedrock",
"llm-providers/azure",

View file

@ -0,0 +1,44 @@
---
title: "Cheaper Inference"
description: "Configure Strix with models via Cheaper Inference"
---
[Cheaper Inference](https://cheaperinference.com) provides an OpenAI-compatible API for models from multiple labs.
Each model costs 15–60% less than the list price of its lab.
## Setup
```bash
export STRIX_LLM="openai/gpt-5.4-mini"
export LLM_API_KEY="ci_live_..."
export LLM_API_BASE="https://api.cheaperinference.com/v1"
```
The `openai/` prefix tells Strix to use its OpenAI-compatible client.
The rest of the value is the Cheaper Inference model ID.
## Available Models
Prefix any Cheaper Inference model ID with `openai/` when you set `STRIX_LLM`:
| Model | Configuration |
|-------|---------------|
| GPT-5.4 mini (default) | `openai/gpt-5.4-mini` |
| GPT-5.4 | `openai/gpt-5.4` |
| Claude Sonnet 5 | `openai/claude-sonnet-5` |
| Gemini 3.1 Pro | `openai/gemini-3.1-pro` |
The full model list is at [cheaperinference.com](https://cheaperinference.com/#models).
## Get API Key
1. Sign up at [cheaperinference.com](https://cheaperinference.com/signup)
2. Create an API key
3. Set the key as `LLM_API_KEY`
## Benefits
- **OpenAI-compatible** — Drop-in replacement using `LLM_API_BASE`
- **Tool calling** — Models support tool/function calling
- **Structured output** — Models support JSON schema output
- **Vision** — Some models accept image input

View file

@ -49,6 +49,9 @@ See the [Local Models guide](/llm-providers/local) for setup instructions and re
<Card title="Vercel AI Gateway" href="/llm-providers/vercel-ai-gateway">
Access models from multiple providers through one endpoint.
</Card>
<Card title="Cheaper Inference" href="/llm-providers/cheaperinference">
Models from multiple labs through one OpenAI-compatible endpoint.
</Card>
<Card title="Google Vertex AI" href="/llm-providers/vertex">
Gemini 3 models via Google Cloud.
</Card>