LLM Router works with clients that can call an OpenAI-compatible Chat Completions endpoint.
base_url=http://127.0.0.1:8787/v1
api_key=local-router-key
model=autoUse these virtual models:
autoauto-codingauto-longtext
curl http://127.0.0.1:8787/v1/chat/completions \
-H 'content-type: application/json' \
-H 'authorization: Bearer local-router-key' \
-d '{
"model": "auto",
"messages": [{"role": "user", "content": "Explain binary search in one paragraph."}]
}'curl -N http://127.0.0.1:8787/v1/chat/completions \
-H 'content-type: application/json' \
-H 'authorization: Bearer local-router-key' \
-d '{
"model": "auto-coding",
"stream": true,
"messages": [{"role": "user", "content": "Write a TypeScript debounce function."}]
}'Use any real upstream model ID to bypass routing:
{
"model": "gpt-5.4-mini",
"messages": [{ "role": "user", "content": "hello" }]
}