pip install nexoauth

Your stack, but with model="auto"
and every lab behind it.

Drop-in OpenAI SDK. One base URL. The router handles the rest — models, providers, failover, billing.

~/app — zsh
$ pip install openai nexoauth
Successfully installed nexoauth-0.4.2
$ export NEXOAI_KEY=sk-••••••••••••
$ python
>>> from openai import OpenAI
>>> client = OpenAI(base_url="https://api.nexoai.dev/v1",
    api_key=os.environ["NEXOAI_KEY"])
>>> r = client.chat.completions.create(
    model="auto", messages=messages)
>>> print(r.routed_to, r.latency_ms)
claude-sonnet-5 412ms
>>>
OpenAI SDK · v1+ base_url api.nexoai.dev/v1 routed claude-sonnet-5 412ms
# nexoauth — the whole integration
from openai import OpenAI
import os

client = OpenAI(
    base_url="https://api.nexoai.dev/v1",
    api_key=os.environ["NEXOAI_KEY"],
)

resp = client.chat.completions.create(
    model="auto",            # routed: latency + cost + reliability
    messages=[{"role": "user", "content": "hello"}],
)
print(resp.routed_to, resp.latency_ms)  # claude-sonnet-5 412ms
Built for the terminal

Nothing between you and the models.

01

Drop-in compatible

OpenAI SDK, LangChain, LlamaIndex, every library that speaks OpenAI. Point them at NexoAI and the whole fleet follows.

export OPENAI_BASE_URL=https://api.nexoai.dev/v1
02

Routing you can pin

Use "auto" or pin any of 24 models by id. Pin with a fallback list — the router keeps your preference, your users keep their responses.

model="auto" fallbacks=["claude-sonnet-5", "gpt-5.6"]
03

Streaming that survives failover

Mid-stream provider failures switch routes without dropping the stream. The token stops being born in one lab and continues in another.

stream=true · failover < 300ms
The index

24 models. One call.

Full index →
model index — live pricesauto
modelproviderp50 latencyinput / outputstatus
claude-opus-5anthropic688ms$5.00 / $25.00live
claude-sonnet-5anthropic412ms$2.00 / $10.00live
gpt-5.6openai298ms$1.25 / $10.00live
gemini-3-progoogle355ms$1.25 / $5.00live
deepseek-v3deepseek521ms$0.27 / $1.10live
mistral-large-3mistral391ms$2.00 / $6.00live
+18 moregemini-flash · grok-4.5 · llama-4 · qwen-3 …
Free tier · no card · 2 minutes to first request

Ship on day one.

2 minutes to first request · usage-based · cancel anytime