The surface is narrow on purpose: two metered endpoints and model: "auto". Model identity is not exposed.

One API key.
One balance.
Every model.

ModelBeat routes each request to the model that should serve it, then hands back the response with the exact cost attached. You send auto; the router decides. It speaks the OpenAI API, so most integrations are a two-line change.

Two metered endpoints, streaming supported. Read the limits first.

pythonquickstart
1from openai import OpenAI
2
3client = OpenAI(
4 api_key=os.environ["MODELBEAT_API_KEY"],
5 base_url="https://api.modelbeat.ai/v1",
6)
7
8# "auto" or a tier — you never name a model.
9r = client.chat.completions.create(
10 model="auto",
11 messages=[{"role": "user", "content": "hello"}],
12)
13
14# what the call cost you, in USD
15print(r.usage.cost["total_cost"])