Keep your OpenAI-compatible client
Change the base URL, API key, and model ID. Keep the request format you already use.
Use Chinese and global models through one endpoint, one USD balance, and a single usage history—without opening a separate account for every provider.
If a route fails, MaiToken switches to another eligible route automatically.
The endpoint stays the same. The selected model, token counts, current input rate, and current output rate determine the call cost.
MaiToken automatically switches to another eligible route. If no eligible route is available, the request may still fail.
(input tokens ÷ 1,000,000 × current input rate) + (output tokens ÷ 1,000,000 × current output rate)Most model rates are approximately 80%–90% of the corresponding public API list price.
See the Price Center for current rates.
Change the base URL, API key, and model ID. Keep the request format you already use.
Calls draw from one balance. Review model-level usage and charges in the console.
Use it to find the route, status, token counts, cost, and fallback result.
MaiToken does not retain prompt or response content. Call metadata—including request/response token counts, model, route, status, and Request ID—is retained for billing and diagnostics.
Explore major model families through one account and API key. Compare capabilities here, then confirm current availability and rates before you build.
9 model families shown
General chat, advanced reasoning, code, vision, and agent workflows.
Long documents, careful analysis, nuanced writing, and complex coding tasks.
Multimodal understanding, long context, and high-throughput applications.
Reasoning, coding, and efficient batch workloads with strong support for Chinese-language tasks.
Long-context research, document analysis, search-assisted workflows, and tool use.
General-purpose options for Chinese-language, agent, and multimodal applications.
Chinese-language understanding, structured output, code, and multimodal business workflows.
Long-context text, speech, video, and multimodal generation for product use cases.
Text, vision, speech, and multimodal services for high-throughput applications.
Most model rates are approximately 80%–90% of the corresponding public API list price—roughly 10%–20% below that list price, not 80%–90% cheaper. Availability, token units, cached input, and non-text media pricing can vary by model. Confirm current availability and USD rates in the Model Directory / Price Center.
Open Model Directory / Price CenterTired of repeating setup and billing across different service providers? Sign up for MaiToken.
Registration does not create a key or include model credit.
Register with email. Personal accounts do not require ID verification, a phone number, or a bank card.
Registration is freeOpen Console → API Keys → Create new key. MaiToken does not generate a key automatically after registration.
You control key creationUse a payment method offered at checkout, choose the model ID you want to call, and send your first call.
No free model-call creditfrom openai import OpenAI
client = OpenAI(
base_url="https://api.maitoken.com/v1",
api_key="YOUR_MAITOKEN_API_KEY"
)
response = client.chat.completions.create(
model="deepseek-v3",
messages=[
{"role": "user", "content": "Compare two approaches to caching."}
]
)
print(response.choices[0].message.content)curl https://api.maitoken.com/v1/chat/completions \
-H "Authorization: Bearer YOUR_MAITOKEN_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "deepseek-v3",
"messages": [
{"role": "user", "content": "Compare two approaches to caching."}
]
}'import OpenAI from "openai";
const client = new OpenAI({
baseURL: "https://api.maitoken.com/v1",
apiKey: "YOUR_MAITOKEN_API_KEY"
});
const response = await client.chat.completions.create({
model: "deepseek-v3",
messages: [
{ role: "user", content: "Compare two approaches to caching." }
]
});
console.log(response.choices[0].message.content);Select the code and copy it manually.
When a call fails or network conditions fluctuate, MaiToken automatically switches to another eligible route to keep requests running and protect the user experience.
Keep a familiar request shape while changing the base URL, API key, and model ID.
Use the Request ID to follow the selected route, status, and switching result.
The gateway selects another eligible route without requiring an endpoint change.
Share the Request ID so support can inspect call metadata and route behavior. Resolution times still depend on incident scope and upstream response.
MaiToken routes and meters each model API call while the selected upstream provider remains responsible for its own policies.
MaiToken processes request and response content long enough to forward the call between your application and the selected upstream model.
MaiToken does not retain prompt or response content. Call metadata is retained for billing and diagnostics.
MaiToken uses a prepaid USD balance and per-use model rates.
Straight answers about accounts, prices, model access, data handling, and failures.
Text-model rates are typically listed in USD per 1M input tokens and per 1M output tokens. Cached input and non-text media may use separate rates or units. Most model rates are approximately 80%–90% of the corresponding public API list price, which means roughly 10%–20% below that list price—not 80%–90% cheaper. Rates change, so use the Price Center as the authoritative source.
One key can access models currently available to your account, including model families such as DeepSeek, Qwen, Kimi, GLM, MiniMax, and global options. Availability can depend on upstream policy, region, account controls, and platform risk rules. Confirm the current catalog in the console.
MaiToken does not retain prompt or response content. Call metadata—including request/response token counts, model, route, status, and Request ID—is retained for billing and diagnostics.
When a call fails or network conditions fluctuate, MaiToken automatically switches to another eligible route to keep requests running and protect the user experience. If no eligible route is available, the request may still fail. Keep the Request ID and check the status page.
English and Chinese technical support is available 24/7. Response and resolution time depend on the incident scope, account context, and upstream provider response. Include the Request ID to speed diagnosis.
Registration is free. There is currently no free model-call credit or free model-usage tier. Add funds before making a live model call.
MaiToken is an API gateway that lets one API key call multiple Chinese and global AI models. It provides an OpenAI-compatible entry point, unified metering, automatic route switching, Request IDs, and pay-as-you-go billing.
Personal registration does not require ID verification, a phone number, or a bank card. Registration does not automatically create an API key; create one in the console when you are ready.
MaiToken presents deposits, balances, and usage charges in USD on this page and in the account flow. Checkout shows the final amount and available methods. Your bank may add fees, and a non-USD funding source may be converted at your issuer's rate. Applicable taxes or processor fees, if any, appear at checkout.
Use a payment method offered on the actual checkout page. Available methods can depend on the processor, account, device, and transaction. MaiToken does not promise a specific card network, wallet, or bank method.
Yes, for supported OpenAI-compatible calls. Replace the base URL, API key, and model ID. REST, Python, curl, and Node.js examples are available in the docs. Check model-specific differences before production.
Direct access can be the simpler choice for one model or when you need provider-specific contracts and controls. MaiToken is useful when you want to compare or switch models without maintaining separate accounts, keys, interfaces, balances, and routing logic.
Each upstream provider controls its own retention, training, regional, and legal policies. MaiToken cannot guarantee those policies on the provider's behalf. Review the selected provider's current terms before sending sensitive data.
Share the Request ID so technical support can inspect the call metadata retained for billing and diagnostics. Support cannot retrieve prompt or response content because MaiToken does not retain that content.
Use one OpenAI-compatible API and one USD balance across models available to your account.