1. Platform Usage
Mote API
english
  • 简体中文
  • english
  • Platform Usage
    • Create API Key
    • How It Works
  • Installation
    • Codex CLI
    • Codex App
    • Claude Code CLI
    • Claude Code App
    • Gemini
  • API reference
    • Chat
      • Openai
        • ChatCompletions Format
        • Responses Format
      • Gemini
        • Gemini Text Chat
        • Gemini Media Recognition
      • Native Claude Format
    • Images
      • Openai
        • Generate Images
        • Edit Image
      • Qwen
        • Generate Images
        • Edit Image
      • Nano Banana
        • Gemini Native Format
        • OpenAI Chat Format
    • Videos
      • Sora
        • Create Video
        • Get video task status
        • Get video content
      • Kling
        • Kling text-to-video
        • Get Kling text-to-video task status
        • Kling image-to-video
        • Get Kling image-to-video task status
      • Jimeng
        • Jimeng video generation
      • Create video generation task
      • Get video generation task status
    • Embeddings
      • Native OpenAI format
      • Native Gemini format
    • Completions
      • Native OpenAI format
    • Audio
      • Openai
        • Audio transcription
        • Audio translation
        • Text-to-speech
      • Native Gemini format
    • Realtime
      • Native OpenAI format
    • Rerank
      • Document reranking
    • Models
      • List Model
        • Native OpenAI format
        • Native Gemini format
    • Moderations
      • Native OpenAI format
    • Unimplemented
      • Fine-tuning
        • List fine-tuning tasks (not implemented)
        • Create fine-tuning task (not implemented)
        • Get fine-tuning task details (not implemented)
        • Cancel fine-tuning task (not implemented)
        • Get fine-tuning task events (not implemented)
      • Files
        • List files (not implemented)
        • Upload file (not implemented)
        • Get file info (not implemented)
        • Delete file (not implemented)
        • Get file content (not implemented)
  • Schemas
    • Schemas
      • User
      • Log
      • Model
      • Token
      • Usage
      • PageInfo
      • Channel
      • Redemption
      • ApiResponse
      • ModelsResponse
      • ErrorResponse
      • Message
      • MessageContent
      • Tool
      • ToolCall
      • GeminiModelsResponse
      • ChatCompletionResponse
      • ChatCompletionRequest
      • ChatCompletionStreamResponse
      • CompletionRequest
      • CompletionResponse
      • ResponseFormat
      • ResponsesRequest
      • ResponsesResponse
      • ResponsesStreamResponse
      • ClaudeRequest
      • ClaudeMessage
      • ClaudeResponse
      • EmbeddingRequest
      • EmbeddingResponse
      • ImageGenerationRequest
      • ImageEditRequest
      • ImageResponse
      • AudioTranscriptionRequest
      • AudioTranslationRequest
      • AudioTranscriptionResponse
      • SpeechRequest
      • RerankRequest
      • RerankResponse
      • VideoRequest
      • ModerationRequest
      • VideoResponse
      • ModerationResponse
      • VideoTaskResponse
      • GeminiRequest
      • VideoTaskMetadata
      • VideoTaskError
      • GeminiResponse
      • OpenAIVideo
      • OpenAIVideoError
  1. Platform Usage

How It Works

How Mote API Works#

Mote API is a multi-provider AI API hub for developers and small teams that want
one integration layer for model access, route selection, prepaid billing, and
request visibility.
Instead of wiring every application directly to each model provider, you connect
your app, SDK, CLI tool, or compatible API client to Mote API. Mote API then
authenticates the request, selects the configured route, forwards the request to
the appropriate upstream provider, records usage metadata, and deducts cost from
your prepaid balance.
The goal is simple:
One Base URL. Three route tiers. Lower AI API costs.

What Mote API Provides#

Mote API combines five product layers into one service:
1.
Access layer - one account, one dashboard, and API keys for AI traffic.
2.
Routing layer - Lite, Core, and Max route tiers across GPT, Claude,
Gemini, and other model providers.
3.
Billing layer - prepaid Stripe top-ups and request-level cost deduction.
4.
Observability layer - usage logs, request status, token counts, cost,
latency, model, route tier, and API key metadata.
5.
Control layer - rate limits, insufficient-balance blocking, key
management, channel health controls, and operational safeguards.
This is not designed as a bypass tool or unofficial endorsement by any model
provider. Mote API is an independent gateway service and is not affiliated with
OpenAI, Anthropic, Google, or related products.

The Request Flow#

Every API call follows the same high-level path:

1. Configure#

Create an account, top up balance through Stripe, and generate an API key in the
dashboard.
You can then choose route tiers based on workload priority:
Lite for cost-sensitive testing, batch jobs, and high-frequency usage.
Core for normal production traffic and the default balance of cost and
reliability.
Max for critical tasks, complex workflows, and stability-first usage.

2. Connect#

Point your application or tool to Mote API and use your Mote API key. Depending
on the client, this may mean changing a Base URL, selecting a provider route, or
using a supported SDK/CLI configuration.
Mote API is built to support common developer workflows, including:
API clients and custom backend services.
Coding tools such as Codex CLI and Codex Desktop.
Gemini CLI and other model-specific command-line workflows.
Provider-compatible API routes where supported.

3. Monitor#

After requests start flowing, the dashboard shows request-level metadata:
model and route tier;
API key used;
timestamp and status;
token usage where available;
estimated or final cost;
latency and failure signals;
remaining balance and balance movements.
Mote API does not store your prompts or model responses. To help with billing,
debugging, and service safety, Mote API keeps request logs for 30 days. These
logs contain operational metadata such as model, route tier, status, latency,
token usage when available, cost, API key identifier, and timestamp.

Route Tiers#

Mote API uses route tiers so teams can match cost, quality, and workload
criticality without changing the application integration every time.
TierPositioningBest forPricing message
LiteCost-first routingTesting, batch jobs, high-frequency calls, cost-sensitive usageSelected Lite routes may start around 12-14% of official pricing depending on model family and route availability.
CoreDefault recommended routingEveryday production traffic, balanced cost and reliabilityCore routes are designed as the default balance of price, availability, and practical stability.
MaxQuality-first routingCritical business flows, complex tasks, stability-sensitive workloadsMax routes prioritize route quality and stability while remaining below official pricing in supported cases.
Pricing varies by provider, model, route tier, upstream availability, and traffic
conditions. The percentages above describe selected route ranges. Current
pricing details are shown in the Mote API dashboard.

Why the Cost Can Be Lower#

Mote API can offer lower-cost route options because it separates workloads by
route tier instead of forcing every request through the same premium path.
The practical effect:
Lite routes optimize for low unit cost.
Core routes optimize for everyday reliability at a lower blended cost.
Max routes optimize for quality and stability when the business impact of
failure is higher.
This structure lets a team run experiments, automation, and bulk traffic through
lower-cost routes while keeping important production flows on stronger routes.

Stability and Failure Handling#

Mote API is more than a forwarding layer. It provides operational controls
around AI API traffic.
Mote API is designed around these operational controls:
API key authentication and key-level controls.
Balance checks before forwarding paid requests.
Rate limits and concurrency limits to protect capacity.
Route health monitoring and upstream channel disablement.
Failure-rate visibility for debugging and admin handling.
Retry or fallback policies where a route supports them.
This follows the same general principle used by established API gateway systems:
the gateway acts as a front door that handles traffic management, authorization,
monitoring, and routing before requests reach backend services.

Speed and Integration Experience#

The target integration experience is under five minutes from account creation to
the first successful API call:
1.
Register or sign in.
2.
Add prepaid balance.
3.
Create an API key.
4.
Configure your client or tool.
5.
Send the first request.
6.
Confirm usage and cost in the dashboard.
Mote API keeps integration fast by making the following easy to copy:
endpoint / route configuration;
API key;
SDK examples;
CLI examples;
curl examples;
common error explanations.

Model and Provider Coverage#

Mote API is positioned as a multi-provider hub, not a single-model wrapper.
Mote API supports:
GPT model routes;
Claude model routes;
Gemini model routes;
additional compatible providers and API routes.
The important point is not only model count. The stronger promise is that teams
can manage model access, route quality, billing, and usage visibility through one
operational surface.

Billing and Balance Ledger#

Mote API uses prepaid billing so users can control spend before production
traffic grows.
Billing behavior:
Top-ups are processed through Stripe.
Balance is posted after confirmed Stripe payment events, not only after a browser
redirect.
Every balance movement is recorded in a ledger.
Every billable request creates usage metadata.
Insufficient balance blocks requests before upstream cost is incurred.
Users can review recent spend, model usage, and request-level deductions.

Security and Compliance Principles#

Mote API operates as a legitimate API infrastructure service. It does not
support or promote bypassing provider rules, sanctions, account restrictions, or
access controls.
Core principles:
Users must comply with applicable laws and third-party provider policies.
API keys must be kept secret and rotated if compromised.
High-risk traffic can be limited, suspended, or reviewed.
Request logs are retained for 30 days and do not include prompts or model
responses.
Mote API publishes clear Terms of Service, Privacy Policy, balance rules, and
support contact information.

Why Developers Use Mote API#

Lower cost#

Lite, Core, and Max route tiers let teams choose the right price and quality
level for each workload instead of paying the same route cost for every request.

Better control#

Prepaid balance, API keys, usage logs, and route tier selection help developers
understand what is being used and what it costs.

More model coverage#

GPT, Claude, Gemini, and additional provider routes can be managed through one
dashboard and one service relationship.

Faster integration#

Developers can start with one account, one key, and a small set of client
configuration changes instead of building separate billing and monitoring logic
for every provider.

Practical reliability#

Route health, rate limits, failure visibility, and admin controls make the
gateway more useful for production workflows than a simple forwarding proxy.

Recommended First-Run Checklist#

Use this checklist when setting up Mote API for a new project:
Create an account.
Add prepaid balance through Stripe.
Create a project-specific API key.
Choose a default route tier: Lite, Core, or Max.
Configure your app, SDK, or CLI tool.
Send a small test request.
Confirm the request appears in usage logs.
Confirm the balance deduction is visible.
Set a low-balance alert when available.
Rotate keys periodically and disable unused keys.

FAQ#

Is Mote API an official provider service?#

No. Mote API is an independent API hub and gateway service. It is not affiliated
with or endorsed by OpenAI, Anthropic, Google, or other model providers unless
explicitly stated in a separate official partnership notice.

Is Lite always the best choice?#

No. Lite is designed for cost-first usage. Core is the recommended default for
most production traffic. Max is better for critical workflows where stability
and route quality matter more than lowest unit cost.

Are prices always a fixed percentage of official pricing?#

No. Route pricing can vary by model family, provider, route tier, availability,
and cost structure. Mote API shows selected route ranges and current pricing
details in the dashboard.

Does Mote API store prompts and responses?#

No. Mote API does not store prompts or model responses. It keeps request logs
for 30 days so users can review billing, troubleshoot failures, and understand
usage. These logs include operational metadata such as model, route tier, status,
latency, token usage when available, cost, API key identifier, and timestamp.

What happens if my balance is too low?#

Requests are blocked before they create upstream cost. The dashboard shows
insufficient-balance errors and helps users top up before continuing.
Modified at 2026-07-13 07:20:44
Previous
Create API Key
Next
Codex CLI
Built with