FAQ
Plain-language answers for people who have never used an AI API, plus links for developers who want the deeper detail. New to all of this? Start at getting started or the glossary.
Using OMA with your tools
Can I use OMA inside Claude Code, OpenCode, or Cursor?
Yes. Use the tool-specific setup at /docs#integrations with your oma_sk_... key, then choose an available catalog model. OpenAI-compatible tools use
https://www.oma-ai.com/api/v1; the Claude Code guide uses OMA's Anthropic-compatible endpoint. OpenCode, Claude Code, Cursor, and Continue are covered in the docs.How do I migrate an app that already uses OpenAI?
You keep your request shape and SDK. Change one thing — the base URL is now
https://www.oma-ai.com/api/v1 — swap the key, and specify an OMA catalog model per request. The openai-python package, Vercel AI SDK, and anything else that speaks OpenAI's API work as-is.Does OMA support the Anthropic Messages API?
Yes — OMA exposes an Anthropic-compatible Messages endpoint alongside the OpenAI-compatible
/chat/completions path, so tools written against the Anthropic SDK can also point at OMA with the same key. See /docs#endpoints for the exact paths and request shapes.Reliability & safety
How does OMA handle provider outages?
A circuit breaker keeps traffic off an unhealthy upstream route. If the requested model has a configured fallback, OMA can retry that mapped model on the fallback route; otherwise the request returns an upstream
/unavailable error rather than silently swapping to a different model. Check /status for live gateway health.Can a provider see my prompts?
OMA itself does not persist prompt or completion content and meters only billing metadata. However, the inference provider that serves a request must process that content to answer it, so provider-side retention guarantees depend on the model and provider you select. See /privacy for the exact data-flow.
What are the rate limits?
There are no per-plan rate tiers; a single shared token-bucket set applies, with per-model overrides where a provider needs them (for example llama-3.3-70b). The docs list the burst, refill, sustained RPM, and per-model token ceilings at /docs#rate-limits. Rate-limit headers are returned on each response.
Getting started
What is OMA?
OMA (Open Model Access) is a single place to use many AI models. Instead of signing up for ten different AI companies, you get one account, one API key, and one way to pay. It works exactly like OpenAI's API, so almost any AI app or tool that supports OpenAI already works with OMA.
Do I need to know how to code?
No. You can try AI immediately in the free chat playground at /chat — no account needed. Guests can use the three OMA-AI website-chat models (DeepSeek V4 Flash, MiniMax M2.7, and GLM-5.3 Flash) with a daily message cap. You only need the API (and an API key) when you want to build something, like an app or a script.
How do I start using it?
Three steps: (1) try free website chat on the three OMA-AI models — no account needed; (2) sign in with your email (magic link) or crypto wallet and create an API key; (3) add a little credit and call the API. The full walkthrough is on the /getting-started page.
What can I do with it?
Chat, write code, summarize documents, analyze data, work with image inputs, and more — across 21 models from many providers through one API.
Is there a free tier?
Yes. Website chat has a daily free allowance on eligible OMA-AI models, subject to availability, with more messages once you sign in. Any signup credit depends on the current offer and is shown in your account. API-key calls always spend credits, even on those models; the free tier lives in browser chat. See /pricing for limits.
Accounts, keys & privacy
What is an API key?
An API key is like a password that lets your code (or an AI tool you use) talk to OMA on your behalf. It looks like `oma_sk_...`. Keep it secret — anyone with your key can spend your credits. You can create and revoke keys from the dashboard.
Do you store my API keys?
Only the SHA-256 hash of your key is stored — not the key itself. The plaintext `oma_sk_...` is shown once when you create it and can't be viewed again; if you lose it, just revoke and create a new one. Keys are tied to your account and can be revoked instantly from the dashboard. OMA does not persist prompt or completion content. See /privacy for the full detail.
What is a magic link?
A magic link is a sign-in without a password. You enter your email, we send you a link, you click it, and you're signed in. No passwords to remember.
What is a wallet (SIWE) and do I need one?
SIWE (Sign-In With Ethereum) lets you sign in with a crypto wallet instead of an email. You do NOT need a wallet for email sign-in or card checkout when it is enabled. A wallet is needed for direct USDC deposits and the advanced x402 settlement flow.
Money & payments
How do I pay?
Deposit USDC on Base, then spend credits per request. Agents can also fund via x402 verify
/settle, which credits the same balance before API spend. Card checkout through Polar is available only when it is enabled, and the credits page shows its live status. Credits never expire while your account is active.What is USDC? What is Base?
USDC is a U.S.-dollar-denominated stablecoin issued by Circle. It aims to track $1, though its market price can vary. Base is the blockchain network where OMA accepts USDC payments. See /glossary for more detail.
What is x402?
x402 is a crypto payment flow for agents: sign a USDC authorization on Base, verify at
/x402/verify, settle to credit your account, then call the API with your key. Inference itself is billed from that credit balance — not attached as a mid-request 402 on chat today.How much does it cost?
API cost depends on the selected model and the number of input and output tokens. Rates are quoted per million tokens, with no subscription or monthly fee. Website chat has a separate daily free allowance on eligible models. See /pricing for current rates and limits.
What happens when I run out of credits?
Requests stop with a clear 'insufficient quota' message until you top up. Your keys, conversations, and account all stay intact.
Models, tokens & chat
What is a model?
A model is a specific AI system with different strengths, context limits, and prices. You pick the model id per request; OMA routes that requested model to its configured upstream provider.
What is a token?
A token is roughly a piece of a word — 'hello' is about one token, 'hello world' about two. AI prices and limits are measured in tokens (and you'll see 'tokens per second' while chatting). A typical English word is 1–2 tokens.
What is a context window?
The context window is how much text a model can see at once — your whole conversation plus the answer. OMA-AI models currently span 200K to 1M tokens, and several Venice-routed models also reach 1M. Each model page shows its exact limit.
What does 'OpenAI-compatible' mean?
It means OMA speaks the same language as OpenAI's API. Apps, libraries, and tutorials written for OpenAI (like the openai-python package or the Vercel AI SDK) work with OMA by changing one line: the API base URL.
What is streaming?
Streaming shows the answer as it's being written, word by word, instead of waiting for the whole thing. Most chat apps use it — it feels faster and you can read along.
Which model should I pick?
Start with a budget or mid model for everyday chat, and reach for frontier models for hard math, code, or long documents. The chat UI shows each model's price, context length, and strengths, and /models has the full catalog.
Do you support multimodal inputs?
Some catalog models accept image inputs alongside text. Check the model page or capability filters before relying on vision support. Compatible models use the OpenAI multimodal `messages[].content` array shape (for example text and image_url parts).
For developers
Where are the API docs?
The full reference is at /docs — quickstart, authentication, every endpoint, streaming, rate limits, errors, and the x402 flow. There's a search box at the top.
Is there a quickstart?
Yes — /getting-started walks through your first request in under five minutes with copy-paste code, and /docs#quickstart has the developer version.
Can my AI agent use OMA?
Yes. /agent-skill.md explains how to sign in, create an API key, discover available models, and make a request. It also documents optional x402 funding, which requires the payer's wallet session and moves real funds. OMA publishes an agent card at /.well-known/agent-card.json and a machine-readable index at /llms.txt.
What happens when a model provider is down?
The circuit breaker stops repeatedly sending traffic to an unhealthy route. If the requested model has a configured fallback, OMA can retry that mapped model there. If no fallback is configured, the request can return an upstream
/unavailable error rather than silently substituting a different model.I got a 402 error — what now?
A 402 means payment is required — usually your credit balance is empty. Top up at /deposit (USDC or card when enabled), or use the x402 verify
/settle flow to credit your balance, then retry with an API key (docs: /docs#x402).How do I troubleshoot x402 errors?
Check the Base network, USDC balance, current treasury recipient, signature fields, validity window, and a fresh nonce. EIP-3009 uses a signed transfer authorization; a separate token approval is not required. Settlement needs a SIWE session for the same wallet that paid. If a transaction hash was returned, check it before authorizing another transfer. See /docs#x402 for the full flow.