Transparent pricing
Prepaid credit. No subscriptions, no seats, no minimums. Every response tells you exactly what it cost and what's left.
$0.40 /1M tok
OpenAI-compatible completions on MLX.
POST /v1/chat/completions
$0.02 /1M tok
384-dim semantic vectors for RAG, search and clustering.
POST /v1/embeddings
$0.001 /query
Private per-key vector store + retrieval. No infra to run.
POST /v1/search
$0.05 /call
Live cross-asset signals + an AI analyst narrative.
POST /v1/briefing
Concrete, at the rates above — so you can size a budget before writing any code.
Reranking and document upserts are billed at the embedding rate, on the tokens they actually embed — there is no separate per-call fee.
Every new key starts with $0.05 of credit, issued instantly by
POST /v1/trial. That's ~125,000 chat tokens — plenty to evaluate the API properly.
One trial key per network per 24 hours.
No. Prepaid credit only — top up what you want, spend it at the metered rates, nothing recurs.
No. It stays on your key until you spend it.
POST /v1/checkout creates a card checkout session; credit lands on the
same key. Keys are never rotated on top-up.
Calls return 402 insufficient_quota and stop billing — no overage, no
surprise invoice. Top up and the same key resumes.
Yes — GET /pricing
is public, and GET /v1/usage returns your balance and spend.
Get an API key with $0.05 free credit — instantly. No card required.
Save it now. Point any OpenAI SDK at api.vellaquant.com/v1.