Everything public about Nexara.
What Nexara is, where it runs, how Compute and the usage limits work, what each plan actually covers — and which plan fits which kind of usage. If you found Nexara through search and want the straight facts before trying it, this is the page.
What is Nexara
Nexara is an AI assistant and creation suite that runs everywhere you work: in the browser, as an Android app, as a Windows desktop app, and from a terminal. It is a real product with real accounts, real billing, and a real Compute currency — every feature below is something you can use today. asknexara.com is Nexara's official, current domain — the same Nexara service that was previously reachable at the nexara-ai-chat.vercel.app address during early development.
Chat with strong models
A full model catalog — frontier, reasoning, coding, speed and multimodal models — all billed from one Compute balance.
Create, don't just describe
Image Studio (GPT Image), Slides, Website and Presentation builder modes, PDFs, games, widgets, 3D models — real artifacts rendered in the chat.
OS-level tools on your devices
The mobile/desktop assistants can read your screen, open apps, toggle flashlight/brightness/media, take screenshots and control device settings end-to-end.
Android app
Distributed as a direct APK (not on the Play Store yet) — plus a Windows desktop build with automatic updates.
NexaraCode + CLI
An IDE desktop shell and a standalone CLI for agent-style coding sessions against the same account.
Builders & agents
Slides, websites and presentations from a bare prompt; Planner and autonomous /goal runs; user-made Nexar personas.
Compute & the usage limits
Compute is Nexara's currency — the same money as a dollar, just renamed. The exchange rate is exact and it never changes: 1 Compute = $0.000001, so 100K Compute = $0.10 and 1M Compute = $1.00, 100% — no hidden spread between the two. If a model's input costs the provider $1 per million tokens, it costs exactly 1,000,000 Compute here. Every model is billed from that pool at its own actual provider rate, and you can subscribe to a plan for a monthly allowance, top up a balance, or earn free Compute.
1M Compute = $1.00
Exactly, always. Compute is Nexara's name for the dollar at a fixed rate — 1M Compute and $1 are the same amount, and you only ever see Compute.
Daily allowance
Plan users get a daily Compute allowance plus a generous weekly hard cap — no monthly limit, no 5-hour window.
Weekly hard cap
A weekly ceiling bounds total burn so one runaway session can't exhaust a month in a day.
12,500 Compute per image
GPT Image generation costs a flat 12,500 Compute per image regardless of size.
What that actually buys you
- One casual chat reply — usually a few hundred to a couple of thousand Compute, depending on the model and how long the answer is.
- One Image Studio picture — exactly 12,500 Compute, any size, any engine.
- One heavy coding turn (agent session, big artifact build) — tens of thousands of Compute and up; this is why the coding plans carry much larger allowances.
- MiniMax M3 is 100% free and never touches your balance — useful when you want an answer without spending anything.
- Exact per-model prices live on the Docs page — every rate is per 1M tokens, at the same fixed exchange rate as above.
How limits actually work
- Free plan. No credit card. A small automatic allowance (about 5M Compute/day) refills daily — enough to try the product, but it will not survive heavy or long coding sessions.
- Refills daily. Paid plans refresh their Compute allowance every day, with a generous weekly cap — your live numbers always show in Settings → Compute.
- Hitting the limit. When the daily allowance or the weekly cap is reached, new requests are paused until the day refills or the week resets on Monday. Nexara refuses a request before it starts if your balance can't cover it — it never silently overcharges.
- Top-ups & overage. You can top up Compute at any time; overage beyond an allowance is billed from the balance at 500K Compute per extra 1M tokens.
- Billing multiplier. Token usage is counted with a flat 5% surcharge covering payment-processing and currency-conversion fees on API credit purchases. There are no other hidden fees.
Plans, prices & limits
Which plan is for you? The $5 Lite plan is for testing Nexara — expect usage limits to be hit quickly under heavy use. $10 Pro is the middle tier: how far it goes depends on what you're doing (normal chat lasts; big builds don't). The higher tiers are for when it makes sense to actually code with Nexara — agent sessions, the IDE, long automation — without hitting usage limits every few minutes.
Lite
$5/moTrying out Nexara
- Daily spend cap
- 12.3M
- Weekly spend cap
- 61.4M
- Images per day
- 35
- Library slots
- 100
The daily and weekly figures are the only limits on this plan — your Compute refills each day and is hard-capped each week. There is no monthly limit and no 5-hour window: the weekly cap refreshes every Monday.
Built for testing the product. Compute is real but limited — heavy or long sessions can hit the limit quickly. If you mainly chat casually, it goes further.
Compare on the pricing pagePro
$10/moThe mid tier
- Daily spend cap
- 24.6M
- Weekly spend cap
- 122.9M
- Images per day
- 70
- Library slots
- 200
The daily and weekly figures are the only limits on this plan — your Compute refills each day and is hard-capped each week. There is no monthly limit and no 5-hour window: the weekly cap refreshes every Monday.
Comfortable for everyday chat and light creation. How far it goes depends on what you're doing: normal conversation lasts, but big image runs or heavy coding will drain it noticeably faster.
Compare on the pricing pagePro+
$20/moThe sweet spot
- Daily spend cap
- 49.1M
- Weekly spend cap
- 245.7M
- Images per day
- 175
- Library slots
- 500
The daily and weekly figures are the only limits on this plan — your Compute refills each day and is hard-capped each week. There is no monthly limit and no 5-hour window: the weekly cap refreshes every Monday.
Recommended for regular creation plus light coding. 2× the Pro allowance at 2× the price.
Compare on the pricing pageMax
$50/moBuilding with Nexara
- Daily spend cap
- 122.9M
- Weekly spend cap
- 614.3M
- Images per day
- Unlimited
- Library slots
- 1,000
The daily and weekly figures are the only limits on this plan — your Compute refills each day and is hard-capped each week. There is no monthly limit and no 5-hour window: the weekly cap refreshes every Monday.
For when you actually code with it — agent sessions, IDE use, big artifact builds — without watching the meter run dry.
Compare on the pricing pageMax (Ultra)
$100/moHeavy, all-day usage
- Daily spend cap
- 491.4M
- Weekly spend cap
- 2.5B
- Images per day
- Unlimited
- Library slots
- 1,000
The daily and weekly figures are the only limits on this plan — your Compute refills each day and is hard-capped each week. There is no monthly limit and no 5-hour window: the weekly cap refreshes every Monday.
20× the Pro allowance. Long agent runs, automation, and serious development work that would otherwise hit usage limits quickly.
Compare on the pricing pageBilling details
- Monthly. Billed every month at the list price.
- Quarterly −10% / Yearly −20%. Longer commitments are charged upfront as one payment for the whole period at a discount.
- Pro+ intro. New Pro+ subscribers get 25% off their first 3 monthly cycles ($15 instead of $20), then it renews at full price automatically.
- Custom plans. Build your own: pick a daily Compute allowance and Nexar slots; the weekly cap is set alongside it (same ×20 basis standard plans use).
- Gifting. Plans can be gifted, and upgrades extend rather than replace.
Images & the Library
Image Studio generates real pictures with OpenAI's GPT Image at a flat 12,500 Compute per image (any size) — that's $0.0125 (1.25¢) per image, roughly 80 images per dollar, with no quality or resolution tiers: one flat price, any size. Per-plan daily image caps and Library storage:
| Plan price | Images per day | Library slots |
|---|---|---|
| Free | 4 | 50 |
| $5 Lite | 35 | 100 |
| $10 Pro | 70 | 200 |
| $20 Pro+ | 175 | 500 |
| $50 Max and up | Unlimited | 1,000 |
Library slot limits on paid plans scale with tier ("1,000" on Max/Ultra); the Free plan has a fixed 50-slot library. Slots can be purchased on top of the plan allowance.
Models
Nexara's catalog covers frontier, reasoning, coding, speed and multimodal models from OpenAI, Anthropic, Google, Qwen, MiniMax and more. No model is plan-gated — every plan unlocks the full lineup; only the Compute allowance changes. What you get is always the provider's full, unmodified model — full context window, full output length, normal reasoning modes. Nexara never substitutes a secretly smaller or watered-down stand-in, and no plan hides a model behind a paywall. MiniMax M3 is 100% free and never touches your balance.
Here is the complete per-model Compute reference — every model in the catalog, with its cost per million tokens for input, output, and cached input, plus image/video input support and context windows. 1M Compute = $1 — the dollar figure under each price is the internal provider rate. The Docs page has the same table with the full media-input breakdown.
Frontier
25 models| Model | Compute / 1M in | Compute / 1M out | Compute / 1M cache | Image | Video | Context |
|---|---|---|---|---|---|---|
GPT-OSS-120B openai/gpt-oss-120b | 100K$0.10 | 400K$0.40 | — | 131K | ||
MiniMax M3Free minimax/minimax-m3:free | Free | Free | — | 1M | ||
Grok 4.5 x-ai/grok-4.5 | 1.74M$1.74 | 5.27M$5.27 | — | 500K | ||
Grok 4.6 x-ai/grok-4.6 | 2M$2.00 | 6M$6.00 | — | 500K | ||
GPT-5.6 Luna openai/gpt-5.6-luna | 200K$0.20 | 1.05M$1.05 | — | 1M | ||
GPT-5.6 Terra openai/gpt-5.6-terra | 2.05M$2.05 | 11.75M$11.75 | — | 1M | ||
GPT-5.6 Sol openai/gpt-5.6-sol | 5M$5.00 | 30M$30.00 | 500K$0.50 | 1M | ||
Claude Sonnet 4.6 anthropic/claude-sonnet-4.6 | 3M$3.00 | 15M$15.00 | 300K$0.30 | 1M | ||
Claude Sonnet 5 anthropic/claude-sonnet-5 | 2M$2.00 | 10M$10.00 | 200K$0.20 | 1M | ||
Claude Opus 4.6 anthropic/claude-opus-4.6 | 5M$5.00 | 25M$25.00 | 500K$0.50 | 1M | ||
Claude Opus 4.7 anthropic/claude-opus-4.7 | 5M$5.00 | 25M$25.00 | 500K$0.50 | 1M | ||
Claude Opus 4.8 anthropic/claude-opus-4.8 | 5M$5.00 | 25M$25.00 | 500K$0.50 | 1M | ||
Claude Opus 5 anthropic/claude-opus-5 | 5M$5.00 | 25M$25.00 | 500K$0.50 | 1M | ||
Claude Fable 5 anthropic/claude-fable-5 | 10M$10.00 | 50M$50.00 | 1M$1.00 | 1M | ||
Kimi K3 moonshotai/kimi-k3 | 3M$3.00 | 15M$15.00 | 300K$0.30 | 262K | ||
Kimi K2.6Locked moonshotai/kimi-k2.6 | 950K$0.95 | 4M$4.00 | — | 262K | ||
Qwen 3.8 Max qwen/qwen3.8-max | 2.55M$2.55 | 7.55M$7.55 | — | 1M | ||
Qwen 3.7 Max qwen/qwen3.7-max | 2.55M$2.55 | 7.55M$7.55 | — | 1M | ||
Qwen 3.6 Max (Preview) qwen/qwen3.6-max-preview | 1.35M$1.35 | 7.85M$7.85 | — | 256K | ||
Qwen 3.5 397B A17B qwen/qwen3.5-397b-a17b | 650K$0.65 | 3.65M$3.65 | — | 256K | ||
Qwen3 Max qwen/qwen3-max | 1.25M$1.25 | 6.05M$6.05 | — | 256K | ||
GLM 5 z-ai/glm-5 | 630K$0.63 | 1.95M$1.95 | — | 200K | ||
GLM 5.1 z-ai/glm-5.1 | 1.42M$1.42 | 4.42M$4.42 | — | 200K | ||
GLM 5.2 z-ai/glm-5.2 | 1.5M$1.50 | 4.52M$4.52 | — | 1M | ||
GLM 5.3 z-ai/glm-5.3 | 1.4M$1.40 | 4.4M$4.40 | 260K$0.26 | 1M |
Reasoning
15 models| Model | Compute / 1M in | Compute / 1M out | Compute / 1M cache | Image | Video | Context |
|---|---|---|---|---|---|---|
Step 3.7 Flash stepfun/step-3.7-flash | 100K$0.10 | 300K$0.30 | — | 256K | ||
Nemotron 3 Nano 30B A3B nvidia/nemotron-3-nano-30b-a3b | 50K$0.05 | 150K$0.15 | — | 256K | ||
DeepSeek V4 Flash 07.31 deepseek/deepseek-v4-flash-0731 | 40K$0.04 | 130K$0.13 | 9K$0.01 | 1M | ||
Xiaomi Mimo V2.5 ProLocked xiaomi/mimo-v2.5-pro:free | 430K$0.43 | 870K$0.87 | — | 128K | ||
Xiaomi Mimo V2.5Locked xiaomi/mimo-v2.5:free | 430K$0.43 | 870K$0.87 | — | 128K | ||
Kimi K2.5Locked moonshotai/kimi-k2.5 | 570K$0.57 | 2.85M$2.85 | — | 262K | ||
Nemotron 3 Nano nvidia/nemotron-3-nano | 100K$0.10 | 400K$0.40 | — | 1M | ||
Nemotron 3 Super nvidia/nemotron-3-super | 250K$0.25 | 1M$1.00 | — | 1M | ||
Nemotron 3 Ultra nvidia/nemotron-3-ultra | 500K$0.50 | 2M$2.00 | — | 1M | ||
Qwen 3.6 27B qwen/qwen3.6-27b | 650K$0.65 | 3.65M$3.65 | — | 256K | ||
Qwen 3.6 35B A3B qwen/qwen3.6-35b-a3b | 298K$0.30 | 1.54M$1.53 | — | 256K | ||
GLM 4.7 z-ai/glm-4.7 | 640K$0.64 | 2.24M$2.24 | — | 200K | ||
DeepSeek V3.2 deepseek/deepseek-v3.2 | 280K$0.28 | 420K$0.42 | — | 131K | ||
DeepSeek V4 Flash deepseek/deepseek-v4-flash | 250K$0.25 | 1M$1.00 | — | 1M | ||
DeepSeek V4 Pro deepseek/deepseek-v4-pro | 600K$0.60 | 2.4M$2.40 | — | 1M |
General
20 models| Model | Compute / 1M in | Compute / 1M out | Compute / 1M cache | Image | Video | Context |
|---|---|---|---|---|---|---|
MiniMax M2 minimax/minimax-m2:free | 300K$0.30 | 1.2M$1.20 | — | 205K | ||
MiniMax M2.1 minimax/minimax-m2.1:free | 300K$0.30 | 1.2M$1.20 | — | 205K | ||
MiniMax M2.5 minimax/minimax-m2.5:free | 300K$0.30 | 1.2M$1.20 | — | 205K | ||
MiniMax M2.7 minimax/minimax-m2.7:free | 300K$0.30 | 1.2M$1.20 | — | 205K | ||
Ministral 14B mistralai/ministral-14b | 200K$0.20 | 200K$0.20 | — | 256K | ||
Mistral Small 26.03 mistralai/mistral-small-2603 | 100K$0.10 | 300K$0.30 | — | 256K | ||
Mistral Medium 3.5 mistralai/mistral-medium-3.5 | 350K$0.35 | 1.05M$1.05 | — | 256K | ||
Mistral Large 3 mistralai/mistral-large-2512 | 2M$2.00 | 6M$6.00 | — | 256K | ||
Ling 3.0 Flash inclusion-ai/ling-3.0-flash | 50K$0.05 | 150K$0.15 | — | 260K | ||
Llama 3.3 70B Instruct meta/llama-3.3-70b-instruct | 150K$0.15 | 600K$0.60 | — | 131K | ||
Llama 3.1 8B Instruct meta/llama-3.1-8b-instruct | 40K$0.04 | 100K$0.10 | — | 16K | ||
Qwen 3.7 Plus qwen/qwen3.7-plus | 450K$0.45 | 1.65M$1.65 | — | 1M | ||
Qwen 3.6 Plus qwen/qwen3.6-plus | 550K$0.55 | 3.05M$3.05 | — | 1M | ||
Qwen 3.5 Plus qwen/qwen3.5-plus | 450K$0.45 | 2.45M$2.45 | — | 1M | ||
Qwen Plus 07.28 qwen/qwen-plus-2025-07-28 | 450K$0.45 | 1.25M$1.25 | — | 1M | ||
GLM 4.5 z-ai/glm-4.5 | 630K$0.63 | 2.23M$2.23 | — | 131K | ||
DeepSeek Chat V3.1 deepseek/deepseek-chat-v3.1 | 280K$0.28 | 420K$0.42 | — | 164K | ||
Llama 3.3 Nemotron Super 49B nvidia/llama-3.3-nemotron-super-49b | 150K$0.15 | 600K$0.60 | — | 131K | ||
SenseNova 6.7 Flash-Lite sensenova/sensenova-6.7-flash-lite | 20K$0.02 | 80K$0.08 | — | 262K | ||
SenseNova 6.8 Flash-Lite sensenova/sensenova-6.8-flash-lite | 20K$0.02 | 80K$0.08 | — | 262K |
Coding
8 models| Model | Compute / 1M in | Compute / 1M out | Compute / 1M cache | Image | Video | Context |
|---|---|---|---|---|---|---|
Devstral Medium mistralai/devstral-medium | 500K$0.50 | 650K$0.65 | — | 256K | ||
Codestral 25.08 mistralai/codestral-2508 | 300K$0.30 | 900K$0.90 | — | 256K | ||
Laguna XS.2 poolside/laguna-xs.2 | 100K$0.10 | 400K$0.40 | — | 131K | ||
Grok Build 0.1 x-ai/grok-build-0.1 | 1M$1.00 | 2M$2.00 | 200K$0.20 | 256K | ||
GPT-5.3 Codex Spark openai/gpt-5.3-codex-spark | 150K$0.15 | 600K$0.60 | — | 128K | ||
Kimi K2.7 CodeLocked moonshotai/kimi-k2.7-code | 700K$0.70 | 3.5M$3.50 | 150K$0.15 | 262K | ||
Qwen3 Coder Plus qwen/qwen3-coder-plus | 1.05M$1.05 | 5.05M$5.05 | — | 1M | ||
GLM 4.6 z-ai/glm-4.6 | 620K$0.62 | 2.22M$2.22 | — | 200K |
Speed
13 models| Model | Compute / 1M in | Compute / 1M out | Compute / 1M cache | Image | Video | Context |
|---|---|---|---|---|---|---|
MiniMax M2.1 High-Speed minimax/minimax-m2.1-highspeed:free | 150K$0.15 | 600K$0.60 | — | 205K | ||
MiniMax M2.5 High-Speed minimax/minimax-m2.5-highspeed:free | 150K$0.15 | 600K$0.60 | — | 205K | ||
MiniMax M2.7 High-Speed minimax/minimax-m2.7-highspeed:free | 150K$0.15 | 600K$0.60 | — | 205K | ||
Ministral 3B mistralai/ministral-3b | 50K$0.05 | 50K$0.05 | — | 128K | ||
Ministral 8B mistralai/ministral-8b | 100K$0.10 | 100K$0.10 | — | 256K | ||
Nemotron Nano 9B V2 nvidia/nemotron-nano-9b-v2 | 30K$0.03 | 100K$0.10 | — | 128K | ||
Llama 3.2 1B Instruct meta/llama-3.2-1b-instruct | 20K$0.02 | 50K$0.05 | — | 16K | ||
Llama 3.2 3B Instruct meta/llama-3.2-3b-instruct | 30K$0.03 | 80K$0.08 | — | 16K | ||
Claude Haiku 4.5 anthropic/claude-haiku-4.5 | 1M$1.00 | 5M$5.00 | 100K$0.10 | 200K | ||
Qwen 3.5 Flash qwen/qwen3.5-flash | 150K$0.15 | 450K$0.45 | — | 1M | ||
GLM 4.5 Air z-ai/glm-4.5-air | 220K$0.22 | 1.12M$1.12 | — | 131K | ||
GLM 5 Turbo z-ai/glm-5-turbo | 1.24M$1.24 | 4.04M$4.04 | — | 200K | ||
GLM 5.3 Flash z-ai/glm-5.3-flash | 150K$0.15 | 500K$0.50 | 30K$0.03 | 1M |
Multimodal
13 models| Model | Compute / 1M in | Compute / 1M out | Compute / 1M cache | Image | Video | Context |
|---|---|---|---|---|---|---|
Gemini 3.6 FlashLocked google/gemini-3.6-flash | 1.5M$1.50 | 7.5M$7.50 | — | 1M | ||
Gemini 3.7 FlashLocked google/gemini-3.7-flash | 380K$0.38 | 1.88M$1.88 | 40K$0.04 | 1M | ||
Gemini 3.5 FlashLocked google/gemini-3.5-flash | 1.5M$1.50 | 9M$9.00 | — | 1M | ||
Gemini 3.1 ProLocked google/gemini-3.1-pro | 2M$2.00 | 12M$12.00 | — | 1M | ||
Gemini 3 FlashLocked google/gemini-3-flash | 500K$0.50 | 3M$3.00 | — | 1M | ||
Gemini 2.5 FlashLocked google/gemini-2.5-flash | 300K$0.30 | 2.5M$2.50 | — | 1M | ||
Gemini 2.5 ProLocked google/gemini-2.5-pro | 1.25M$1.25 | 10M$10.00 | — | 1M | ||
Qwen 3.5 Omni Plus qwen/qwen3.5-omni-plus | 1.45M$1.45 | 11.05M$11.05 | — | 128K | ||
Qwen 3.5 Omni Flash qwen/qwen3.5-omni-flash | 450K$0.45 | 3.05M$3.05 | — | 128K | ||
Qwen3 VL Plus qwen/qwen3-vl-plus | 250K$0.25 | 1.65M$1.65 | — | 256K | ||
Qwen3 Omni Flash qwen/qwen3-omni-flash | 480K$0.48 | 1.71M$1.71 | — | 128K | ||
Ox Alpha stealth/ox-alpha-free | 500K$0.50 | 1.25M$1.25 | — | 1M | ||
Nemotron 3 Nano Omni nvidia/nemotron-3-nano-omni | 50K$0.05 | 150K$0.15 | — | 256K |
Prices are Compute per million tokens at each model's actual provider rate. The cache column only appears when the model's provider advertises a separate cache-read rate — otherwise cached input bills at the full input rate. Context is the model's maximum window in tokens.
How requests reach the models. Nexara mostly routes model calls through third-party model gateways (routers) — OpenRouter is the main one — and for some models it calls the official provider's own API directly instead. A router is just the delivery path: it forwards your request to the real provider (OpenAI, Anthropic, Google, and so on) and relays the answer back. Routing never replaces, shrinks, or swaps the model you picked — the model that answers is the model named in your conversation, whichever gateway carried the request. If a gateway is down, a request may fall over to the next route in that model's chain, and the chat tells you when that happens.
Your chats, settings, personas and preferences follow your account across the web, Android, Windows and the CLI. Guests can try the chat without an account; Compute, the Library, and billing features require signing in.
Where Nexara runs
Web app
The full experience in the browser — chat, builders, image studio, library, settings.
Android (APK)
Direct APK install with OS-level assistant tools. Not on the Play Store yet.
Windows desktop
Electron shell with automatic updates — new versions appear in-app when a release is published.
NexaraCode (IDE)
A desktop IDE shell for agent-driven coding sessions.
Nexara CLI
A standalone command-line tool sharing the same account and models.
Agent Visualization SDK
Connect your own front-end to local agents over a small WebSocket protocol.
Security & routing, plainly
Two things people ask us about are how their data is stored and which AI providers actually answer their messages. Here is exactly how both work — no fine print.
Your data in the database
- Chats live in Supabase — a managed Postgres database — tied to your account and synced across web, Android, Windows, and the CLI.
- Row Level Security is on. Every thread and message row is scoped to its owner (
auth.uid()); the database only ever returns rows that belong to the signed-in user. A user cannot read or write another user's chats by guessing IDs. - The only exceptions are opt-in: threads you explicitly mark public (share links) and threads where you invited a collaborator. Nothing else opens.
- The rules ship with the code. The RLS policies live in the web app's version-controlled database migrations — anyone can read them at github.com/K1NGMR/nexara-ai-chat under
supabase/.
Which provider answers you
- Your message goes to a real provider. The chat calls the model you picked — either straight to that provider's own API (Claude through Anthropic, GPT through OpenAI) or through an open model router.
- Routers are delivery layers. A router like OpenRouter aggregates many providers behind one API; it forwards your request to the actual model host and relays the answer back. It never swaps, shrinks, or rewrites the model — the model named in your chat is the model that answers.
- The model shown is the model used. The chat always labels which model served each reply, and if a gateway is down it tells you when a request falls over to another route.
- The full catalog is on the Docs page, with per-model pricing and reasoning options. Any model on any plan — nothing is gated or hidden.
Why there's no routing map
We can't give you a fixed map of which provider serves each model — we adjust routing constantly, sometimes daily. Even inside OpenRouter the providers behind a model change: a cheaper host for the same model can appear at any moment, and when it does we move the route. We can't honestly guarantee a 24/7 routing map, because the whole point of Nexara is to give you the cheapest possible API access — and that isn't achievable if we lock ourselves to a single provider inside a router.
What never changes is the contract you see: the model is always the model you picked, and the chat always tells you which one served the reply. Only the behind-the-scenes delivery path moves — always toward the same model at the best real-world price we can find.
Why the backend isn't open-source
Secrets are only half the story. Even with keys tucked safely in environment variables, publishing the backend source would hand anyone the full code to hunt for logic flaws — auth bypasses, rate-limit gaps, IDOR bugs — and then probe our live deployment with exactly what they found. Public frontend code is normal; public backend code is free vulnerability research against us.
So here's the honest balance we strike: we were planning to open-source parts of the backend, but we won't publish the server code itself. What you can always inspect is everything that governs your data and your money — the database security rules, the Compute and pricing logic shown openly on this page, and the entire CLI source. Our security model is public even when our code isn't.
What actually leaves your device for a given request is the minimum context needed to answer it — spelled out further in the Privacy Policy.
The team behind Nexara
Nexara is built by a small group of independent developers — between two and five of us, depending on the week. There is no big company and no investor money behind it: every model integration, app, builder, and bug fix you see here is designed, coded, and shipped by that handful of people.
We also plan to keep AI usage as cheap as we possibly can. Every model is billed at its actual provider rate, we price cached tokens at the cache rate, and we run regular optimization passes to cut what each turn really costs — so the Compute you buy goes as far as possible.
Who actually runs it? Nexara is run by a small team of developers — and we keep our individual identities private, by choice, for our own safety and peace of mind. There's no named CEO and no corporate brand to point at: the people who code Nexara are the people who run it.
Email support@asknexara.com or use the in-app feedback or report an error options — both route straight to the small, trusted team that fixes issues, never shown publicly.
Current goal
Nexara Cowork gets its own basic computer — a dedicated machine so the agent can run longer, heavier jobs without competing with your device.
Every paid plan gives $1 to people who are blind
With every paid plan billing — yours and everyone else's — we donate $1 to support people who are blind or have low vision. It's a commitment we make from the subscription, never an extra charge on top.
But please don't buy a plan purely because it feels like a donation — that's not what plans are for. Buy one because you genuinely need a service with a large lineup of AI models: chat LLMs, image generation, and soon video models too. If you simply want to help, the donate page is the honest way to do it.
Any donation helps a lot
A small team means every donation goes a long way — it covers real model inference and server costs, and lets us keep a generous free tier for people who can't pay. If Nexara has been useful to you, even a few dollars makes a real difference.
Frequently asked
🔒 Privacy — what happens to my chats and files?
Your conversations and files belong to you. We don't read them, we don't sell them, and we don't train on them. Chats are saved to your account in Supabase so you can pick them up on any device, and memory features only ever keep what you explicitly choose.
The people running Nexara don't browse conversations. What reaches us during normal use is technical usage only — typically a webhook notice like "user just used GPT-5.6 Sol" so we can watch costs and reliability. No message text or file content rides along with those notices, and no other data about your activity is saved.
Your name and email only ever leave your account if you send us feedback or report an error — and even then they are never shown publicly. They stay with the small, trusted team that fixes the issue, only for as long as the fix needs them.
🔴 What should I never paste into a chat?
The short rule is the same as with ChatGPT, Claude, or Gemini: if it's a secret, don't paste it. Passwords and API keys, private documents, personal information (ID numbers, addresses), proprietary code, and confidential business information should stay out of any AI chat — including Nexara's. To answer you, the relevant parts of a conversation travel to the model provider for that request (the Privacy Policy spells out who sees what); a secret pasted into a chat is a secret that has left your control.
Need help with sensitive code or documents? Paste a redacted version with the real secrets replaced by placeholders — that's enough for almost every question.
💰 Compute — how much usage do I actually get?
Compute is Nexara's currency: 1M Compute = $1.00, and every model bills from that pool at its own provider rate. The live number you actually have is always in Settings → Compute — daily allowance, weekly cap, and balance.
To make it concrete (rough figures — they vary with model and message length): a short casual reply usually costs a few hundred to a couple of thousand Compute; a long coding turn that builds something big can run tens of thousands; one Image Studio picture is exactly 12,500. So a few million Compute is hundreds of everyday messages — or roughly ten images plus casual chatting — and it goes fast under heavy agent/IDE sessions.
The free plan refills about 5M/day. The daily allowance and weekly cap for every paid plan are in the plan cards above.
🤖 Model access — full models or restricted versions?
Full models — the real thing. Every plan can use the entire catalog, and what you get is the provider's actual model with its full context window and reasoning modes. Nexara never serves secretly smaller, quantized, or watered-down stand-ins; names like "Lite" or "Flash" in the catalog are the providers' own model names, not downgrades we apply.
Plans don't restrict which models you can pick — they only change how much Compute you have to spend. The free tier gets the same full models with a small daily allowance. If a provider model is briefly down, a request may fall back to another capable model, and the chat always tells you when that happens.
⚡ Speed & reliability — does it stay usable when lots of people are online?
Mostly — but we'll be straight about it: Nexara is run by a small team, not a company with giant infrastructure budgets. When a lot of people are online at once, replies can slow down, and there are real limits (the daily allowances and weekly caps) that keep the service from being swamped. We scale what we can within our budget and watch load continuously.
In normal use it stays usable. If we ever hit a rough patch, you'll see it acknowledged honestly in the Updates log rather than hidden.
🛠️ Features — are the coding and agent tools actually good?
They're real and people use them every day — file-building artifacts with live previews, the agent loop, websites and slides, NexaraCode, the CLI, and the desktop and Android assistants. We wouldn't claim we're the best at everything; bigger tools with bigger teams exist. What we can promise is that we genuinely try to make ours good and keep improving them.
The IDE and NexaraCode are still in Beta — you will hit rough edges. Try everything yourself, and if something is weak, tell us exactly what was weak: feedback is the roadmap.
📈 Longevity — will Nexara still be around in a year?
That's the plan. We intend to keep Nexara running for a very long time. Realistically, staying up long-term depends on Nexara getting enough support (plans, donations) and feedback to cover the model bills and keep the team shipping.
As long as that support keeps coming, we plan to keep going — and if things ever change, you'll hear it from us clearly and in advance, never find out after the fact.
📊 Rate limits — every Compute number, for every plan
All limits below are in Compute — the same money as dollars (1M Compute = $1). Paid plans refill a daily allowance with a hard weekly cap — that is the whole limit system: no monthly limit, no rolling 5-hour window. One image always costs exactly 12,500. The live numbers always win and live in your Settings → Compute.
| Plan | Daily allowance | Weekly cap |
|---|---|---|
| Free | 5M | — |
| Lite · $5/mo | 12.3M | 61.4M |
| Pro · $10/mo | 24.6M | 122.9M |
| Pro+ · $20/mo | 49.1M | 245.7M |
| Max · $50/mo | 122.9M | 614.3M |
| Max (Ultra) · $100/mo | 491.4M | 2.5B |
These are the only limits: the daily allowance refills every day, and the weekly cap resets every Monday. There is no monthly limit and no 5-hour window.
🔴 Billing & refunds — renewals, cancelling, unused Compute
Automatic renewal? Yes — paid plans are subscriptions billed automatically on the interval you chose at checkout. Monthly renews monthly; quarterly (−10%) and yearly (−20%) are charged upfront as one payment and renew at the end of their period.
Cancelling is easy. Turn off future renewals any time from your billing controls — no contracts and no cancellation fees. The subscription simply stops at the end of the paid period you already bought, and your account drops to the free tier: chats, files, Library and personas stay exactly where they are.
Unused Compute & refunds. You're never charged past the period you paid for. Refunds and credits for unused balance follow the terms shown at checkout and applicable law — if you have a question, send it via feedback or report-an-error and a real person on the small team will sort it out with you.
Payment processor. Checkout and subscriptions run through Stripe, a reputable, widely used processor — card details go to Stripe, and Nexara never sees or stores your full card number.
🟠 Agent permissions — what can the AI actually do?
It depends on the surface, and the honest picture is: the web chat is sandboxed; the on-device tools are powerful by design.
In the web chat an agent sees your conversation and anything you attach, and it can create artifacts — files, HTML pages, images, games — that render inside the chat and that you download. It cannot reach your device's filesystem or run commands on your computer from the website. Web lookups happen only through an explicit, visible web-search step mid-conversation, and memory features save only what you choose to keep.
Connected services. Automations, browser actions, and integrations only work with accounts you explicitly connect, and they act under the permission you grant when connecting. Don't connect anything you wouldn't want an agent touching.
On your device the desktop and Android assistants and the IDE/CLI harness deliberately have more power: they can read your screen, open apps, toggle device settings, and in NexaraCode and CLI sessions read, write, and run commands inside the working folder you open. That's the point of an on-device agent — and those runs happen only in sessions you start yourself on your own machine.
The universal rule: never hand any agent a password, API key, or secret. Keep secrets in places only you control, and check what a session can reach before letting it loose on important files.
🟠 Vendor lock-in — could I leave without losing my work?
Short answer: yes, you could move tomorrow without losing your work. Nexara doesn't hold what you create hostage.
Everything the AI makes is ordinary, open stuff you already own: website and code artifacts download as files or a .zip, documents and slides download in their standard formats, images save to your device or Library, and 3D models export to .glb/.obj/.stl. Nothing is encrypted to your Nexara account or stored in a proprietary format that only Nexara can open.
Chats are plain text on your account: you can copy any message or artifact out at any time, and share links exist for anything you want to publish. There isn't a one-click "export every chat" button yet — treat Nexara like any tool and download what you care about. Account deletion is handled through in-app support under the Privacy Policy, and if you ever cancel you keep your account, chats, and anything you already downloaded.
🤔 What are model names like "GPT-5.6 Luna" or "Claude Fable 5"?
They're identifiers in our model catalog, not claims that OpenAI or Anthropic publish those exact retail names. The catalog follows the standard router convention — provider/name — the same format aggregators like OpenRouter use, and the prefix simply names the provider family a model belongs to.
Every model is served through the real provider pipeline: either a direct provider connection (Claude models through Anthropic, GPT models through OpenAI) or an open model router, depending on the model. No model is a custom stand-in we trained — you get the full, unmodified model behind each catalog entry, and the chat tells you exactly which model served your reply.
