Public info · last updated September 2026

Everything public about Nexara.

What Nexara is, where it runs, how Compute and the usage limits work, what each plan actually covers — and which plan fits which kind of usage. If you found Nexara through search and want the straight facts before trying it, this is the page.

What is Nexara

Nexara is an AI assistant and creation suite that runs everywhere you work: in the browser, as an Android app, as a Windows desktop app, and from a terminal. It is a real product with real accounts, real billing, and a real Compute currency — every feature below is something you can use today. asknexara.com is Nexara's official, current domain — the same Nexara service that was previously reachable at the nexara-ai-chat.vercel.app address during early development.

Chat with strong models

A full model catalog — frontier, reasoning, coding, speed and multimodal models — all billed from one Compute balance.

Create, don't just describe

Image Studio (GPT Image), Slides, Website and Presentation builder modes, PDFs, games, widgets, 3D models — real artifacts rendered in the chat.

OS-level tools on your devices

The mobile/desktop assistants can read your screen, open apps, toggle flashlight/brightness/media, take screenshots and control device settings end-to-end.

Android app

Distributed as a direct APK (not on the Play Store yet) — plus a Windows desktop build with automatic updates.

NexaraCode + CLI

An IDE desktop shell and a standalone CLI for agent-style coding sessions against the same account.

Builders & agents

Slides, websites and presentations from a bare prompt; Planner and autonomous /goal runs; user-made Nexar personas.

Compute & the usage limits

Compute is Nexara's currency — the same money as a dollar, just renamed. The exchange rate is exact and it never changes: 1 Compute = $0.000001, so 100K Compute = $0.10 and 1M Compute = $1.00, 100% — no hidden spread between the two. If a model's input costs the provider $1 per million tokens, it costs exactly 1,000,000 Compute here. Every model is billed from that pool at its own actual provider rate, and you can subscribe to a plan for a monthly allowance, top up a balance, or earn free Compute.

1M Compute = $1.00

Exactly, always. Compute is Nexara's name for the dollar at a fixed rate — 1M Compute and $1 are the same amount, and you only ever see Compute.

Daily allowance

Plan users get a daily Compute allowance plus a generous weekly hard cap — no monthly limit, no 5-hour window.

Weekly hard cap

A weekly ceiling bounds total burn so one runaway session can't exhaust a month in a day.

12,500 Compute per image

GPT Image generation costs a flat 12,500 Compute per image regardless of size.

What that actually buys you

  • One casual chat reply — usually a few hundred to a couple of thousand Compute, depending on the model and how long the answer is.
  • One Image Studio picture — exactly 12,500 Compute, any size, any engine.
  • One heavy coding turn (agent session, big artifact build) — tens of thousands of Compute and up; this is why the coding plans carry much larger allowances.
  • MiniMax M3 is 100% free and never touches your balance — useful when you want an answer without spending anything.
  • Exact per-model prices live on the Docs page — every rate is per 1M tokens, at the same fixed exchange rate as above.

How limits actually work

  • Free plan. No credit card. A small automatic allowance (about 5M Compute/day) refills daily — enough to try the product, but it will not survive heavy or long coding sessions.
  • Refills daily. Paid plans refresh their Compute allowance every day, with a generous weekly cap — your live numbers always show in Settings → Compute.
  • Hitting the limit. When the daily allowance or the weekly cap is reached, new requests are paused until the day refills or the week resets on Monday. Nexara refuses a request before it starts if your balance can't cover it — it never silently overcharges.
  • Top-ups & overage. You can top up Compute at any time; overage beyond an allowance is billed from the balance at 500K Compute per extra 1M tokens.
  • Billing multiplier. Token usage is counted with a flat 5% surcharge covering payment-processing and currency-conversion fees on API credit purchases. There are no other hidden fees.

Plans, prices & limits

Which plan is for you? The $5 Lite plan is for testing Nexara — expect usage limits to be hit quickly under heavy use. $10 Pro is the middle tier: how far it goes depends on what you're doing (normal chat lasts; big builds don't). The higher tiers are for when it makes sense to actually code with Nexara — agent sessions, the IDE, long automation — without hitting usage limits every few minutes.

Lite

$5/mo

Trying out Nexara

Daily spend cap
12.3M
Weekly spend cap
61.4M
Images per day
35
Library slots
100

The daily and weekly figures are the only limits on this plan — your Compute refills each day and is hard-capped each week. There is no monthly limit and no 5-hour window: the weekly cap refreshes every Monday.

Built for testing the product. Compute is real but limited — heavy or long sessions can hit the limit quickly. If you mainly chat casually, it goes further.

Compare on the pricing page

Pro

$10/mo

The mid tier

Daily spend cap
24.6M
Weekly spend cap
122.9M
Images per day
70
Library slots
200

The daily and weekly figures are the only limits on this plan — your Compute refills each day and is hard-capped each week. There is no monthly limit and no 5-hour window: the weekly cap refreshes every Monday.

Comfortable for everyday chat and light creation. How far it goes depends on what you're doing: normal conversation lasts, but big image runs or heavy coding will drain it noticeably faster.

Compare on the pricing page

Pro+

$20/mo

The sweet spot

Daily spend cap
49.1M
Weekly spend cap
245.7M
Images per day
175
Library slots
500

The daily and weekly figures are the only limits on this plan — your Compute refills each day and is hard-capped each week. There is no monthly limit and no 5-hour window: the weekly cap refreshes every Monday.

Recommended for regular creation plus light coding. 2× the Pro allowance at 2× the price.

Compare on the pricing page

Max

$50/mo

Building with Nexara

Daily spend cap
122.9M
Weekly spend cap
614.3M
Images per day
Unlimited
Library slots
1,000

The daily and weekly figures are the only limits on this plan — your Compute refills each day and is hard-capped each week. There is no monthly limit and no 5-hour window: the weekly cap refreshes every Monday.

For when you actually code with it — agent sessions, IDE use, big artifact builds — without watching the meter run dry.

Compare on the pricing page

Max (Ultra)

$100/mo

Heavy, all-day usage

Daily spend cap
491.4M
Weekly spend cap
2.5B
Images per day
Unlimited
Library slots
1,000

The daily and weekly figures are the only limits on this plan — your Compute refills each day and is hard-capped each week. There is no monthly limit and no 5-hour window: the weekly cap refreshes every Monday.

20× the Pro allowance. Long agent runs, automation, and serious development work that would otherwise hit usage limits quickly.

Compare on the pricing page

Billing details

  • Monthly. Billed every month at the list price.
  • Quarterly −10% / Yearly −20%. Longer commitments are charged upfront as one payment for the whole period at a discount.
  • Pro+ intro. New Pro+ subscribers get 25% off their first 3 monthly cycles ($15 instead of $20), then it renews at full price automatically.
  • Custom plans. Build your own: pick a daily Compute allowance and Nexar slots; the weekly cap is set alongside it (same ×20 basis standard plans use).
  • Gifting. Plans can be gifted, and upgrades extend rather than replace.

Images & the Library

Image Studio generates real pictures with OpenAI's GPT Image at a flat 12,500 Compute per image (any size) — that's $0.0125 (1.25¢) per image, roughly 80 images per dollar, with no quality or resolution tiers: one flat price, any size. Per-plan daily image caps and Library storage:

Plan priceImages per dayLibrary slots
Free450
$5 Lite35100
$10 Pro70200
$20 Pro+175500
$50 Max and upUnlimited1,000

Library slot limits on paid plans scale with tier ("1,000" on Max/Ultra); the Free plan has a fixed 50-slot library. Slots can be purchased on top of the plan allowance.

Models

Nexara's catalog covers frontier, reasoning, coding, speed and multimodal models from OpenAI, Anthropic, Google, Qwen, MiniMax and more. No model is plan-gated — every plan unlocks the full lineup; only the Compute allowance changes. What you get is always the provider's full, unmodified model — full context window, full output length, normal reasoning modes. Nexara never substitutes a secretly smaller or watered-down stand-in, and no plan hides a model behind a paywall. MiniMax M3 is 100% free and never touches your balance.

What the model names mean: catalog names like GPT-5.6 Luna or Claude Fable 5 are identifiers in our model catalog, using the same provider/name convention routers like OpenRouter use. The prefix identifies the provider family — models are served through a mix of direct provider connections (Claude models through Anthropic, GPT models through OpenAI) and open model routers, depending on the model. The chat always shows you which model actually served your reply.

Here is the complete per-model Compute reference — every model in the catalog, with its cost per million tokens for input, output, and cached input, plus image/video input support and context windows. 1M Compute = $1 — the dollar figure under each price is the internal provider rate. The Docs page has the same table with the full media-input breakdown.

Frontier

25 models
ModelCompute / 1M inCompute / 1M outCompute / 1M cacheImageVideoContext
GPT-OSS-120B
openai/gpt-oss-120b
100K$0.10400K$0.40131K
MiniMax M3Free
minimax/minimax-m3:free
FreeFree1M
Grok 4.5
x-ai/grok-4.5
1.74M$1.745.27M$5.27500K
Grok 4.6
x-ai/grok-4.6
2M$2.006M$6.00500K
GPT-5.6 Luna
openai/gpt-5.6-luna
200K$0.201.05M$1.051M
GPT-5.6 Terra
openai/gpt-5.6-terra
2.05M$2.0511.75M$11.751M
GPT-5.6 Sol
openai/gpt-5.6-sol
5M$5.0030M$30.00500K$0.501M
Claude Sonnet 4.6
anthropic/claude-sonnet-4.6
3M$3.0015M$15.00300K$0.301M
Claude Sonnet 5
anthropic/claude-sonnet-5
2M$2.0010M$10.00200K$0.201M
Claude Opus 4.6
anthropic/claude-opus-4.6
5M$5.0025M$25.00500K$0.501M
Claude Opus 4.7
anthropic/claude-opus-4.7
5M$5.0025M$25.00500K$0.501M
Claude Opus 4.8
anthropic/claude-opus-4.8
5M$5.0025M$25.00500K$0.501M
Claude Opus 5
anthropic/claude-opus-5
5M$5.0025M$25.00500K$0.501M
Claude Fable 5
anthropic/claude-fable-5
10M$10.0050M$50.001M$1.001M
Kimi K3
moonshotai/kimi-k3
3M$3.0015M$15.00300K$0.30262K
Kimi K2.6Locked
moonshotai/kimi-k2.6
950K$0.954M$4.00262K
Qwen 3.8 Max
qwen/qwen3.8-max
2.55M$2.557.55M$7.551M
Qwen 3.7 Max
qwen/qwen3.7-max
2.55M$2.557.55M$7.551M
Qwen 3.6 Max (Preview)
qwen/qwen3.6-max-preview
1.35M$1.357.85M$7.85256K
Qwen 3.5 397B A17B
qwen/qwen3.5-397b-a17b
650K$0.653.65M$3.65256K
Qwen3 Max
qwen/qwen3-max
1.25M$1.256.05M$6.05256K
GLM 5
z-ai/glm-5
630K$0.631.95M$1.95200K
GLM 5.1
z-ai/glm-5.1
1.42M$1.424.42M$4.42200K
GLM 5.2
z-ai/glm-5.2
1.5M$1.504.52M$4.521M
GLM 5.3
z-ai/glm-5.3
1.4M$1.404.4M$4.40260K$0.261M

Reasoning

15 models
ModelCompute / 1M inCompute / 1M outCompute / 1M cacheImageVideoContext
Step 3.7 Flash
stepfun/step-3.7-flash
100K$0.10300K$0.30256K
Nemotron 3 Nano 30B A3B
nvidia/nemotron-3-nano-30b-a3b
50K$0.05150K$0.15256K
DeepSeek V4 Flash 07.31
deepseek/deepseek-v4-flash-0731
40K$0.04130K$0.139K$0.011M
Xiaomi Mimo V2.5 ProLocked
xiaomi/mimo-v2.5-pro:free
430K$0.43870K$0.87128K
Xiaomi Mimo V2.5Locked
xiaomi/mimo-v2.5:free
430K$0.43870K$0.87128K
Kimi K2.5Locked
moonshotai/kimi-k2.5
570K$0.572.85M$2.85262K
Nemotron 3 Nano
nvidia/nemotron-3-nano
100K$0.10400K$0.401M
Nemotron 3 Super
nvidia/nemotron-3-super
250K$0.251M$1.001M
Nemotron 3 Ultra
nvidia/nemotron-3-ultra
500K$0.502M$2.001M
Qwen 3.6 27B
qwen/qwen3.6-27b
650K$0.653.65M$3.65256K
Qwen 3.6 35B A3B
qwen/qwen3.6-35b-a3b
298K$0.301.54M$1.53256K
GLM 4.7
z-ai/glm-4.7
640K$0.642.24M$2.24200K
DeepSeek V3.2
deepseek/deepseek-v3.2
280K$0.28420K$0.42131K
DeepSeek V4 Flash
deepseek/deepseek-v4-flash
250K$0.251M$1.001M
DeepSeek V4 Pro
deepseek/deepseek-v4-pro
600K$0.602.4M$2.401M

General

20 models
ModelCompute / 1M inCompute / 1M outCompute / 1M cacheImageVideoContext
MiniMax M2
minimax/minimax-m2:free
300K$0.301.2M$1.20205K
MiniMax M2.1
minimax/minimax-m2.1:free
300K$0.301.2M$1.20205K
MiniMax M2.5
minimax/minimax-m2.5:free
300K$0.301.2M$1.20205K
MiniMax M2.7
minimax/minimax-m2.7:free
300K$0.301.2M$1.20205K
Ministral 14B
mistralai/ministral-14b
200K$0.20200K$0.20256K
Mistral Small 26.03
mistralai/mistral-small-2603
100K$0.10300K$0.30256K
Mistral Medium 3.5
mistralai/mistral-medium-3.5
350K$0.351.05M$1.05256K
Mistral Large 3
mistralai/mistral-large-2512
2M$2.006M$6.00256K
Ling 3.0 Flash
inclusion-ai/ling-3.0-flash
50K$0.05150K$0.15260K
Llama 3.3 70B Instruct
meta/llama-3.3-70b-instruct
150K$0.15600K$0.60131K
Llama 3.1 8B Instruct
meta/llama-3.1-8b-instruct
40K$0.04100K$0.1016K
Qwen 3.7 Plus
qwen/qwen3.7-plus
450K$0.451.65M$1.651M
Qwen 3.6 Plus
qwen/qwen3.6-plus
550K$0.553.05M$3.051M
Qwen 3.5 Plus
qwen/qwen3.5-plus
450K$0.452.45M$2.451M
Qwen Plus 07.28
qwen/qwen-plus-2025-07-28
450K$0.451.25M$1.251M
GLM 4.5
z-ai/glm-4.5
630K$0.632.23M$2.23131K
DeepSeek Chat V3.1
deepseek/deepseek-chat-v3.1
280K$0.28420K$0.42164K
Llama 3.3 Nemotron Super 49B
nvidia/llama-3.3-nemotron-super-49b
150K$0.15600K$0.60131K
SenseNova 6.7 Flash-Lite
sensenova/sensenova-6.7-flash-lite
20K$0.0280K$0.08262K
SenseNova 6.8 Flash-Lite
sensenova/sensenova-6.8-flash-lite
20K$0.0280K$0.08262K

Coding

8 models
ModelCompute / 1M inCompute / 1M outCompute / 1M cacheImageVideoContext
Devstral Medium
mistralai/devstral-medium
500K$0.50650K$0.65256K
Codestral 25.08
mistralai/codestral-2508
300K$0.30900K$0.90256K
Laguna XS.2
poolside/laguna-xs.2
100K$0.10400K$0.40131K
Grok Build 0.1
x-ai/grok-build-0.1
1M$1.002M$2.00200K$0.20256K
GPT-5.3 Codex Spark
openai/gpt-5.3-codex-spark
150K$0.15600K$0.60128K
Kimi K2.7 CodeLocked
moonshotai/kimi-k2.7-code
700K$0.703.5M$3.50150K$0.15262K
Qwen3 Coder Plus
qwen/qwen3-coder-plus
1.05M$1.055.05M$5.051M
GLM 4.6
z-ai/glm-4.6
620K$0.622.22M$2.22200K

Speed

13 models
ModelCompute / 1M inCompute / 1M outCompute / 1M cacheImageVideoContext
MiniMax M2.1 High-Speed
minimax/minimax-m2.1-highspeed:free
150K$0.15600K$0.60205K
MiniMax M2.5 High-Speed
minimax/minimax-m2.5-highspeed:free
150K$0.15600K$0.60205K
MiniMax M2.7 High-Speed
minimax/minimax-m2.7-highspeed:free
150K$0.15600K$0.60205K
Ministral 3B
mistralai/ministral-3b
50K$0.0550K$0.05128K
Ministral 8B
mistralai/ministral-8b
100K$0.10100K$0.10256K
Nemotron Nano 9B V2
nvidia/nemotron-nano-9b-v2
30K$0.03100K$0.10128K
Llama 3.2 1B Instruct
meta/llama-3.2-1b-instruct
20K$0.0250K$0.0516K
Llama 3.2 3B Instruct
meta/llama-3.2-3b-instruct
30K$0.0380K$0.0816K
Claude Haiku 4.5
anthropic/claude-haiku-4.5
1M$1.005M$5.00100K$0.10200K
Qwen 3.5 Flash
qwen/qwen3.5-flash
150K$0.15450K$0.451M
GLM 4.5 Air
z-ai/glm-4.5-air
220K$0.221.12M$1.12131K
GLM 5 Turbo
z-ai/glm-5-turbo
1.24M$1.244.04M$4.04200K
GLM 5.3 Flash
z-ai/glm-5.3-flash
150K$0.15500K$0.5030K$0.031M

Multimodal

13 models
ModelCompute / 1M inCompute / 1M outCompute / 1M cacheImageVideoContext
Gemini 3.6 FlashLocked
google/gemini-3.6-flash
1.5M$1.507.5M$7.501M
Gemini 3.7 FlashLocked
google/gemini-3.7-flash
380K$0.381.88M$1.8840K$0.041M
Gemini 3.5 FlashLocked
google/gemini-3.5-flash
1.5M$1.509M$9.001M
Gemini 3.1 ProLocked
google/gemini-3.1-pro
2M$2.0012M$12.001M
Gemini 3 FlashLocked
google/gemini-3-flash
500K$0.503M$3.001M
Gemini 2.5 FlashLocked
google/gemini-2.5-flash
300K$0.302.5M$2.501M
Gemini 2.5 ProLocked
google/gemini-2.5-pro
1.25M$1.2510M$10.001M
Qwen 3.5 Omni Plus
qwen/qwen3.5-omni-plus
1.45M$1.4511.05M$11.05128K
Qwen 3.5 Omni Flash
qwen/qwen3.5-omni-flash
450K$0.453.05M$3.05128K
Qwen3 VL Plus
qwen/qwen3-vl-plus
250K$0.251.65M$1.65256K
Qwen3 Omni Flash
qwen/qwen3-omni-flash
480K$0.481.71M$1.71128K
Ox Alpha
stealth/ox-alpha-free
500K$0.501.25M$1.251M
Nemotron 3 Nano Omni
nvidia/nemotron-3-nano-omni
50K$0.05150K$0.15256K

Prices are Compute per million tokens at each model's actual provider rate. The cache column only appears when the model's provider advertises a separate cache-read rate — otherwise cached input bills at the full input rate. Context is the model's maximum window in tokens.

How requests reach the models. Nexara mostly routes model calls through third-party model gateways (routers) — OpenRouter is the main one — and for some models it calls the official provider's own API directly instead. A router is just the delivery path: it forwards your request to the real provider (OpenAI, Anthropic, Google, and so on) and relays the answer back. Routing never replaces, shrinks, or swaps the model you picked — the model that answers is the model named in your conversation, whichever gateway carried the request. If a gateway is down, a request may fall over to the next route in that model's chain, and the chat tells you when that happens.

How a message travels. When you send a message, it goes to a model router (or straight to the provider's own API for models that aren't on a router). The router hands the request to the provider, the provider generates the response, and the response streams back to you — typically in a few seconds. Nexara is the client and orchestrator in that loop, like Cursor is for coding: we connect you to the models, we don't stand in for them.
What we keep. Nexara itself retains no conversation data beyond what your account needs — we don't warehouse prompts or train on them. Your chat history is stored in Supabase, tied to your account, so it syncs across web, Android, Windows, and CLI and is deletable whenever you want. What actually reaches a model provider is the minimum context needed to answer your current request (see the Privacy Policy for exactly who sees what).

Your chats, settings, personas and preferences follow your account across the web, Android, Windows and the CLI. Guests can try the chat without an account; Compute, the Library, and billing features require signing in.

Where Nexara runs

Web app

The full experience in the browser — chat, builders, image studio, library, settings.

Android (APK)

Direct APK install with OS-level assistant tools. Not on the Play Store yet.

Windows desktop

Electron shell with automatic updates — new versions appear in-app when a release is published.

NexaraCode (IDE)

A desktop IDE shell for agent-driven coding sessions.

Nexara CLI

A standalone command-line tool sharing the same account and models.

Agent Visualization SDK

Connect your own front-end to local agents over a small WebSocket protocol.

All downloads

Security & routing, plainly

Two things people ask us about are how their data is stored and which AI providers actually answer their messages. Here is exactly how both work — no fine print.

Your data in the database

  • Chats live in Supabase — a managed Postgres database — tied to your account and synced across web, Android, Windows, and the CLI.
  • Row Level Security is on. Every thread and message row is scoped to its owner (auth.uid()); the database only ever returns rows that belong to the signed-in user. A user cannot read or write another user's chats by guessing IDs.
  • The only exceptions are opt-in: threads you explicitly mark public (share links) and threads where you invited a collaborator. Nothing else opens.
  • The rules ship with the code. The RLS policies live in the web app's version-controlled database migrations — anyone can read them at github.com/K1NGMR/nexara-ai-chat under supabase/.

Which provider answers you

  • Your message goes to a real provider. The chat calls the model you picked — either straight to that provider's own API (Claude through Anthropic, GPT through OpenAI) or through an open model router.
  • Routers are delivery layers. A router like OpenRouter aggregates many providers behind one API; it forwards your request to the actual model host and relays the answer back. It never swaps, shrinks, or rewrites the model — the model named in your chat is the model that answers.
  • The model shown is the model used. The chat always labels which model served each reply, and if a gateway is down it tells you when a request falls over to another route.
  • The full catalog is on the Docs page, with per-model pricing and reasoning options. Any model on any plan — nothing is gated or hidden.

Why there's no routing map

We can't give you a fixed map of which provider serves each model — we adjust routing constantly, sometimes daily. Even inside OpenRouter the providers behind a model change: a cheaper host for the same model can appear at any moment, and when it does we move the route. We can't honestly guarantee a 24/7 routing map, because the whole point of Nexara is to give you the cheapest possible API access — and that isn't achievable if we lock ourselves to a single provider inside a router.

What never changes is the contract you see: the model is always the model you picked, and the chat always tells you which one served the reply. Only the behind-the-scenes delivery path moves — always toward the same model at the best real-world price we can find.

Why the backend isn't open-source

Secrets are only half the story. Even with keys tucked safely in environment variables, publishing the backend source would hand anyone the full code to hunt for logic flaws — auth bypasses, rate-limit gaps, IDOR bugs — and then probe our live deployment with exactly what they found. Public frontend code is normal; public backend code is free vulnerability research against us.

So here's the honest balance we strike: we were planning to open-source parts of the backend, but we won't publish the server code itself. What you can always inspect is everything that governs your data and your money — the database security rules, the Compute and pricing logic shown openly on this page, and the entire CLI source. Our security model is public even when our code isn't.

What actually leaves your device for a given request is the minimum context needed to answer it — spelled out further in the Privacy Policy.

The team behind Nexara

Nexara is built by a small group of independent developers — between two and five of us, depending on the week. There is no big company and no investor money behind it: every model integration, app, builder, and bug fix you see here is designed, coded, and shipped by that handful of people.

We also plan to keep AI usage as cheap as we possibly can. Every model is billed at its actual provider rate, we price cached tokens at the cache rate, and we run regular optimization passes to cut what each turn really costs — so the Compute you buy goes as far as possible.

Where we're headed: Nexara is a very new project — the website went live on August 3rd, 2026 — and we're planning to keep it going long term, adding more tools as we go. We're working toward running our own local servers so we don't rely on routers as we head into 2027, and our dev team hopes to expand the project even further. Reach us at support@asknexara.com, and a Discord server is in the works (about four weeks out — the bots are taking more time than we expected). Thanks for being here this early. — Nexara Team 🙂

Who actually runs it? Nexara is run by a small team of developers — and we keep our individual identities private, by choice, for our own safety and peace of mind. There's no named CEO and no corporate brand to point at: the people who code Nexara are the people who run it.

Email support@asknexara.com or use the in-app feedback or report an error options — both route straight to the small, trusted team that fixes issues, never shown publicly.

Current goal

Nexara Cowork gets its own basic computer — a dedicated machine so the agent can run longer, heavier jobs without competing with your device.

Support this goal

Every paid plan gives $1 to people who are blind

With every paid plan billing — yours and everyone else's — we donate $1 to support people who are blind or have low vision. It's a commitment we make from the subscription, never an extra charge on top.

But please don't buy a plan purely because it feels like a donation — that's not what plans are for. Buy one because you genuinely need a service with a large lineup of AI models: chat LLMs, image generation, and soon video models too. If you simply want to help, the donate page is the honest way to do it.

$1 from each paid billing

Any donation helps a lot

A small team means every donation goes a long way — it covers real model inference and server costs, and lets us keep a generous free tier for people who can't pay. If Nexara has been useful to you, even a few dollars makes a real difference.

Donate

Frequently asked

🔒 Privacy — what happens to my chats and files?

Your conversations and files belong to you. We don't read them, we don't sell them, and we don't train on them. Chats are saved to your account in Supabase so you can pick them up on any device, and memory features only ever keep what you explicitly choose.

The people running Nexara don't browse conversations. What reaches us during normal use is technical usage only — typically a webhook notice like "user just used GPT-5.6 Sol" so we can watch costs and reliability. No message text or file content rides along with those notices, and no other data about your activity is saved.

Your name and email only ever leave your account if you send us feedback or report an error — and even then they are never shown publicly. They stay with the small, trusted team that fixes the issue, only for as long as the fix needs them.

🔴 What should I never paste into a chat?

The short rule is the same as with ChatGPT, Claude, or Gemini: if it's a secret, don't paste it. Passwords and API keys, private documents, personal information (ID numbers, addresses), proprietary code, and confidential business information should stay out of any AI chat — including Nexara's. To answer you, the relevant parts of a conversation travel to the model provider for that request (the Privacy Policy spells out who sees what); a secret pasted into a chat is a secret that has left your control.

Need help with sensitive code or documents? Paste a redacted version with the real secrets replaced by placeholders — that's enough for almost every question.

💰 Compute — how much usage do I actually get?

Compute is Nexara's currency: 1M Compute = $1.00, and every model bills from that pool at its own provider rate. The live number you actually have is always in Settings → Compute — daily allowance, weekly cap, and balance.

To make it concrete (rough figures — they vary with model and message length): a short casual reply usually costs a few hundred to a couple of thousand Compute; a long coding turn that builds something big can run tens of thousands; one Image Studio picture is exactly 12,500. So a few million Compute is hundreds of everyday messages — or roughly ten images plus casual chatting — and it goes fast under heavy agent/IDE sessions.

The free plan refills about 5M/day. The daily allowance and weekly cap for every paid plan are in the plan cards above.

🤖 Model access — full models or restricted versions?

Full models — the real thing. Every plan can use the entire catalog, and what you get is the provider's actual model with its full context window and reasoning modes. Nexara never serves secretly smaller, quantized, or watered-down stand-ins; names like "Lite" or "Flash" in the catalog are the providers' own model names, not downgrades we apply.

Plans don't restrict which models you can pick — they only change how much Compute you have to spend. The free tier gets the same full models with a small daily allowance. If a provider model is briefly down, a request may fall back to another capable model, and the chat always tells you when that happens.

⚡ Speed & reliability — does it stay usable when lots of people are online?

Mostly — but we'll be straight about it: Nexara is run by a small team, not a company with giant infrastructure budgets. When a lot of people are online at once, replies can slow down, and there are real limits (the daily allowances and weekly caps) that keep the service from being swamped. We scale what we can within our budget and watch load continuously.

In normal use it stays usable. If we ever hit a rough patch, you'll see it acknowledged honestly in the Updates log rather than hidden.

🛠️ Features — are the coding and agent tools actually good?

They're real and people use them every day — file-building artifacts with live previews, the agent loop, websites and slides, NexaraCode, the CLI, and the desktop and Android assistants. We wouldn't claim we're the best at everything; bigger tools with bigger teams exist. What we can promise is that we genuinely try to make ours good and keep improving them.

The IDE and NexaraCode are still in Beta — you will hit rough edges. Try everything yourself, and if something is weak, tell us exactly what was weak: feedback is the roadmap.

📈 Longevity — will Nexara still be around in a year?

That's the plan. We intend to keep Nexara running for a very long time. Realistically, staying up long-term depends on Nexara getting enough support (plans, donations) and feedback to cover the model bills and keep the team shipping.

As long as that support keeps coming, we plan to keep going — and if things ever change, you'll hear it from us clearly and in advance, never find out after the fact.

📊 Rate limits — every Compute number, for every plan

All limits below are in Compute — the same money as dollars (1M Compute = $1). Paid plans refill a daily allowance with a hard weekly cap — that is the whole limit system: no monthly limit, no rolling 5-hour window. One image always costs exactly 12,500. The live numbers always win and live in your Settings → Compute.

PlanDaily allowanceWeekly cap
Free5M
Lite · $5/mo12.3M61.4M
Pro · $10/mo24.6M122.9M
Pro+ · $20/mo49.1M245.7M
Max · $50/mo122.9M614.3M
Max (Ultra) · $100/mo491.4M2.5B

These are the only limits: the daily allowance refills every day, and the weekly cap resets every Monday. There is no monthly limit and no 5-hour window.

🔴 Billing & refunds — renewals, cancelling, unused Compute

Automatic renewal? Yes — paid plans are subscriptions billed automatically on the interval you chose at checkout. Monthly renews monthly; quarterly (−10%) and yearly (−20%) are charged upfront as one payment and renew at the end of their period.

Cancelling is easy. Turn off future renewals any time from your billing controls — no contracts and no cancellation fees. The subscription simply stops at the end of the paid period you already bought, and your account drops to the free tier: chats, files, Library and personas stay exactly where they are.

Unused Compute & refunds. You're never charged past the period you paid for. Refunds and credits for unused balance follow the terms shown at checkout and applicable law — if you have a question, send it via feedback or report-an-error and a real person on the small team will sort it out with you.

Payment processor. Checkout and subscriptions run through Stripe, a reputable, widely used processor — card details go to Stripe, and Nexara never sees or stores your full card number.

🟠 Agent permissions — what can the AI actually do?

It depends on the surface, and the honest picture is: the web chat is sandboxed; the on-device tools are powerful by design.

In the web chat an agent sees your conversation and anything you attach, and it can create artifacts — files, HTML pages, images, games — that render inside the chat and that you download. It cannot reach your device's filesystem or run commands on your computer from the website. Web lookups happen only through an explicit, visible web-search step mid-conversation, and memory features save only what you choose to keep.

Connected services. Automations, browser actions, and integrations only work with accounts you explicitly connect, and they act under the permission you grant when connecting. Don't connect anything you wouldn't want an agent touching.

On your device the desktop and Android assistants and the IDE/CLI harness deliberately have more power: they can read your screen, open apps, toggle device settings, and in NexaraCode and CLI sessions read, write, and run commands inside the working folder you open. That's the point of an on-device agent — and those runs happen only in sessions you start yourself on your own machine.

The universal rule: never hand any agent a password, API key, or secret. Keep secrets in places only you control, and check what a session can reach before letting it loose on important files.

🟠 Vendor lock-in — could I leave without losing my work?

Short answer: yes, you could move tomorrow without losing your work. Nexara doesn't hold what you create hostage.

Everything the AI makes is ordinary, open stuff you already own: website and code artifacts download as files or a .zip, documents and slides download in their standard formats, images save to your device or Library, and 3D models export to .glb/.obj/.stl. Nothing is encrypted to your Nexara account or stored in a proprietary format that only Nexara can open.

Chats are plain text on your account: you can copy any message or artifact out at any time, and share links exist for anything you want to publish. There isn't a one-click "export every chat" button yet — treat Nexara like any tool and download what you care about. Account deletion is handled through in-app support under the Privacy Policy, and if you ever cancel you keep your account, chats, and anything you already downloaded.

🤔 What are model names like "GPT-5.6 Luna" or "Claude Fable 5"?

They're identifiers in our model catalog, not claims that OpenAI or Anthropic publish those exact retail names. The catalog follows the standard router convention — provider/name — the same format aggregators like OpenRouter use, and the prefix simply names the provider family a model belongs to.

Every model is served through the real provider pipeline: either a direct provider connection (Claude models through Anthropic, GPT models through OpenAI) or an open model router, depending on the model. No model is a custom stand-in we trained — you get the full, unmodified model behind each catalog entry, and the chat tells you exactly which model served your reply.

🌐 Which provider actually serves each model?

Claude models are served through Anthropic, GPT models through OpenAI, and the rest of the catalog through open model routers such as OpenRouter — the exact serving model is always shown in the chat. When a model shows as "Locked", it means the upstream provider is having an outage, not that Nexara removed it.

🖥️ Is computer control (NexaraClaw) safe to use?

Destructive actions — deleting, overwriting, sending, anything that can't be undone — require your explicit confirmation, and are blocked entirely in Plan mode. Like any automation tool, treat it as an agent with access: start with things you can afford to lose, review what it proposes to do, and don't leave it unattended while it works on important systems or accounts.

🗑️ What happens to my data once I delete my account?

You can delete conversations from within Nexara, and account deletion is processed on request through the in-app support channel — chats, files and personal data are removed as described in section 6 of the Privacy Policy.

What is one unit of Compute worth?

$0.000001 — so 100,000 Compute is $0.10 and 1,000,000 Compute is $1.00. Every model price is expressed at that fixed rate.

Why do different models cost different amounts of Compute?

Each model bills at its actual provider rate. A small flash model costs a fraction of a frontier flagship because that's what the provider charges Nexara.

Why did I hit my limit even though the meter looked low?

Paid plans refill a daily Compute allowance with a weekly hard cap. Usage recorded earlier in the week still counts against that cap, so the weekly number is what actually stops you once it's spent — Settings → Compute always shows both.

Is the $5 plan enough?

It's for testing Nexara. Casual chat works, but usage limits can be hit quickly — heavy or long sessions will run into the cap. Upgrade anytime; no contract.

What happens when I run out of Compute?

New requests pause until the next daily refill or Monday's week reset, or you top up balance. Nexara checks the balance before starting a request and refuses up front if it can't be covered.