Support
Frequently asked questions
Browse common answers or ask the assistant above for a more specific path.
NanoGPT uses account funds for model usage. Add funds from Balance with credit card or supported crypto, then use chat, media generation, API, or other tools.
You can start with an anonymous local session. Creating an account is optional, but it makes balances easier to keep across devices.
Open a new chat, choose a text model or Auto Model, and type your message in the input box. Conversations are separate by default, although Context Memory and Global Memory can carry selected context forward when enabled.
No. NanoGPT can be used with an anonymous local session. Accounts are useful when you want a balance that is easier to keep across devices.
Conversations are stored locally in your browser unless you explicitly export or share them.
NanoGPT can be installed as a Progressive Web App. The NanoGPT Browser Assistant is available for Chrome and Firefox, letting you attach pages, elements, images, or marked areas and then summarize, explain, translate, or chat with that context while you browse.
A model is the AI system that answers, reasons, writes code, analyzes images, or generates media. NanoGPT lets you choose between many text, image, video, and audio models.
Different models have different strengths, speeds, context sizes, prices, and media capabilities.
Some models are strongest at reasoning, some at coding, some at long context, some at vision, and some at low-cost high-volume work. Price usually reflects capability, provider cost, and performance.
Model pages show capabilities and pricing. Leaderboards help compare model quality across text, image, and video tasks.
Auto Model lets NanoGPT select a suitable Basic, Standard, or Premium model tier for the task. It is useful when you care more about the result than manually choosing a specific model.
Yes, NanoGPT supports prompt caching on most models. It can speed up repeated long-context requests and reduce costs when the selected model supports caching.
Vision means a text model can accept images or screenshots alongside text. Vision models can describe images, read text, analyze screenshots, and reason about visual content.
NanoGPT connects supported models to external AI model providers. Availability and routing depend on the selected model and may change over time.
Check the model page or provider-selection controls for the options available for that model.
Ask the Help assistant with the task and required inputs. It can compare published model capabilities and context limits, then point to suitable model detail pages. Use those pages to compare current pricing and availability. If you prefer an automatic choice, select Auto Model and choose the Basic, Standard, or Premium tier that fits the task.
For manual selection, check the capabilities that matter: coding agents often need tool calling and enough context for repository files; PDFs and screenshots need the corresponding PDF or vision support; and roleplay or translation requests should name the desired language, tone, and content requirements.
Model availability and capabilities change, so verify the recommendation on the current model detail page before relying on it.
Not every model exposes reasoning text, and some third-party clients ignore separate reasoning fields even when the model returns them. The normal API uses the `reasoning` field, while older clients may expect `reasoning_content`.
If you can change the request, set `reasoning.delta_field` to `reasoning_content`, or use the `reasoning_delta_field` or `reasoning_content_compat` shorthand. If you cannot change the request payload, use `https://nano-gpt.com/api/v1legacy/chat/completions` for the legacy field name.
Use `https://nano-gpt.com/api/v1thinking/chat/completions` when a client ignores reasoning-specific fields and should receive reasoning in the normal content stream. This is the preferred reasoning-compatible endpoint for JanitorAI.
Use the Media page in image mode. Enter a prompt, choose an image model, adjust settings such as aspect ratio or resolution when available, and start generation.
The help assistant can also recommend an image model for a specific prompt and link you directly to the right Media page setup.
Yes. Some image models support image-to-image or editing workflows. Choose a compatible model from the image generation flow and upload the source image where the UI offers image input.
As between you and NanoGPT, and to the extent permitted by applicable law, NanoGPT assigns its rights in generated output to you. Generated images may be used commercially, subject to applicable law and any restrictions in the selected model provider's terms.
You are responsible for checking third-party rights and provider-specific restrictions. Generated output is not guaranteed to be copyrightable, exclusive, or free of third-party rights.
Use the Media page in video mode. Choose a video model, enter the prompt, and adjust model-specific settings such as duration, aspect ratio, resolution, or image input when supported.
Some video models support image-to-video workflows for animating an existing image.
Image and video settings depend on the selected model. The Media page shows available settings such as aspect ratio, resolution, duration, rendering speed, and model-specific options.
NanoGPT has Context Memory and Global Memory features for carrying useful context forward. Memory can be controlled through conversation and system prompt settings.
Use the memory documentation for the exact setup and behavior.
Conversations are stored locally on your device and can be accessed from the Conversations page. A conversation can be continued later from the same browser/device.
Conversation history stays in the browser on the device where it was created by default. Open Conversations on that device to find and continue old chats.
Optional cloud sync can keep selected user data and conversations available across signed-in devices. The Storage tab in Settings controls NanoGPT-hosted sync or your own remote storage.
Unsynced local chats are not automatically recoverable in another browser. A synced chat can be restored after sign-in; an exported backup must be imported manually.
The Workspace groups related chats, files, notes, and tasks into projects. Create or open a project from Workspace, then start project-linked conversations so the project context and selected files stay organized together.
Projects are separate from ordinary chat history. Use Workspace for ongoing work with shared project context, and Conversations for the full list of individual chats.
Attach supported files from the chat input or add reusable files to a Workspace project. The selected model must support the file or document workflow; PDF, image, and other capabilities vary by model.
If a model cannot read an attachment, try a model whose detail page lists the required capability, reduce an oversized file, or add the extracted text directly. Project files can be reused by project-linked conversations.
Conversations can be exported from the Conversations page. If you use NanoGPT without an account, your local session identifier is what ties the browser session to your balance.
Creating an account is the easiest way to keep balances across devices.
NanoGPT is designed so you can use the site without creating an account. Conversations are stored locally by default.
When you send a message to a model, the prompt and relevant conversation context are sent to the selected model provider so the model can answer. Provider-specific privacy and retention policies may apply.
NanoGPT does not attach IP addresses to prompts, model-provider requests, usage records, support tickets, or bug reports. Raw IP addresses are used temporarily in rate-limit and abuse-prevention systems and are automatically deleted after the relevant security window expires.
Retention varies by control, and NanoGPT does not publish individual thresholds because doing so could weaken those protections. Password-reset abuse protection may also store a keyed one-way identifier derived from an IP address as a separate security record.
NanoGPT does not intentionally write raw IP addresses into application log messages. Vercel separately processes network information as the hosting provider, and its public privacy policy does not give one universal IP-specific retention period that NanoGPT can promise on Vercel’s behalf.
Web search is separate from model inference. The selected search provider handles the search query. You can change web search settings when the feature is available for your workflow.
TEE model protections apply to model execution and do not automatically extend to web-search provider requests.
Open Balance and choose a payment method. NanoGPT supports credit card payments and supported crypto deposits.
Crypto deposits are credited based on the current USD exchange rate. Nano deposits are converted to USD at the live rate, and new Nano deposits are no longer held as a Nano balance.
NanoGPT is pay-as-you-go by default. Text models are generally priced by input and output tokens. Media models are priced by model-specific generation settings such as resolution, duration, or rendering mode.
Pricing pages and model detail pages show the relevant costs.
The Pro plan currently advertises 60 million included input-token units per week and 100 included images per day. Most included text models consume one allowance unit per input token, while models marked with a 2x multiplier consume two. These figures can change, so check the live Subscription page before purchasing or planning usage. The weekly allowance resets Monday at 00:00 UTC; purchasing, renewing, resuming, or being billed does not reset it.
The included model list can change. Use the Subscription page for the current text and image model lists. Video generation, voice, web search, extended memory, TEE variants, and other paid extras are not included unless the live Subscription page explicitly says otherwise.
If an included limit is reached, included usage pauses until reset unless “Use balance after limits” is enabled, in which case eligible overage can use the normal pay-as-you-go balance.
For API overage, the API key must also allow balance spending: its billing mode must not be “Subscription only” and it must not have a $0 spend limit. Otherwise the request continues to return 429 until the included limit resets.
Yes. Website and API requests share the same subscription allowance. Use the subscription endpoint and an included model when you want the request covered by the subscription.
Use `GET /api/subscription/v1/models?detailed=true` for the current subscription-included text models and `POST /api/subscription/v1/chat/completions` for subscription-covered chat requests. Explicit provider selection is always pay-as-you-go and does not count as subscription-covered usage.
No. A subscription is for one person and is not a pooled team allowance. It cannot be shared, split, resold, or used to serve multiple users. Each person who wants subscription-covered usage needs their own account and subscription.
For shared or commercial team workloads, use pay-as-you-go access and the team controls instead of a personal subscription.
NanoGPT does not currently offer a separate pause action. To stop the next renewal while keeping access through the paid period, schedule a cancellation instead.
Open Subscription and choose Cancel subscription. Stripe subscriptions are managed through the Stripe Customer Portal; balance-funded subscriptions can be stopped from renewing directly on NanoGPT.
After cancellation is scheduled, access remains active until the displayed cancellation or current-period end date. If a Resume option is available before that date, use it to restore renewal. Canceling does not reset the weekly included-input quota.
Only the payment methods currently shown on the Balance or Subscription checkout are available. Availability can vary by region and payment method. Reopen the relevant page, confirm you are using the same NanoGPT session or account that made the payment, and check whether the card or crypto payment is still pending.
If a completed payment is not reflected, do not pay again immediately. Open a private support ticket from the affected session and include the payment method, approximate time, amount, receipt or transaction reference, and the Support Key shown by NanoGPT. Do not post full card details or private wallet credentials.
Yes. If your organization needs a written offer, quote, or pro-forma document for funding approval, grants, procurement, or internal purchasing, contact support with the billing entity, required reference numbers, requested credit amount, currency, and any deadline.
These documents are usually prepared for prepaid NanoGPT credits. Credits on NanoGPT are denominated in USD, so an offer can state the funding amount in another currency and clarify that the credited balance is converted at the exchange rate used at the time of payment or deposit.
Final VAT or tax details are stated on the invoice where applicable.
Yes. NanoGPT has a testimonials page with user reviews and links to original sources where available.
Yes. NanoGPT provides OpenAI-compatible API endpoints. Create an API key from the API page and use it with the documented base URL and model IDs.
NanoGPT works with many OpenAI-compatible clients. Create an API key, follow the dedicated integration guide when one exists, and use the exact base URL or endpoint required by that client.
The integrations documentation includes setup guides for JanitorAI, SillyTavern, OpenCode, OpenWebUI, Cursor, Cline, Codex CLI, Claude Code, and other tools. A third-party client can still have its own limitations, so first verify the API key, endpoint, model ID, subscription-versus-pay-as-you-go mode, and the client’s reasoning-field support.
Start with the HTTP status and machine-readable error code. Fix credentials for 401, add funds or disable paid extras for 402, check model permissions for 403, verify the model or resource ID for 404, and respect `Retry-After` for 429.
Timeouts and temporary server errors such as 408, 500, 503, or 504 can usually be retried with exponential backoff. A content-policy error is not a transient network failure and should not be retried unchanged. When contacting support, include the `X-Request-ID` response header when available.
The Usage page shows account activity, costs, model usage, subscription usage, and media-generation records. Use its time range and model or API-key filters to narrow the results, and open a row for the available request or generation details.
Request/response body logging is currently available for new `/v1/chat/completions` requests made with an API key after logging is enabled for that key. The setting does not retroactively create logs and does not capture every NanoGPT API endpoint.
Usage and billing rows are an audit trail and cannot be selectively deleted from the Usage page. Stored request/response logs are separate. Disabling logging stops future captures and immediately deletes retained logs for that API key. The “Clear all account request logs” action deletes all stored account request logs without changing which API keys will log future requests.
Use the request’s output-token limit, such as `max_tokens`, `max_completion_tokens`, or the endpoint’s documented equivalent, to cap generated output. This is different from counting input tokens or changing the model’s context window.
For thinking models, reasoning tokens may consume part or all of that output budget. A very low limit can therefore produce a truncated answer or little to no visible answer. When debugging, omit the limit or leave enough headroom for both reasoning and the final response.
For supported reasoning models, use `reasoning_effort` or `reasoning.effort` with a documented level such as `none`, `low`, `medium`, or `high`. Not every model supports every level, so check the model details and API documentation.
Use the Support page to create a private ticket tied to your current session. The help assistant can also prefill a ticket draft, but it will only create the ticket after you confirm.
Use the Terms of Service and Privacy Policy for NanoGPT account, usage, payment, data-handling, and provider terms. Provider retention differs by route; Zero Data Retention applies only where the selected provider and route explicitly support it.
Do not assume every model or provider is ZDR. For a compliance or security-audit requirement that is not answered by the published policies, contact support with the specific control or evidence you need.
Use the newsletter signup on the Help page to get occasional NanoGPT updates, including major product changes and new model additions.
Newsletter emails are separate from account emails and include an unsubscribe link.
NanoGPT works with many model providers. If you offer models NanoGPT should add, or can offer better pricing or reliability, contact the team through support.
NanoGPT exists to make many AI models available through one pay-as-you-go interface. The goal is to reduce friction for people who want access to strong models without managing many subscriptions or provider accounts.
NanoGPT also supports privacy-conscious and crypto-friendly payment paths for users who cannot or do not want to use traditional subscription flows.