You can start without paying anything

Set up RefineLoop in about 3 minutes

For your first 14 days you can skip this page entirely: press Start on the app's first screen and corrections run through RefineLoop's own connection: no key, no account, nothing to create. When you are ready for your own connection (free, and yours forever), the app walks you through it: Settings (Ctrl+Alt+O) → Get connected free. This page is the same guidance in long form, for reading ahead or helping somebody else.

GroqFree, no card, and one key covers writing, speech, and voice.Free tier · no cardGoogle GeminiFree Flash models; some accounts are asked to set up billing.Free tier · variesOllama on your PCNothing leaves your computer. Needs a gaming-class GPU to feel fast.Free forever · offlineAzure, OpenAI, ClaudeBest quality and the choice for work data.Pay per use · cents a day

Free path 1: a Groq key (no credit card)

The fastest way to have RefineLoop working today. Groq's free tier has no credits system and no card, just rate limits.

  1. Open console.groq.com and sign in with Google, GitHub, or email.
  2. Go to API Keys and click Create API key. Copy it; it starts with gsk_.
  3. In RefineLoop press Ctrl+Alt+O, set Provider to groq, and paste the key.
  4. Leave the model as openai/gpt-oss-120b, click Test connection, then Save all.
  5. For read-aloud only: Groq gates its voices behind a one-time approval. Open console.groq.com/playground, select the Orpheus voice model, and accept the terms. Then press Test voice in section 3 of Settings. Speech-to-text needs no such step.

One key runs everything. Groq also serves speech-to-text (Whisper) and read-aloud (Orpheus voices) on the same free key, so dictation, voice messages, and natural read-aloud work immediately, with no Python and no extra deployments. Set “Speech-to-text runs” to Same as AI model in section 2 if you would rather not install Python. Free limits, measured on a real key: about 1,000 corrections, 2,000 dictations, and 8 hours of audio a day (far more than one person uses), but read-aloud is about 100 plays a day. When those run out, RefineLoop switches to the Windows voice by itself, so the sound never stops, it just gets plainer until the next day. Groq runs open-source models (Llama, Qwen, GPT-OSS), which handle English corrections well; a paid Gemini/OpenAI/Claude key is still noticeably better for Vietnamese explanations. Switching later is one dropdown.

Free path 2: a Google Gemini key

Also free for the Flash models we use, though Google has been tightening this. If AI Studio asks you to set up billing, use Groq above instead; nothing in RefineLoop depends on which one you pick.

  1. Open aistudio.google.com/apikey and sign in with any Google account.
  2. Click Create API key. Choose a project if it asks; any project works.
  3. Copy the key. It looks like AIza… and is shown only once, so paste it somewhere safe.
  4. In RefineLoop press Ctrl+Alt+O, set Provider to gemini, and paste the key into API key.
  5. Leave the model as gemini-2.5-flash, click Test connection, then Save all. That is it. Try Ctrl+Alt+B on a sentence.

Be aware: on Google's free tier, your prompts may be used to improve their products. That is fine for practising English and everyday chat. For confidential work text, use a paid key or Ollama; the app treats every provider the same.

Free and fully private: Ollama on your own PC

No key, no account, no internet needed after the download. RefineLoop sets it up for you; there is no terminal in this path. Speed is the catch: this needs a dedicated graphics card. On a laptop with built-in graphics a word lookup takes around 20 seconds, where a free cloud key answers in one or two.

  1. In RefineLoop press Ctrl+Alt+O and set Provider to ollama.
  2. RefineLoop checks what your PC is missing and offers to fix it. Press Set it up and it downloads Ollama, runs the installer, starts it, and pulls the model. Nothing downloads until you press that button.
  3. Pick the model when asked: llama3.1 (about 4.7 GB) is the recommended one, gemma3:4b (about 3.3 GB) suits smaller laptops.
  4. When it says local AI is ready, click Save all. From then on RefineLoop starts Ollama by itself when it needs it.

Local models are weaker than cloud ones, especially for Vietnamese and for the “why” explanations. Many people start on Gemini and switch to Ollama later for sensitive work.

Advanced: your own AI gateway (OmniRoute, LiteLLM, LM Studio…)

If you already run a local gateway that speaks the OpenAI API and routes to providers you chose, RefineLoop can use it like any other endpoint. This path is for people comfortable running a server. Nobody needs it, and the paths above stay simpler.

  1. In RefineLoop press Ctrl+Alt+O, set Provider to custom, and enter your gateway's address as the Endpoint: for OmniRoute that is http://localhost:20128/v1, with an API key from its dashboard and auto as the model.
  2. Turn the gateway's prompt compression OFF (OmniRoute: Dashboard → Settings → Compression → off, globally). Compression rewrites “verbose” text before sending, and your imperfect sentences are exactly what RefineLoop must deliver untouched. A compressed prompt silently breaks corrections.
  3. In section 2, set Speech-to-text runs on to This PC. Gateways generally do not carry audio, and the bundled local speech works regardless.
  4. Press Test connection, then Save all.

Two honest limits. RefineLoop treats a localhost endpoint as a local model, so it uses the shorter prompts written for laptop-class models even when your gateway routes to a big cloud one. Quality can be a shade below connecting to that provider directly. And privacy is what your gateway does, not what the address looks like: localhost here is a hop, and your text continues to whichever provider the gateway picks.

Speech: nothing to configure

Dictation, voice messages, instant replay, and shadowing work out of the box.

  1. The app runs a small speech server on your PC and starts it for you, hidden. Your voice never leaves the computer.
  2. It needs Python, and if it is missing, the first speech hotkey simply offers to install it for you (about 25 MB, one click). Installing yourself from python.org with “Add python.exe to PATH” ticked works too.
  3. The very first use downloads the speech model (about 460 MB), once. After that it is instant and offline.

If something goes wrong

Connection refused (localhost:11434)Provider is set to ollama but Ollama is not running. Start Ollama, or switch the provider to gemini.
DeploymentNotFoundAzure: the Deployment field must contain your deployment name from the Azure portal, not the model name.
401 / UnauthorizedThe API key is wrong, expired, or has a stray space. Paste it again, and use Show keys to check it.
Nothing happens on a hotkeyAnother app owns that shortcut. The tray menu shows a red warning; click it to rebind.
Read-aloud sounds roboticYou are hearing the Windows fallback voice. Press Test voice in Settings section 3 - it says exactly why. On Groq, accept the voice model terms in their playground; on Azure, add a tts deployment.
Speech features do nothingPython is missing. Press a speech hotkey and choose Set it up to let the app install it; or see logs\whisper-server.log.

Still stuck? Email support@refineloop.app with what you see on screen; a screenshot is perfect.

← Back to RefineLoop