AI

Private local AI with Ollama: what it is, what hardware you need, and how to connect it

· 7 min read · বাংলায় পড়ুন

Status: Cinematic Recorder currently takes your own key for OpenAI or Google Gemini. Support for Ollama is coming soon — this guide gets you ready for it.

Cloud AI is convenient, but every request leaves your home. Ollama local AI takes a different path: it runs open models on your own computer, so prompts and answers never go to an outside company. This guide explains what Ollama is, what hardware you realistically need, and how to connect it to Cinematic Recorder on your phone.

What is Ollama?

Ollama is a free program for Windows, macOS and Linux that downloads and runs AI models locally. You install it, pull a model with one command, and it serves that model on your computer. Other apps, including apps on your phone, can then send requests to it over your home network.

The models are "open-weight" models, which means their files can be downloaded and run by anyone. They range from small ones that run on an ordinary laptop to large ones that need a powerful graphics card.

Why choose local AI?

  • Privacy. Your prompts stay on hardware you control. This is the main reason most people choose it.
  • No per-request bill. After setup, you only pay for electricity.
  • No account or card needed. Useful if international payments are hard for you.
  • Works without internet once the model is downloaded, as long as your phone and computer share a local network.

The honest downsides

  • Local models are usually smaller than top cloud models, so answers can be less polished, especially in Bangla.
  • Speed depends on your computer. A slow machine means slow replies.
  • Your computer must be on and reachable whenever you want AI on your phone.
  • Setup takes a little more effort than pasting a key.

What hardware do you need?

There's no single answer, because it depends on the model size. As a rough guide:

ComputerWhat to expect
Laptop with 8 GB RAM, no graphics cardSmall models (a few billion parameters) work for short text tasks, slowly
16 GB RAM, or an Apple Silicon MacMedium models run comfortably for titles, summaries and drafts
Desktop with a recent graphics card with plenty of video memoryLarger models and much faster replies

Free disk space matters too: each model can take several gigabytes. Start with a small model, see how it feels, and move up only if you need to. Our guide on choosing an AI model explains why bigger isn't always better.

Setting up Ollama on your computer

  1. Download Ollama from the official website and install it.
  2. Open a terminal (Command Prompt on Windows, Terminal on Mac) and pull a model, for example with ollama pull followed by the model name listed on Ollama's model library.
  3. Test it with ollama run and the model name. Type a question; if it answers, the model works.
  4. By default Ollama only listens to the computer itself. To let your phone reach it, set the OLLAMA_HOST environment variable to 0.0.0.0 and restart Ollama. Ollama's documentation explains this for each operating system.
  5. Allow Ollama through your computer's firewall for your private (home) network only.
  6. Find your computer's local IP address (it usually looks like 192.168.x.x). Ollama normally uses port 11434.
Only open Ollama on a trusted home or office network. Don't expose it to the public internet or to public Wi-Fi, because anyone who can reach it could use your computer's AI.

Connecting Ollama to Cinematic Recorder (once it is supported)

  1. Connect your phone to the same Wi-Fi as the computer.
  2. Open Cinematic Recorder and go to the AI or "bring your own key" settings.
  3. Choose Ollama as the provider.
  4. Enter your computer's address, for example http://192.168.1.20:11434 (use your own IP).
  5. Choose the model you pulled, then save.

If it doesn't connect, check that both devices are on the same network, that Ollama is running, and that the firewall allows it. Some routers isolate devices on guest networks, so use your main Wi-Fi.

Read more about provider options on the features page. Keep in mind that captions from speech currently run on our server with credits, or with your own OpenAI or Gemini key. AI editing in plain words and AI translation are still coming soon; setting up Ollama now means they can use your local model when they arrive.

When local AI makes the most sense

Work content you can't share

If you record internal tools, client dashboards or school records, local AI keeps that text on your own machine. Pair it with careful recording habits from our guide to privacy before you record.

Lots of small, repeated jobs

If you generate many titles or summaries each week, local AI avoids a growing bill. It's one of the best ways to keep AI costs low.

Learning and experimenting

Trying different open models is a good way to understand what AI can and can't do, without worrying about cost.

FAQ

Can Ollama run directly on my Android phone?

Cinematic Recorder will connect to Ollama running on a computer on your network. Running large models on a phone itself is slow and drains the battery, so a computer is the practical choice.

Is Ollama completely private?

The model runs on your computer, so requests don't go to an AI company. Your privacy then depends on keeping that computer and network secure.

Is Ollama free?

Ollama and many open models are free to download. Always check each model's licence if you plan to use it for business.

Why are answers slow?

The model may be too big for your hardware. Try a smaller model, close heavy programs, or use a computer with a graphics card.

Local AI with Ollama trades a bit of convenience for real control. If privacy matters to you, it's well worth an afternoon of setup — and the editor keeps working fully even when your computer is off.

Try Cinematic Recorder free — record, then make it cinematic with auto zoom and a smart cursor. Download · All features