Turn text trapped in a picture into something editable.
A photographed notice, a screenshot or a scanned page can contain the exact words you need, but copying them by hand is slow. OGAD (Off Grid AI Desktop) lets a local vision model read the image on your Mac and return text you can edit. After the app and model downloads, the analysis can run without a cloud AI upload.

What would you like to do with Off Grid AI?
Have a feature or use case you would like us to support? Tell us what you want to do and which device you use.
Write to support@offgridmobileai.co, join our Slack community, or talk to us on Reddit.
Get words out of the image
Use a clear photo of a printed notice for your first attempt. Ask for transcription before you ask for a summary. This keeps a reading mistake separate from the model’s interpretation.
Transcribe the visible text exactly. Preserve paragraph breaks and list items. Write [unclear] where you cannot read a word. Do not correct spelling or fill missing text.
The useful result is a first draft you can check. You supply the image and the question; the model supplies an answer for review.
Choose a model that can read images
Open Models and choose a local vision-capable model that fits your computer. The chat interface identifies options such as Qwen3-VL 2B or Gemma E4B when image support is missing. Download the complete model and its required vision component, then select it for the chat.
A text-only model cannot inspect the image. An image-generation model creates pictures; it is not the model you need for this analysis. Local chat and vision are part of OGAD’s free desktop app.
Use the model recommendations shown for your hardware. The installed file size is not the amount of working memory the model needs while answering. If a large model cannot load, choose a smaller supported vision option and close other heavy applications.
Get your first answer
- Save the image locally. Crop excess background, but keep the text and context needed for the task.
- In OGAD, open Chat and select the downloaded local vision model.
- Open the + composer menu and select Add image.
- Choose the file and check the attachment previews.
- Send the example question above, adapted to your image.
- Compare the answer with the original before you reuse it.
If the small text is unreadable, attach a close crop of that section and ask again. More confident wording in the prompt cannot restore detail absent from the image.
Check a notice before turning it into an action
Suppose the image is a notice about a room change. Ask for the text first and compare the room number, day, and time with the picture. Watch for small characters that can be confused, such as a zero and the letter O, or a one and a lowercase l.
If one line is unclear, crop that line with a little surrounding text and attach the crop in a new request. Name the exact question: “Does this line say Room 105 or Room 10S?” The model can still be wrong, so use the sharper image to make your own check.
Once the transcription is correct, paste the checked text and ask for a short action note. For example: “List the new room, the date it applies, and anything the notice asks me to do. Leave missing details marked as not stated.” This second stage should use the verified words, not silently repair the first answer from memory.
For a photographed table, keep the row and column headings visible. Read a small region at a time and check how each value relates to its heading. A correct number in the wrong column can be as misleading as a misread number.
Copy the final text into the place where you need it and keep the source image until you have checked the result. This is a way to save retyping while retaining a clear route back to the original.
This is vision-model transcription, not a promise of perfect OCR. Handwriting, tiny text, curved pages and low contrast can produce errors. For long documents with selectable text, attach the original document rather than photographing each page.
Keep the analysis on your computer
Select the local model and use files already on your computer. You can disconnect internet access after setup and send a short image question to check the local route. A remote model or an online tool changes that data path, so leave those out of this workflow.
Local analysis avoids sending the image to a cloud AI provider. Separate file backup and device-sync settings still apply if you have enabled them.
When the answer is not useful
| Symptom | Useful next step |
|---|---|
| Image input is unavailable | Select a vision model and complete its required downloads |
| Small text is misread | Attach a sharper crop of the relevant area |
| The answer is vague | Ask about one visible detail or one concrete decision |
| The model invents details | Request visible evidence and verify it against the source |
| The model cannot load | Try a smaller supported model that fits available memory |
These steps use the image-input workflow in OGAD 0.0.51. Keep the app current and use a complete local model.
Try it with your own image
Download OGAD, choose a vision model, and transcribe one short notice. Compare its lines with the original. Keep the source beside the answer. You get a useful starting point while the image stays within your local analysis workflow.