Chat with a local model
On first launch Draggy sizes a model to your graphics card and downloads it. Three answering modes, and you can swap the model whenever you like.
Draggy talks, browses and works with your files using models on your own machine. It drives a local Ollama, so there is no account and no API key.
Free software under the GPL. Windows, macOS and Linux. Version 1.2.6.
What it does
Answering questions is the easy half. Draggy also reads your files, writes new ones, searches, and opens pages for itself.
On first launch Draggy sizes a model to your graphics card and downloads it. Three answering modes, and you can swap the model whenever you like.
Word, PowerPoint, Excel, PDF, code and plain text, saved where you can open them. Attach a document and it reads that back.
It can search the web, open a result and read it, or drive a real browser session when a page needs clicking through.
Point it at a folder. The contents are indexed on your own machine and searched on meaning and keywords at once.
Python and JavaScript run in a sandbox, so the model can work an answer out rather than guess at it.
Thirty-four Model Context Protocol servers in a catalogue, none of them switched on until you switch one on.
A continuous voice mode that works out when you have finished a sentence, answers out loud, and stops when you cut in.
English, French, Spanish, German, Italian, Portuguese, Dutch, Russian, Chinese, Japanese, Korean and Arabic.
Updates download in the background and install the next time you open the app. Nothing is installed behind your back.
Voice mode
Voice mode listens continuously and works out when your sentence has ended, so there is no button to hold and no wake word to remember.
Built-in browser
Links open in a browser window inside the app, running uBlock Origin's filter lists on Ghostery's engine. The model reads the same clean page you do.
History
Every chat is stored on your own disk and searchable from the history list. Any one of them exports to a Markdown file in a click.
Requirements
The model runs on your hardware, so your hardware decides how fast it answers. Ollama has to be installed; Draggy offers to install it if it cannot find it.
| System | Windows 10 64-bit (1809 or newer), macOS 11, 64-bit Linux with glibc 2.28 or newer |
| Processor | Intel Core i5-8250U, AMD Ryzen 3 3200U, Apple M1 |
| Memory | 8 GB |
| Graphics | Intel UHD 620, AMD Radeon Vega 8 |
| Disk | 10 GB free |
| System | Windows 11, macOS 14, a current Linux |
| Processor | Intel Core i5-12400, AMD Ryzen 5 5600, Apple M2 |
| Memory | 16 GB |
| Graphics | GeForce RTX 3070, Radeon RX 7600, Intel Arc A750 |
| Disk | 20 GB free, on an SSD |
Privacy
Model downloads, searches you or the model trigger, pages the browser opens, the speech models on first use, the ad blocker's filter lists, site icons for search results, the update check, and any extension you switch on. Everything else stays local. There is no telemetry, and nowhere for it to go.