Msty
Desktop app for local LLMs โ model management, RAG, and chat in one polished interface
๐ Overview
Msty is a desktop application for running local LLMs with a polished ChatGPT-like interface โ featuring built-in model management, RAG with document upload, and multi-model chat. Unlike command-line tools (Ollama) or basic web UIs (WebUI), Msty provides a native desktop experience with drag-and-drop model installation, automatic GPU/CPU detection, and a clean conversation interface. The RAG feature lets you upload documents (PDF, TXT, markdown) and query them alongside normal chat. Available on Windows, macOS, and Linux with offline-first design.
โจ Key Features
- โขPolished native desktop app โ Windows, macOS, Linux
- โขOne-click model download with automatic quantization
- โขBuilt-in RAG: document upload + vector search
- โขAutomatic GPU detection (CUDA, Metal, ROCm) with CPU fallback
- โขMultiple concurrent conversations with history
- โขOffline-first design โ no cloud required
- โขDrag-and-drop document ingestion for knowledge bases
๐ฏ The Problem It Solves
Local LLM setup remains fragmented โ you need Ollama for model serving, a separate UI for chat, another tool for RAG, and manual config for each. Msty consolidates all of these into a single desktop application with a polished interface that non-technical users can navigate.
๐ง How It Works
Msty bundles Ollama model management with a native desktop chat interface. Models are downloadable with one click, with automatic quantization selection based on your hardware. The chat interface supports multiple concurrent conversations, conversation history, and model switching. RAG mode documents the sidebar, chunks them, embeds them into a local vector DB, and serves retrieval-augmented responses. GPU acceleration is auto-detected (CUDA, Metal, ROCm), with CPU fallback. No cloud required โ everything runs locally.
๐ Installation & Quick Start
Installation
Download from https://msty.appQuick Start
- Download and install
- Browse models in-app
- Download a model
- Start chatting
โ Pros
- โขMost polished local LLM desktop interface
- โขOne-click model management โ no terminal needed
- โขBuilt-in RAG with document upload
- โขAutomatic GPU/CPU detection and optimization
- โขOffline-first โ full data sovereignty
- โขCross-platform: Windows, macOS, Linux
- โขAffordable Pro tier at /mo
โ Cons
- โขSmaller community than LMStudio or Jan
- โขAdvanced RAG features require Pro subscription
- โขNo API server mode โ desktop-only
- โขVery large models (70B+) still require high-end hardware
- โขProprietary license โ no self-hosting or source access
๐ฌ Practitioner Verdict
โMsty is the most polished local LLM desktop app available โ the model management alone saves significant setup time, and the built-in RAG works well for personal knowledge bases. The trade-off: smaller communityLMStudio or Jan have larger user bases, advanced features require Pro subscription, and very large models (70B+) still need significant hardware. For non-technical users who want local AI without terminal commands, Msty is the best starting point.โ
Self-Hosted (Free)
Open source, MIT/Apache licensed. Run it yourself.
โญ Star & Clone on GitHubFree forever. Your infrastructure, your data.
Msty Free
Full chat, basic RAG
- Full chat
- Basic RAG
- Model downloads
Msty Pro
Unlimited models, advanced RAG
- Unlimited models
- Advanced RAG
- Priority support
๐ Specifications
- Language
- TypeScript
- License
- Proprietary
- Platform
- Linux, macOS, Windows
- Supported Models
- REST API, CLI
๐ฐ Pricing Reality
Free tier: full chat, basic RAG, limited model downloads. Pro /mo: unlimited models, advanced RAG with reranking, priority support, early features. No API usage fees โ you run your own models. Optional cloud backup/sync is extra.