Molten brings AI to your devices—completely local, completely private.
Run LLMs Offline
Chat with models from Ollama, Swama (MLX-optimized for Apple Silicon), or built-in On Device Models. No cloud API, no tracking—your data stays your devices.
Key Features
Seamless self-hosted server switching and model selection, including on-device models.
Real-time streaming responses with performance analytics (tokens/sec, eval rates).
Native SwiftUI interface with markdown rendering, syntax highlighting, dark mode, and keyboard shortcuts.
Floating panel for quick chats on macOS.
Privacy Obsessed
100% local and private processing. Open-source for full audit on GitHub. No telemetry or cloud dependency.
Download Molten and own your AI today.