{"repo":"AtomicBot-ai/Atomic-Chat","free":true,"listed":false,"github":"https://github.com/AtomicBot-ai/Atomic-Chat","clone":"git clone https://github.com/AtomicBot-ai/Atomic-Chat.git","description":"Local AI app and inference engine for agents. Run open-weight LLMs locally — private, 100% offline on your computer. Join our Discord: https://discord.com/invite/8wGSsvmg4V","language":"TypeScript","stars":1309,"topics":["ai-chat","ai-tools","apple-silicon","chatgpt","deepseek","desktop-app","gemma","gguf","gpt-oss","llamacpp"],"license":null,"category":"machine-learning","readme_excerpt":"Atomic Chat Local AI app and inference engine for agents. Run open-weight LLMs locally — private, on your machine. &nbsp; &nbsp; &nbsp; &nbsp; &nbsp; Getting Started · Hugging Face · Discord · X / Twitter · Bug Reports --- 📦 Download Desktop &nbsp; &nbsp; Mobile &nbsp; --- 🔌 Use It as an API Atomic Chat runs an OpenAI-compatible server at http://localhost:1337/v1 — a drop-in replacement for the OpenAI SDK. Load a model in the app, then point any client at it: Bound to 127.0.0.1 by default; set host: 0.0.0.0 to expose it on your LAN. Works with any agent, CLI, or IDE plugin that speaks the OpenAI API — see Launch With below. --- ✨ Features Local models - Run open-weight LLMs locally from HuggingFace — Llama, Gemma, Qwen, Mistral, Phi, and others - Multi-Token Prediction (MTP) speculative decoding — 30–70% throughput boost on supported models, up to 3× on Gemma 4 - DFlash block-diffusion decoding — up to 6× faster on Qwen 3.6, Gemma 4, Kimi K2.5 - Flash Attention toggle ( on / off / auto ) - Automatic reasoning-context tracking for chain-of-thought models - Auto context-window expansion with overflow notifications - EAGLE-3 speculative decoding for Gemma 4 on Apple Silicon (MLX) - MTP on MLX for Qwen 3.5 / 3.6 and DeepSeek V4 - TurboQuant KV cache ( turbo3 / turbo4 ) on llama.cpp — now on Windows & Linux too, not just macOS: up to 4.3× smaller KV cache footprint, CPU and GPU (CUDA / Vulkan) - TurboQuant KV cache on MLX-VLM — smaller memory footprint via RHT-correct fast paths","default_branch":null,"files":null,"tree":[],"storefront":"/r/AtomicBot-ai","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/AtomicBot-ai/Atomic-Chat/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}