One-click install
Pick a model from the library and Gotchi pulls it down, verifies the checksum, and loads it - ready to chat in seconds.
Browse the full open-source model catalog, download what you need, and run it on your hardware - no cloud account required.
Gotchi handles downloads, updates, quantization, and hardware detection so you can focus on using AI, not configuring it.
Pick a model from the library and Gotchi pulls it down, verifies the checksum, and loads it - ready to chat in seconds.
Gotchi detects your GPU, VRAM, and unified memory to recommend the best quantization level. Apple Silicon, NVIDIA, and CPU-only all supported.
Choose between Q4, Q5, Q6, Q8, and full-precision weights. Smaller quants run faster on limited hardware with minimal quality loss.
When a model publisher releases a new version, Gotchi notifies you and offers a one-click upgrade - no manual downloading or swapping.
See exactly how much disk each model uses. Delete, archive, or move models between drives without breaking your setup.
Built-in tok/s benchmarking for every model on your hardware. Compare models side by side before committing to a workflow.
New models are available the day they launch. Here are some of the most popular families.
Load multiple models simultaneously and route tasks to the best one. Use a small model for quick answers and a large one for deep analysis.
Metal acceleration on Apple Silicon and CUDA/AVX2 on Windows give near-native speed for 7B–70B parameter models. Gotchi detects your hardware and picks the right build automatically.
Download Gotchi and start exploring the open-source model library - local, private, and always available.