The easiest way to run an open model on your own machine: one command downloads and starts it, plus a local REST API and libraries for Python and JavaScript.
Open models & inferenceunslothai
unsloth
A tool for running and fine-tuning models while saving memory: there is a desktop app, a web UI and a library. Fine-tuning works on a modest GPU.
Who it is for
For people who want to fine-tune a model on their data without expensive hardware.
How to start
- Download Unsloth Desktop from the releases page or install:
curl -fsSL https://unsloth.ai/install.sh | sh. - For code:
uv venv unsloth_env --python 3.13, thenuv pip install unsloth --torch-backend=auto. - Pick a ready notebook from Free Notebooks and swap in your data.
Steps are taken from the README. Check the current version in the repository before running them.
Stars over the last 30 days
Author's description
Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4, MiniMax-H3, Gemma 4, FLUX and more.
More in «Open models & inference»
The core library for working with models: one codebase for text, vision, audio and multimodal tasks, for both inference and training. Most new open models ship through it.
A self-hosted ChatGPT-style interface that connects to Ollama and any OpenAI-compatible API. Installs with a single Docker command and runs on your own server.
A C/C++ engine that runs language models on ordinary hardware: laptops, phones, GPU-less servers. Much of local AI, Ollama included, is built on it.
Repository of the large open DeepSeek-V3 mixture-of-experts model: description, benchmark results and run instructions. It showed an open model can stand next to closed ones.
The foundation under most modern neural networks: GPU-accelerated tensors and automatic differentiation, wrapped in ordinary Python. Transformers, Llama and nearly every open model are written on it.
