- 1
Ollama - Open-source local LLM runtime that runs models privately on your machine with a simple CLI and REST API, no restrictions on model choice or usage. - 2
LM Studio - Desktop app and headless server for running large language models locally with full privacy, supporting multiple model formats and an OpenAI-compatible API. - 3
Lemonade - AMD-backed open-source local AI server that runs LLMs, image generation, and speech on your GPU/NPU with OpenAI API compatibility and no cloud dependency. - 4
AI Backends - Self-hosted open-source API server exposing unified REST endpoints for multiple LLM providers with full local control. - 5
whichllm - CLI tool that auto-detects your hardware and ranks the best local LLMs from HuggingFace that fit your system, helping you find unrestricted models optimized for your machine.
i want unrestricted LLM to use it without ant limitation
- 1
Odysseus - Self-hosted AI workspace that integrates with Ollama, llama.cpp, and vLLM so you can download and chat with any open-weight model without provider-imposed content restrictions. - 2
ModelHub - macOS menu-bar app that discovers, downloads, and manages local LLMs from Hugging Face for Ollama, llama.cpp, and vLLM, letting you run whichever unrestricted weights you choose. - 3
Reame - Lean llama.cpp inference server exposing an OpenAI-compatible API that serves whatever GGUF model you supply, bypassing external moderation layers entirely. - 4
OpenJarvis - Local-first agent framework built around Ollama integration that runs entirely on your own hardware using the open-weight models you select. - 5
Locally AI β Local AI Chat - Free Apple app that runs open-weight Hugging Face models privately on-device without an account, allowing you to sideload unrestricted weights for offline use.
Have a tool question of your own? Describe what you need in plain English and let two models search our database for you.