LM-Kit One
The vendor describes LM-Kit One as a self-hosted AI application server that exposes OpenAI, Anthropic, and Ollama-compatible endpoints, so…
Local inference runtimes (Ollama, LM Studio, llama.cpp).
The vendor describes LM-Kit One as a self-hosted AI application server that exposes OpenAI, Anthropic, and Ollama-compatible endpoints, so…
Select your rig from a list of 56 tracked hardware configs — Mac unified-memory devices, multi-GPU setups up to 8x, or custom specs — and…
Open WebUI is a self-hosted chat interface that connects to local models via Ollama, cloud providers like OpenAI and Anthropic, or any…
Pinokio is an open-source desktop launcher that wraps open-source AI tools — image generators, audio DAWs, TTS engines, video models — in…
The repository ships three concrete layers: Python export recipes for popular Hugging Face models, reusable PyTorch primitives for…
The installer handles the assembly: LLM inference via Ollama, a chat interface, voice input/output, RAG over private documents, local image…
llama.cpp is a C/C++ inference engine that runs quantized LLMs entirely on local hardware, from an Apple Silicon laptop to an H100 cluster…
LocalAI is a self-hosted, MIT-licensed stack that exposes an OpenAI-compatible REST API from your own hardware. Language model inference…
LM Studio, built by Element Labs Inc., is a desktop and server runtime for running open-source LLMs — Qwen, Gemma, DeepSeek, gpt-oss, and…