Tokenstead
Select your rig from a list of 56 tracked hardware configs — Mac unified-memory devices, multi-GPU setups up to 8x, or custom specs — and…
Local inference runtimes (Ollama, LM Studio, llama.cpp).
Select your rig from a list of 56 tracked hardware configs — Mac unified-memory devices, multi-GPU setups up to 8x, or custom specs — and…
Open WebUI is a self-hosted chat interface that connects to local models via Ollama, cloud providers like OpenAI and Anthropic, or any…
The repository ships three concrete layers: Python export recipes for popular Hugging Face models, reusable PyTorch primitives for…
The installer handles the assembly: LLM inference via Ollama, a chat interface, voice input/output, RAG over private documents, local image…