You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
Run any local LLM engine, auto-tuned to your GPU — polished web UI + OpenAI/Anthropic-compatible API. Point Claude Code at your own machine in one command. No Electron, no Python, offline-first.
BoxedLLaMA is a specialized software toolkit designed for developers to integrate local artificial intelligence into Windows applications. It functions as a comprehensive wrapper for llama.cpp, automating complex tasks such as server installation, version updates, and model management from Hugging Face.
为DSH接入本地大模型能力:在「设置→插件」页一键启停本地 llama.cpp 大模型(双槽x双模态x双预设),卡片内配置、一条命令安装、自动注册,装完即用| Enable local large model capabilities for DSH: One-click start/stop for local llama.cpp models (dual‑slot × dual‑modal × dual‑preset) right in Settings → Plugins; configure within the card, install with a single command, automatically register, and ready to use
llama.cpp instance manager — spin up GGUF models with a web UI, REST API, and full flag control. Pull from HuggingFace, chat, manage instances from browser or CLI.
Profile-based launcher for local LLM servers - start, stop and switch models on llama.cpp, Ollama and LM Studio from one CLI/TUI. Zero resident overhead, optional MCP control plane, importable as a Go library.