






LMForge
Run AI Models on Your Own Machine
Download a model and within minutes you have a fast, OpenAI-compatible AI endpoint running entirely on your machine. LMForge detects your GPU, VRAM, and drivers, installs the best inference engine automatically, and keeps your models served by an always-on daemon — closing the app never stops them.
- Private — nothing leaves your device
- One-command install, auto engine selection
- OpenAI + Ollama compatible API
- Multi-model orchestration
- Rust
- Tokio
- axum
- SvelteKit
- Tauri



