MIT LLM Serve Dashboard I am making open source
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
| A single-file, dependency-free live dashboard for your local LLM serving box — GPU utilization, per-model throughput, KV/context fill, and system stats for llama.cpp and vLLM, in one green terminal-styled page. No framework, no build step, no external requests. The frontend is one https://github.com/NHClimber87/llm-serve-dashboard What it shows
It's a work in progress but I know a few people asked for this dashboard in my other posts so please try it out and I will do my best to respond to questions and requests. This really helps me increase my observability. I am especially happy with the thought tap that displays the chain of thought the models have. Critical to have when using teacher model distills! [link] [comments] |
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.