llama-dash - Local LLM Ops
I've been building llama-dash, a single-pane dashboard and logging proxy for a self-hosted local inference stack. I run llama-swap + llama.cpp on a box at home and got tired of having zero visibility — no request log, no idea which model was loaded when, no way to hand out scoped access without exposing the raw backend. So llama-dash sits in front as one public port: it proxies the OpenAI/Anthropic-compatible /v1/* endpoints unchanged (streaming SSE passes straight through), logs every reque...
Original Source
Read the full article at Dev →KhanList aggregates and links to publicly available news content. We do not host full articles from third-party sources. Always verify important information with original sources.