Dashboards
17 results| Name | Popularity | Author | Quality | Modified |
|---|---|---|---|---|
| vLLM Monitoring - V2vllminferencellm vLLM 추론 서버 모니터링 | 0 0 | miensoap | Good | |
| Ollama LLM Inferenceollamallminference Ollama dashboard for the ollama-exporter written in go (https://github.com/maravexa/ollama-exporter) | 0 0 | maravexa | Good | |
| Claude Code Metrics (Prometheus)anthropicclaude-codecost-tracking Claude Code usage analytics on Prometheus: cost, token consumption, sessions, active coding time, lines of code, commits, pull requests, tool decisions, cache hit ratio, and leaderboards by user, model, and session. Consumes Anthropic's OpenTelemetry metrics via OTLP. Works with Prometheus, VictoriaMetrics, Mimir, and Thanos. Requires Grafana 11+. | 0 0 | rockdarko | Good | |
| Claude Code Metricsclaude-codeanthropicllm Comprehensive monitoring dashboard for Claude Code CLI usage. Track sessions, tokens, costs, commits, pull requests, lines of code, and developer productivity metrics. Compatible with Prometheus, VictoriaMetrics, Mimir, and Thanos. Metrics required: - claude_code_session_count_total - claude_code_token_usage_tokens_total - claude_code_cost_usage_USD_total - claude_code_active_time_seconds_total - claude_code_lines_of_code_count_total - claude_code_commit_count_total - claude_code_pull_ | 0 0 | shashwatsaxena | Good | |
| vLLM Inference Monitor (中文汉化版)vllmllminference 基于原版 #24755 的完整中文汉化版本。覆盖 vLLM 推理核心指标:调度器效率、KV Cache 使用率、TTFT/TPOT 延迟、Prefix Cache 命中率、请求完成原因分布等。已适配 Grafana 10.4+,开箱即用。 | 0 0 | q673475900 | Good | |
| Pulze LLM Application Overviewpulzepulze.aillm Pulze.ai Dashboard around LLM application performance, costs, and usage associated with various providers and models. | 0 0 | pulze.ai | Good | |
| LLM Inference Dashboard (SGLang / vLLM)llminferencesglang 同时支持 SGLang 和 vLLM 的统一 LLM 推理监控 Grafana Dashboard,数据源为 Prometheus。该模板将两类推理后端的可比指标放在同一视图中展示,覆盖请求量、Token 吞吐、延迟分位数、队列状态、缓存行为以及 Engine 或 TP Rank 维度的分布情况。适用于混合推理集群、后端迁移、压测对比和统一运维场景,帮助团队用一致的视角观察不同推理框架的服务表现。 | 0 0 | patientflounder2051 | Good | |
| vLLM Monitoring - V1vllminferencellm vLLM 추론 서버 모니터링 대시보드 | 0 0 | miensoap | Good | |
| llama.cpp Monitoringllama.cppllamacppinference llama.cpp LLM server metrics | 0 0 | Paulo Castro | Good | |
| Neurix — Ollama & NVIDIA GPUollamagpunvidia Ollama LLM runtime metrics — models, VRAM, GPU telemetry, process stats. Requires Neurix ollama_exporter. | 0 0 | diyrex | Good | |
| NVIDIA GPU for AI Workloads — Efficiency, Cost & Health (DCGM)gpunvidiadcgm GPU monitoring for self-hosted AI: are your tensor cores actually working, what does each run cost in kWh, and is your VRAM degrading? Built and battle-tested on a local LLM rig (100B+ parameter models). Requires dcgm-exporter with DCP/profiling metrics enabled. Multi-GPU via UUID variable. | 0 0 | orangeoctopus1069 | Good | |
| LiteLLM Proxy — SRElitellmllmproxy LiteLLM proxy panels Fully compatible with LiteLLM 1.92.0 | 0 0 | punyvulture387 | Fair | |
| LangGraph API Metricslanggraphllmapi This is a dashboard for Langraph API System metrics | 0 0 | murtazayusufali1 | Good | |
| LLM Simulation - vLLM Serving Overviewllmvllmsimulation vLLM serving overview: first-token and inter-token latency, throughput, queue depth, KV cache, prefix-cache reuse and a TTFT error budget, every panel by (model_name). V1 metric names; several panels need the llm:* recording rules. | 0 0 | ChrisAdkin | Good | |
| Open WebUIopen-webuillmobservability Usage, cost, and quality metrics for Open WebUI — users, chats, tokens, spend by model and by user, and per-model satisfaction. Requires the Open WebUI Prometheus exporter (polls the REST API, no database access). | 0 0 | baselmathar | Good | |
| local-ai Metricslocal-aitradingllm Metrics for Local AI | 0 0 | fdu0157077 | Good | |
| GodAgent v2 — Agentic Pipelinegodagentagentsllm GodAgent v2 — agentic pipeline observability. Tracks LLM costs, token usage, Motor Memory hit rate, circuit breakers, pipeline CLEAR score. | 0 0 | aramiskessler | Good |