Skip to content
ai101.tools
navigateopenescclose
Claude+372Whisper+228LangChain+168Codex+223NotebookLM+276DALL-E 3+192DeepL+249n8n+208Topaz Video AI+153LlamaIndex+161
MLflow Skills logo
Skill

MLflow Skills

Closes the loop for agent work: instrument, trace, evaluate, change, then verify the change helped.

0
SaveVisit website ↗

Install

OfficialShips scripts
Claude Code: /plugin marketplace add anthropics/claude-plugins-official → /plugin install mlflow@claude-plugins-official
Triggers on

Bir LLM uygulamasına izleme eklenirken, iz üzerinden arıza aranırken ya da değerlendirme puanlayıcısı kurulurken

Needs tools
ReadWriteBashGrep
Author
MLflow
License
Apache-2.0

The pieces are ordinary MLflow — tracing for Python and TypeScript, trace retrieval, metric queries, failure analysis — but the arrangement is the point, because without a scorer an agent improvement is just an opinion. The scorer skill refuses to start from the catalogue: it works out what the application does, extracts a small set of atomic quality criteria from that, and only then picks the cheapest scorer that can judge each one, treating the user as the authority on what counts as wrong. Aiming for a runnable and inspectable prototype rather than a perfect first pass is what keeps the loop turning.

Comments(0)

Sign in to comment

No comments yet — be the first.

Similar skills

Report this comment

Why are you reporting this?