Skip to content
ai101.tools
navigateopenescclose
Claude+372Whisper+228LangChain+168Codex+223NotebookLM+276DALL-E 3+192DeepL+249n8n+208Topaz Video AI+153LlamaIndex+161
DeepEval Skills logo
Skill

DeepEval Skills

Evaluation, tracing and OpenTelemetry export for LLM apps, split into three non-overlapping skills.

0
SaveVisit website ↗

Install

Official
Claude Code: /plugin marketplace add anthropics/claude-plugins-official → /plugin install deepeval@claude-plugins-official
Triggers on

Bir LLM uygulamasına değerlendirme, izleme ya da OpenTelemetry dışa aktarımı eklenirken

Needs tools
ReadWriteBash
Author
Confident AI
License
Apache-2.0

Three skills, and the interesting part is the boundary between them. One builds evaluation suites: goldens, datasets, pytest runs, metrics, and the loop where failures drive prompt and retrieval changes. One instruments an application with DeepEval's own tracing and the framework integrations that ship with it. One covers plain OpenTelemetry export for teams that do not want the SDK. Each frontmatter states what it must not handle and names the sibling that should, which is a cheap answer to the usual problem of three adjacent skills all firing on the word eval.

Comments(0)

Sign in to comment

No comments yet — be the first.

Similar skills

Report this comment

Why are you reporting this?