← Browse
deepeval
confident-ai ★ 18344
DeepEval is an open-source LLM evaluation framework that provides metrics and tools for unit testing AI applications, agents, RAG pipelines, and chatbots using LLM-as-a-judge evaluation and locally-running NLP models. The plugin adds three skills to enable DeepEval evaluations, tracing, datasets, and integration with the Confident AI observability platform directly into Claude Code workflows.
Install
> /plugin marketplace add confident-ai/deepeval
> /plugin install deepeval
What it's made of
3 skills
- Commands
- 0
- Agents
- 0
- Skills
- 3
- MCP servers
- 0
- Hooks
- 0
What it needs & plugs into
- API keys
- none
- Paid services
- Confident AI
- External tools
- none
- Talks to
- nothing external detected
Facts extracted from the plugin's files. Prose generated by claude-haiku-4-5-20251001.
Is this your plugin?
Claim it to keep the card accurate and enter the weekly contest. Requires signing in as the GitHub owner (confident-ai).