
llm-evals
vanman2024/ai-dev-marketplaceby AI Dev Marketplace
v1.0.0development7 stars
LLM testing and evaluation framework with promptfoo, DeepEval, golden datasets, and Supabase-backed eval tracking
Install
/plugin install llm-evals@vanman2024-ai-dev-marketplaceContentsdiscovered from source
commands ×2
agents ×11
Featured
Keeps your Mac awake while Claude Code, Codex or Cursor works. Lets it sleep when they are done.
Try free for 7 days →Connect your Claude agent to live crypto prices and trading routes via 1inch
Get the MCP →Block distracting apps from your iPhone permanently without a 3rd party app. Free and open source.
Block now (100% free) →belt cli automatically finds the best tools and skills for your agent. image, video, music, tts...
one prompt install →Free an open-source skills that make AI coding agents easier to understand, verify, and control.
Download for free →More from vanman2024/ai-dev-marketplaceView source
llm-testing · evaluation · promptfoo · deepeval · golden-datasets · regression-testing · prompt-testing · metrics · faithfulness · relevance · ci-cd · pytest
More from vanman2024/ai-dev-marketplace
All 25 plugins →- mem0v2.1.0
- ml-trainingv2.0.0
- mobilev2.0.0
- nextjs-frontendv3.0.0
- openrouterv2.0.0
- paymentsv2.0.0
- plugin-docs-loaderv2.0.0
- rag-pipelinev3.2.0
- redisv2.0.0
- resendv2.0.0
- supabasev2.0.0
- sveltekit-frontendv2.0.0
- vercel-ai-sdkv3.0.0
- website-builderv2.1.0
- a2a-protocolv2.0.0
- bullmqv1.0.0
- celeryv2.0.0
- claude-agent-sdkv2.1.0
- clerkv2.1.0
- digitaloceanv1.0.0
- elevenlabsv2.0.0
- fastapi-backendv2.0.0
- google-adkv2.0.0
- langgraphv1.0.0