vanman2024 avatar

llm-evals

vanman2024/ai-dev-marketplaceby AI Dev Marketplace
v1.0.0development7 stars

LLM testing and evaluation framework with promptfoo, DeepEval, golden datasets, and Supabase-backed eval tracking

Install

/plugin install llm-evals@vanman2024-ai-dev-marketplace
AI writes the code. CodeRabbit catches the slop.
Try For Free →
Keeps your Mac awake while Claude Code, Codex or Cursor works. Lets it sleep when they are done.
Try free for 7 days →
Monitor with ease. Code with confidence.
Start Free Trial →
Connect your Claude agent to live crypto prices and trading routes via 1inch
Get the MCP →
Block distracting apps from your iPhone permanently without a 3rd party app. Free and open source.
Block now (100% free) →
Protect your code quality, stop the AI slop.
Try For Free →
belt cli automatically finds the best tools and skills for your agent. image, video, music, tts...
one prompt install →
Free an open-source skills that make AI coding agents easier to understand, verify, and control.
Download for free →

Contentsdiscovered from source

commands ×2

agents ×11

More from vanman2024/ai-dev-marketplaceView source

llm-testing · evaluation · promptfoo · deepeval · golden-datasets · regression-testing · prompt-testing · metrics · faithfulness · relevance · ci-cd · pytest