
eval-1337
yzavyas/claude-1337by yzavyas
v0.2.1MITdevelopment5 stars
Write rigorous evals for LLM agents, skills, MCP servers, and prompts. Use when: building test suites, measuring effectiveness, choosing frameworks. Covers: DeepEval, Braintrust, RAGAS, precision/recall, F1.
Install
/plugin install eval-1337@yzavyas-claude-1337Contentsdiscovered from source
agents ×16
Featured
Make coding agent sessions - Searchable, Shareable, Vendor-neutral & Scored.
Try For Free →Your agent targets a perfect 10 Code Health score. Deterministic. Every commit.
Try For Free →Integrate web data into your AI product. One API to scrape website & brand data.
Get API Key Now →belt cli automatically finds the best tools and skills for your agent. image, video, music, tts...
one prompt install →Plug Mailtrap into your AI workflow and let it handle the email.
Connect Mailtrap MCP →Agent, run crypto. Access onchain data & trade routes via 1inch.
Install now →More from yzavyas/claude-1337View sourceHomepage
evals · evaluation · testing · agents · skills · mcp · precision · recall · f1 · deepeval · braintrust · ragas