Kalyanikhandare29 avatar

agent-evaluation

community1 stars

Evaluation frameworks and LLM-as-judge techniques for testing and validating AI agent systems

Install

/plugin install agent-evaluation@kalyanikhandare29-agent-skills-for-context-engineering
AI writes the code. CodeRabbit catches the slop.
Try For Free β†’
Make coding agent sessions - Searchable, Shareable, Vendor-neutral & Scored.
Try For Free β†’
Your agent targets a perfect 10 Code Health score. Deterministic. Every commit.
Try For Free β†’
Integrate web data into your AI product. One API to scrape website & brand data.
Get API Key Now β†’
belt cli automatically finds the best tools and skills for your agent. image, video, music, tts...
one prompt install β†’
create and run specialised agents in minutes
build now β†’
Plug Mailtrap into your AI workflow and let it handle the email.
Connect Mailtrap MCP β†’
Agent, run crypto. Access onchain data & trade routes via 1inch.
Install now β†’