AgentEval — security grade SAFE, quality 67/100

Security audit verdict: SAFE · quality 67/100

No red flags found in any of the 11 categories — no credential harvesting, no data exfiltration, no curl-pipe-shell installer. Scanned against the SlowMist agent-security taxonomy, refreshed every 8 hours. Full audit →

by AgentEvalHQ · Agent Tool · ★ 154

Last updated: · Indexed by AgentSkillsHub · Auto-synced every 8h

🔒 Is AgentEval safe to install? View the security audit →

About AgentEval

# AgentEval The .NET Evaluation Toolkit for AI Agents AgentEval is the comprehensive .NET toolkit for AI agent evaluation—tool usage validation, RAG quality m

agentagenticevalsevaluationsframeworknetred-teamingtestingworkflows

Quick Facts

Stars154
Forks16
LanguageC#
CategoryAgent Tool
LicenseMIT
Quality Score66.7678485646016/100
Open Issues4
Last Updated2026-10-03
Created2026-01-02
Platformsdotnet
Est. Tokens~25k

Compatible Skills

These tools work well together with AgentEval for enhanced workflows:

  • Lucene.Net.Analysis.PanGu — semantic(0.38)+complementary+same_lang+similar_pop+shared_platform (58%)
  • MCPSharp — semantic(0.16)+complementary+same_lang+similar_pop+shared_platform (51%)

AgentEval alternative? Top 6 similar tools

Looking for a AgentEval alternative? If you're comparing AgentEval with other agent tool tools, these 6 projects are the closest alternatives on Agent Skills Hub — ranked by topic overlap, star count, and community traction.

  • SimpleLLMFunc by NiJingzhe · ⭐ 77

    A simple and well-tailored LLM application framework that enables you to seamlessly integrate LLM capabilities

  • n8n-claw by freddy-schuetz · ⭐ 550

    OpenClaw-inspired autonomous AI agent built entirely in n8n. Adaptive RAG-powered memory, Skills via MCP templ

  • deep-seek by dzhng · ⭐ 518

    LLM powered retrieval engine designed to process a ton of sources to collect a comprehensive list of entities.

  • Autono by vortezwohl · ⭐ 213

    A ReAct-Based Highly Robust Autonomous Agent (Harness) Framework.

  • sunpeak by Sunpeak-AI · ⭐ 210

    Server-agnostic MCP testing framework and full-stack MCP App framework for ChatGPT Apps, Claude Connectors, an

  • antigravity-workflows by harikrishna8121999 · ⭐ 187

    Community-driven workflows for Antigravity AI. Like Claude Skills - reusable prompts and automation for AI cod

More Agent Tool Tools

Explore other popular agent tool tools:

View all Agent Tool tools →

Popular C# Agent Tools

Frequently Asked Questions

What is AgentEval?

AgentEval is AgentEval is the comprehensive .NET toolkit for AI agent evaluation—tool usage validation, RAG quality metrics, stochastic evaluation, and model comparison—built first for Microsoft Agent Framework (M. It is categorized as a Agent Tool with 154 GitHub stars.

What programming language is AgentEval written in?

AgentEval is primarily written in C#. It covers topics such as agent, agentic, evals.

How do I install or use AgentEval?

You can find installation instructions and usage details in the AgentEval GitHub repository at github.com/AgentEvalHQ/AgentEval. The project has 154 stars and 16 forks, indicating an active community.

What license does AgentEval use?

AgentEval is released under the MIT license, making it free to use and modify according to the license terms.

What are the best alternatives to AgentEval?

The top alternatives to AgentEval on Agent Skills Hub include SimpleLLMFunc, n8n-claw, deep-seek. Each offers a different approach to the same problem space — compare them side-by-side by stars, quality score, and community activity.

How this security grade is produced

Grades come from a rule-based scan built on the SlowMist agent-security taxonomy, covering 11 red-flag categories including credential harvesting, data exfiltration, and curl | sh installers. It is a first-layer scan, not a manual audit — we say so rather than overstate it.

The scale of the problem is documented independently: Liu et al. (2026), in a study of 31,132 agent skills, report that 26.1% contain security vulnerabilities. Our own full-catalog census is published as a citable open dataset.

Sources & who's responsible:

View on GitHub → Browse Agent Tool tools