No red flags found in any of the 11 categories — no credential harvesting, no data exfiltration, no curl-pipe-shell installer. Scanned against the SlowMist agent-security taxonomy, refreshed every 8 hours. Full audit →
by Alqemist-labs · Agent Tool · ★ 62
Last updated: · Indexed by AgentSkillsHub · Auto-synced every 8h
🔒 Is ruby_llm-tribunal safe to install? View the security audit →
RubyLLM::Tribunal ⚖️ LLM evaluation framework for Ruby, powered by RubyLLM. Tribunal provides tools for evaluating and testing LLM outputs, detecting hallucinations, measuring response quality, and ensuring safety. Perfect for RAG systems, chatbots, and any LLM-powered application. Inspired by the excellent Tribunal library for Elixir. Features 🎯 Deterministic assertions - Fast, free evaluations (contains, regex, JSON validation...) 🤖 LLM-as-Judge - AI-powered quality assessment (faithfulness, relevance, hallucination detection...) 🔐 Safety testing - Toxicity, bias, jailbreak, and PII detection 🎭 Red Team attacks - Generate adversarial prompts to test your LLM's defenses 📊 Multiple reporters - Console, JSON, HTML, JUnit, GitHub Actions 🧪 Test framework integration - Works with RSpec and Minitest Installation Add to your Gemfile: ruby gem 'rubyllm-tribunal' Required: RubyLLM for LLM-as-judge evaluations gem 'rubyllm', ' 1.0' Optional: for embedding-based similarity (assertsimilar) gem 'neighbor', ' 0.6' Opt
| Stars | 62 |
| Forks | 2 |
| Language | Ruby |
| Category | Agent Tool |
| License | MIT |
| Quality Score | 74.3082874008476/100 |
| Last Updated | 2026-04-09 |
| Created | 2026-01-15 |
| Platforms | ruby |
| Est. Tokens | ~5k |
Explore other popular agent tool tools:
ruby_llm-tribunal is LLM evaluation framework for Ruby, powered by RubyLLM. Tribunal provides tools for evaluating and testing LLM outputs, detecting hallucinations, measuring response quality, and ensuring safety. Perfec. It is categorized as a Agent Tool with 62 GitHub stars.
ruby_llm-tribunal is primarily written in Ruby.
You can find installation instructions and usage details in the ruby_llm-tribunal GitHub repository at github.com/Alqemist-labs/ruby_llm-tribunal. The project has 62 stars and 2 forks, indicating an active community.
ruby_llm-tribunal is released under the MIT license, making it free to use and modify according to the license terms.
Grades come from a rule-based scan built on the SlowMist agent-security taxonomy, covering 11 red-flag categories including credential harvesting, data exfiltration, and curl | sh installers. It is a first-layer scan, not a manual audit — we say so rather than overstate it.
The scale of the problem is documented independently: Liu et al. (2026), in a study of 31,132 agent skills, report that 26.1% contain security vulnerabilities. Our own full-catalog census is published as a citable open dataset.
Sources & who's responsible: