jeeves — security grade SAFE, quality 60/100

Security audit verdict: SAFE · quality 60/100

No red flags found in any of the 11 categories — no credential harvesting, no data exfiltration, no curl-pipe-shell installer. Scanned against the SlowMist agent-security taxonomy, refreshed every 8 hours. Full audit →

by PostHog · AI Tool · ★ 400

Last updated: · Indexed by AgentSkillsHub · Auto-synced every 8h

🔒 Is jeeves safe to install? View the security audit →

About jeeves

Jeeves – Reasoning improves Jev-like decision models A reasoning Jev-style classifier with a diffusion drafter, trained with SFT and CISPO. Acknowledgements Inspired by Kev. Highlights A 9B Jev-like model (Qwen3.5-9B, LoRA, pointer head) that thinks before it decides, with a block-4 diffusion drafter and the full training code and train/dev/test data. Beats Kev-9B and Jev on test data it was never trained on (0.889 vs 0.822 and 0.857) and on JevBench's public tiers (0.935 vs 0.866 for Jev). Supports yes/no (), multiple-choice (), and rating () questions in the same request, through a Jev-compatible API. About 0.3 s per request without thinking and a 3.3 s median with it on one H100. Can be sped up by truncating chain length. Runs on CUDA (Hopper for the FP8 kernel). Problem Jev-like models give calibrated decision probabilities, but at low accuracy. A lot of pipelines therefore rely on a reasoning model as a fallback. Jeeves trains a Jev-like Qwen3.5-9B (LoRA and a pointer head) using CISPO t

Quick Facts

Stars400
Forks20
LanguagePython
CategoryAI Tool
LicenseMIT
Quality Score59.9661093905306/100
Open Issues5
Last Updated2026-10-01
Created2026-09-29
Platformspython
Est. Tokens~17k

jeeves alternative? Top 6 similar tools

Looking for a jeeves alternative? If you're comparing jeeves with other ai tool tools, these 6 projects are the closest alternatives on Agent Skills Hub — ranked by topic overlap, star count, and community traction.

  • EDRSilencer by netero1010 · ⭐ 1.9k

    A tool uses Windows Filtering Platform (WFP) to block Endpoint Detection and Response (EDR) agents from report

  • perf-map-agent by jvm-profiling-tools · ⭐ 1.7k

    A java agent to generate method mappings to use with the linux `perf` tool

  • shaping-skills by rjs · ⭐ 1.4k

    Skills I use with Claude for shaping

  • pivotnacci by blackarrowsec · ⭐ 726

    A tool to make socks connections through HTTP agents

  • MedgeClaw by xjtulyc · ⭐ 641

    Open-source AI research assistant for biomedicine — chat to run RNA-seq, drug discovery, clinical analysis, an

  • SciToolAgent by HICAI-ZJU · ⭐ 428

    SciToolAgent: A Knowledge Graph-Driven Scientific Agent for Multi-Tool Integration

More AI Tool Tools

Explore other popular ai tool tools:

View all AI Tool tools →

Popular Python Agent Tools

Frequently Asked Questions

What is jeeves?

jeeves is Jeeves – Reasoning improves Jev-like decision models. It is categorized as a AI Tool with 400 GitHub stars.

What programming language is jeeves written in?

jeeves is primarily written in Python.

How do I install or use jeeves?

You can find installation instructions and usage details in the jeeves GitHub repository at github.com/PostHog/jeeves. The project has 400 stars and 20 forks, indicating an active community.

What license does jeeves use?

jeeves is released under the MIT license, making it free to use and modify according to the license terms.

What are the best alternatives to jeeves?

The top alternatives to jeeves on Agent Skills Hub include EDRSilencer, perf-map-agent, shaping-skills. Each offers a different approach to the same problem space — compare them side-by-side by stars, quality score, and community activity.

How this security grade is produced

Grades come from a rule-based scan built on the SlowMist agent-security taxonomy, covering 11 red-flag categories including credential harvesting, data exfiltration, and curl | sh installers. It is a first-layer scan, not a manual audit — we say so rather than overstate it.

The scale of the problem is documented independently: Liu et al. (2026), in a study of 31,132 agent skills, report that 26.1% contain security vulnerabilities. Our own full-catalog census is published as a citable open dataset.

Sources & who's responsible:

View on GitHub → Browse AI Tool tools