autoresearch — security grade SAFE, quality 68/100

Security audit verdict: SAFE · quality 68/100

No red flags found in any of the 11 categories — no credential harvesting, no data exfiltration, no curl-pipe-shell installer. Scanned against the SlowMist agent-security taxonomy, refreshed every 8 hours. Full audit →

by uditgoenka · Claude Skill · ★ 5.8k

Last updated: · Indexed by AgentSkillsHub · Auto-synced every 8h

🔒 Is autoresearch safe to install? View the security audit →

About autoresearch

Claude Autoresearch Skill Turn Claude Code into a relentless improvement engine. Based on Karpathy's autoresearch — the principle that constraint + mechanical metric + autonomous iteration = compounding gains. What Is This? A Claude Code skill that makes Claude iterate autonomously on ANY task with a measurable outcome — like Karpathy's autoresearch, but generalized beyond ML. The loop: LOOP (FOREVER or N times): Review current state + git history + results log Pick the next change (based on what worked, what failed, what's untried) Make ONE focused change Git commit (before verification) Run mechanical verification (tests, benchmarks, scores) If improved → keep. If worse → git revert. If crashed → fix or skip. Log the r

aiautonomous-agentautoresearchclaudeclaude-codeiterationkarpathyproductivityskill

Quick Facts

Stars5,821
Forks442
LanguageShell
CategoryClaude Skill
LicenseMIT
Quality Score68.309535786371/100
Open Issues1
Last Updated2026-08-12
Created2026-03-13
Platformsclaude-code, cli
Est. Tokens~20k

autoresearch alternative? Top 6 similar tools

Looking for a autoresearch alternative? If you're comparing autoresearch with other claude skill tools, these 6 projects are the closest alternatives on Agent Skills Hub — ranked by topic overlap, star count, and community traction.

  • claude-code-tips by ykdojo · ⭐ 10.1k

    45+ tips for getting the most out of Claude Code, from basics to advanced - includes a custom status line scri

  • openpencil by ZSeven-W · ⭐ 6.1k

    The world's first open-source AI-native vector design tool and the first to feature concurrent Agent Teams. De

  • commands by wshobson · ⭐ 2.6k

    Archived. The original Claude Code slash commands, replaced by the plugin marketplace at wshobson/agents.

  • codex-autoresearch by leo-lilinxiao · ⭐ 2.5k

    Codex Autoresearch Skill — A self-directed iterative system for Codex that continuously cycles through: modify

  • autocontext by greyhaven-ai · ⭐ 1.3k

    a recursive self-improving harness designed to help your agents (and future iterations of those agents) succee

  • ai-guide by liyupi · ⭐ 20.6k

    程序员鱼皮的 AI 资源大全 + Vibe Coding 零基础教程,分享 OpenClaw 保姆级教程、大模型玩法(DeepSeek / GPT / Gemini / Claude / GLM)、最新 AI 资讯、Pr

More Claude Skill Tools

Explore other popular claude skill tools:

View all Claude Skill tools →

Popular Shell Agent Tools

Frequently Asked Questions

What is autoresearch?

autoresearch is Claude Autoresearch Skill — Autonomous goal-directed iteration for Claude Code. Inspired by Karpathy's autoresearch. Modify → Verify → Keep/Discard → Repeat forever.. It is categorized as a Claude Skill with 5.8k GitHub stars.

What programming language is autoresearch written in?

autoresearch is primarily written in Shell. It covers topics such as ai, autonomous-agent, autoresearch.

How do I install or use autoresearch?

You can find installation instructions and usage details in the autoresearch GitHub repository at github.com/uditgoenka/autoresearch. The project has 5.8k stars and 442 forks, indicating an active community.

What license does autoresearch use?

autoresearch is released under the MIT license, making it free to use and modify according to the license terms.

What are the best alternatives to autoresearch?

The top alternatives to autoresearch on Agent Skills Hub include claude-code-tips, openpencil, commands. Each offers a different approach to the same problem space — compare them side-by-side by stars, quality score, and community activity.

How this security grade is produced

Grades come from a rule-based scan built on the SlowMist agent-security taxonomy, covering 11 red-flag categories including credential harvesting, data exfiltration, and curl | sh installers. It is a first-layer scan, not a manual audit — we say so rather than overstate it.

The scale of the problem is documented independently: Liu et al. (2026), in a study of 31,132 agent skills, report that 26.1% contain security vulnerabilities. Our own full-catalog census is published as a citable open dataset.

Sources & who's responsible:

View on GitHub → Browse Claude Skill tools