datasets-for-start — security grade SAFE, quality 63/100

Security audit verdict: SAFE · quality 63/100

No red flags found in any of the 11 categories — no credential harvesting, no data exfiltration, no curl-pipe-shell installer. Scanned against the SlowMist agent-security taxonomy, refreshed every 8 hours. Full audit →

by pplonski · Agent Tool · ★ 72

Last updated: · Indexed by AgentSkillsHub · Auto-synced every 8h

🔒 Is datasets-for-start safe to install? View the security audit →

About datasets-for-start

📦 Datasets for Start A curated collection of simple, ready-to-use datasets for machine learning, data analysis, and tutorials. These datasets are designed to: ✅ be easy to load ✅ require minimal preprocessing ✅ work great for beginners and demos 🚀 Why this repo? This repository helps you: learn machine learning faster practice exploratory data analysis (EDA) build quick prototypes create tutorials and demos 👉 No heavy data cleaning — just start working with data. 🧠 Use with MLJAR Studio These datasets work perfectly with MLJAR Studio. MLJAR Studio is a desktop application designed for data science, combining AI and Python in one place. It lets users easily load data, build machine learning models, and generate reports without complex setup. It is especially beginner-friendly, helping users move from data to insights quickly while still giving advanced users full control. 👉 https://mljar.com/ 📊 Dataset Overview 🔵 Binary Classification tabular

ai-toolsautomldata-analysisdata-sciencedata-visualization

Quick Facts

Stars72
Forks80
CategoryAgent Tool
LicenseMIT
Quality Score62.9133265975638/100
Last Updated2026-08-20
Created2017-03-30
Est. Tokens~16k

Compatible Skills

These tools work well together with datasets-for-start for enhanced workflows:

datasets-for-start alternative? Top 6 similar tools

Looking for a datasets-for-start alternative? If you're comparing datasets-for-start with other agent tool tools, these 6 projects are the closest alternatives on Agent Skills Hub — ranked by topic overlap, star count, and community traction.

  • openwebui-extensions by Fu-Jie · ⭐ 303

    A collection of enhancements, plugins, and prompts for Open WebUI, developed and curated for personal use to e

  • osint-ai-guide by atlas-bear · ⭐ 60

    Comprehensive guide to AI applications in OSINT workflows and intelligence analysis

  • ClaudeR by IMNMV · ⭐ 337

    Connect RStudio to Claude Code, Codex, Gemini, and other LLM agents via MCP. Multi-agent orchestration, automa

  • llms-tools by PetroIvaniuk · ⭐ 327

    A list of LLMs Tools & Projects

  • datastoria by FrankChen021 · ⭐ 326

    AI-native ClickHouse console for your cluster diagnostics and query generation, optimization and data visualiz

  • naksha-studio by Adityaraj0421 · ⭐ 319

    A virtual design team for Claude Code, Cursor, Windsurf, Gemini CLI, and Copilot — 26 roles, 62 commands, 15,0

More Agent Tool Tools

Explore other popular agent tool tools:

View all Agent Tool tools →

Frequently Asked Questions

What is datasets-for-start?

datasets-for-start is A curated collection of simple, ready-to-use datasets for machine learning, data analysis, and tutorials.. It is categorized as a Agent Tool with 72 GitHub stars.

How do I install or use datasets-for-start?

You can find installation instructions and usage details in the datasets-for-start GitHub repository at github.com/pplonski/datasets-for-start. The project has 72 stars and 80 forks, indicating an active community.

What license does datasets-for-start use?

datasets-for-start is released under the MIT license, making it free to use and modify according to the license terms.

What are the best alternatives to datasets-for-start?

The top alternatives to datasets-for-start on Agent Skills Hub include openwebui-extensions, osint-ai-guide, ClaudeR. Each offers a different approach to the same problem space — compare them side-by-side by stars, quality score, and community activity.

How this security grade is produced

Grades come from a rule-based scan built on the SlowMist agent-security taxonomy, covering 11 red-flag categories including credential harvesting, data exfiltration, and curl | sh installers. It is a first-layer scan, not a manual audit — we say so rather than overstate it.

The scale of the problem is documented independently: Liu et al. (2026), in a study of 31,132 agent skills, report that 26.1% contain security vulnerabilities. Our own full-catalog census is published as a citable open dataset.

Sources & who's responsible:

View on GitHub → Browse Agent Tool tools