No red flags found in any of the 11 categories — no credential harvesting, no data exfiltration, no curl-pipe-shell installer. Scanned against the SlowMist agent-security taxonomy, refreshed every 8 hours. Full audit →
by carloslfu · Codex Skill · ★ 412
Last updated: · Indexed by AgentSkillsHub · Auto-synced every 8h
🔒 Is slotstream safe to install? View the security audit →
slotstream Run a 105 GB AI model on a Mac that can't hold it. Slotstream runs Qwen3.8-Flash-Next, a 125-billion-parameter open model, on Macs with 16 to 64 GB of memory. It keeps most of the model on the SSD and loads the parts it needs as it writes. Our 48 GB M5 Pro measured 15.86 tokens per second at a 22 GB memory target (how it was measured). Chat with it, ask it about pictures, or code with it: starts Claude Code on the local model, and Codex, Pi, opencode and Hermes work the same way. Developers can connect their own apps through its Ollama-, OpenAI- and Anthropic-compatible APIs or its Swift library. After a one-time download it works offline, with no Python and no cloud account. The whole engine is one native Swift program on Apple's MLX and Metal; see Built native. Every published number has a recorded method, and the experiments that failed stay in the measurements. Get started · Speed · Guides · Get help I'm building Sevra on Slotstream: private, personal AI optimized for your computer
| Stars | 412 |
| Forks | 28 |
| Language | Swift |
| Category | Codex Skill |
| License | MIT |
| Quality Score | 64.2752115731681/100 |
| Open Issues | 4 |
| Last Updated | 2026-10-03 |
| Created | 2026-08-28 |
| Platforms | claude-code, cli, codex |
| Est. Tokens | ~19k |
These tools work well together with slotstream for enhanced workflows:
Looking for a slotstream alternative? If you're comparing slotstream with other codex skill tools, these 6 projects are the closest alternatives on Agent Skills Hub — ranked by topic overlap, star count, and community traction.
Up to 4× faster LLM decoding on Apple Silicon, lossless. Native MLX port of DeepSeek's DSpark & z-lab's DFlash
A private Claude-Code-style coding agent for Apple Silicon — run chat, code, and local model workflows on-devi
Native LLM inference server for Apple Silicon. OpenAI + Anthropic API compatible. No Python. Zig backend, Swif
High-performance OpenAI and Anthropic compatible LLM inference server for Apple Silicon. Native MLX, continuou
Frontier-class open models on a free Kaggle TPU v5e-8: GLM-5.3-Flash 320B MoE (~64 tok/s, our own JAX engine)
Synthetic Autonomic Mind - An AI assistant for everyone.
Explore other popular codex skill tools:
slotstream is Run a 105 GB AI model on a Mac that can't hold it. Slotstream streams Qwen3.8-Flash-Next (125B mixture of experts) from your SSD and caches the busiest experts in memory, so it runs on Macs with 16 to. It is categorized as a Codex Skill with 412 GitHub stars.
slotstream is primarily written in Swift. It covers topics such as apple-silicon, claude-code, inference-engine.
You can find installation instructions and usage details in the slotstream GitHub repository at github.com/carloslfu/slotstream. The project has 412 stars and 28 forks, indicating an active community.
slotstream is released under the MIT license, making it free to use and modify according to the license terms.
The top alternatives to slotstream on Agent Skills Hub include mlx-dspark, ovo-local-llm, mlx-serve. Each offers a different approach to the same problem space — compare them side-by-side by stars, quality score, and community activity.
Grades come from a rule-based scan built on the SlowMist agent-security taxonomy, covering 11 red-flag categories including credential harvesting, data exfiltration, and curl | sh installers. It is a first-layer scan, not a manual audit — we say so rather than overstate it.
The scale of the problem is documented independently: Liu et al. (2026), in a study of 31,132 agent skills, report that 26.1% contain security vulnerabilities. Our own full-catalog census is published as a citable open dataset.
Sources & who's responsible: