Claude Octopus 🤖 Agent Open source
Run multiple AI models against one task and surface disagreements before you ship
- GitHub stars
- 4.1k
- Stars this week
- +40
- Forks
- 391
- Licence
- MIT
- Last push
- 2026-10-02
- Maintainer
- nyldn
claude plugin install octo@nyldn-pluginsThird-party subagents & agents run with your permissions. Read the source before installing, and prefer pinned versions.
Works with
About Claude Octopus
What it does
Claude Octopus runs a task through multiple AI models and surfaces where they disagree before you ship. Claude Code handles the ordinary path; Octopus stays dormant until you explicitly invoke an /octo:* command, then escalates to multi-model workflows: adversarial review, structured provider debates, consensus gating, and a "Dark Factory" mode that takes a spec and runs research, definition, development, and delivery autonomously. It integrates a dozen external providers (Codex, Antigravity CLI, Copilot, Qwen, Ollama, Perplexity, OpenRouter, OrcaRouter, OpenCode, Cursor CLI, Grok, and Kimi Code) alongside the built-in Claude host, but needs none of them to start.
What is inside
The plugin ships dozens of specialized personas (security-auditor, backend-architect, and the like), a large set of slash commands, and reusable skill modules, plus engineering methods for architecture, TDD, and debugging. Workflows include a Double Diamond discover-to-deliver pipeline, a multi-LLM council with goal modes and adversarial styles, parallel workstreams in isolated git worktrees, and multi-model PR review with inline comments. Optional session memory integrations persist decisions across sessions. Diagnostics (doctor, capabilities, repair, security-audit) check and repair the installation offline.
Works with
Claude Code and Codex as plugins, Cursor as an MCP server, plus OpenCode and Factory Droid. External providers are added one at a time and run only inside explicit workflows.
How to install
claude plugin marketplace add https://github.com/nyldn/plugins.git
claude plugin install octo@nyldn-plugins
Maintenance and safety
MIT licensed and very actively developed. Nothing routes automatically: ordinary prompts are untouched unless you opt into a suggest or invoke router, and the invoke mode can start paid external-provider workflows and share prompt context with those providers, so provider-side data retention follows each account's policy. It is an independent project, not affiliated with Anthropic.
Who should use it
Developers who want cross-model second opinions and adversarial review on high-stakes work, and who are comfortable managing multiple provider credentials.
Pros
- Adversarial multi-model review and consensus gating on demand
- Dormant by default; ordinary prompts are untouched
Cons
- Invoke-mode routing can start paid external-provider workflows and share prompt context
- Large surface area with many providers to configure
Similar subagents & agents
All agent workflows & frameworks →Spec Kit 🤖 AgentFree
GitHub's toolkit for spec-driven development with AI coding agents
Task Master 🤖 AgentFree
AI task management that turns a PRD into tasks your coding agent works through
Fast Agent 🤖 AgentOpen source
Python framework for building, orchestrating and evaluating MCP-native AI agents
OpenSpec 🤖 AgentFree
Lightweight spec-driven development: agree on changes before the agent codes
CC Safety Net 🤖 AgentOpen source
Pre-execution guard that blocks destructive commands and secret access for coding agents
BMAD Method 🤖 AgentFree
Agile AI-driven development with analyst, PM, architect, developer and UX agents