Sep 20 edition/Reporting & analysis
AgentsBusinessSafetyInfrastructure

AgentsAutonomy & tool use

Hermes Super Kanban pitch packages multi-agent project orchestration, but proof of autonomous delivery remains thin

A promotional Hermes demo says one approved idea can trigger agent teams to plan, build and preview projects. Public documentation supports the Kanban-based orchestration architecture, but the reviewed evidence does not independently verify output quality, completion rates or deployment safety.

THE CORE IDEAS3 TAKEAWAYS
01

The linked demo presents Hermes Super Kanban as an approval-gated workflow: a user enters one idea, reviews a generated plan, assigns work to agent roles and sees outputs collected for preview. [3]

02

Project materials describe a real multi-agent Kanban architecture, including worker lanes, task lifecycle management, dispatcher-style execution, board tooling and durable collaboration features. [6] [7] [8] [10]

03

The reviewed research does not include independent benchmarks for the promoted workflow, while security-related sources raise concerns about default configuration risks and Kanban/Docker behavior that warrants sandboxing and review. [4] [11]

WHY IT MATTERS

Evidence supports Hermes as stateful agent orchestration rather than just a chat interface. The business implication is faster handoff from idea to task execution, but production use should depend on measured quality, controls and security testing.

Executive brief

The consequential fact is not that Hermes has Kanban—it is that the September 19 pitch claims a human-approved project board can spawn agent teams that plan, build, and “ship” from one sentence, but public evidence remains mostly vendor/community material, not independent evaluation. The transcript shows a demo narrative: idea intake, plan approval, researcher/designer/coder-style agent assignment, live board movement, memory logging, and gallery preview. Public Hermes docs support the underlying multi-agent Kanban architecture, worker lifecycle, SQLite-backed boards, and dispatcher model; they do not independently verify the marketed “finished project while you do nothing” outcome.

What changed and event timeline

  1. Security audit disclosed high-severity Hermes Agent risks

    Cloud Security Alliance later summarized an independent audit alleging critical/high findings in default configuration, including shell execution and persistent skill-injection vectors.

  2. Hermes Kanban setup guide documented v0.12+ access

    AI Profit Boardroom’s guide said Hermes v0.12 or later was required, with hermes kanban init, profile assignment, kanban create, live watching, comments, and output review.

  3. Durable multi-agent Kanban appeared in release materials

    Release notes described “durable multi-profile collaboration board,” multiple boards, worker handoffs, heartbeats, reclaim, retry budgets, and a hallucination gate.

  4. “Super Kanban Agent” story promoted approval-gated self-driving workflows

    The video transcript says users enter one idea, review a plan, approve it, then agents build while the board advances and saves outputs to a gallery.

Capabilities and access

Exact “Hermes Super Kanban Agent” version is not stated in the Reddit excerpt or transcript. Public setup material says Kanban requires Hermes v0.12+; release notes later describe durable multi-agent Kanban in v0.13.0. The transcript claims a human approval gate, memory log/vault, gallery, live preview, and agent roles such as researcher, designer, and coder.

Read the full section

Exact “Hermes Super Kanban Agent” version is not stated in the Reddit excerpt or transcript. Public setup material says Kanban requires Hermes v0.12+; release notes later describe durable multi-agent Kanban in v0.13.0. Access in the video is framed as an AI Profit Boardroom package: full zip file, walkthrough, prompts, and coaching, not a standalone independently reviewed release. The transcript claims a human approval gate, memory log/vault, gallery, live preview, and agent roles such as researcher, designer, and coder. Hermes Kanban Setup, 03:01

Technical analysis for researchers and developers

Documented Hermes Kanban is a dispatcher-based multi-agent work system, not merely a UI board. Workers report through kanban_* tools and reviewers gate completion. No public reproducible evaluation, benchmark suite, or ablation study verifies success rate, task quality, or autonomy claims for the “Super Kanban” workflow.

Read the full section

Documented Hermes Kanban is a dispatcher-based multi-agent work system, not merely a UI board. Official docs describe worker lanes with assignees, spawn mechanisms, per-task workspaces, environment variables, and lifecycle truth owned by the Kanban kernel: ready → running → blocked / done / archived. Workers report through kanban_* tools and reviewers gate completion. CLI docs describe per-board SQLite storage and many-board support. No public reproducible evaluation, benchmark suite, or ablation study verifies success rate, task quality, or autonomy claims for the “Super Kanban” workflow. Worker lanes, Kanban docs

Claims and evidence

  • Vendor-reported: One sentence can become a plan, agent assignments, build, gallery item, and live preview after human approval.
  • Documented by project materials: Hermes Kanban supports multi-profile workers, task lifecycle, handoffs, heartbeats, and board tools.
  • Independent/third-party: Hermex review reports mobile access to self-hosted Hermes sessions, workspaces, dispatcher runs, and Kanban boards, but does not validate “Super Kanban” build quality. Hermex review
Read the full section
  • Vendor-reported: One sentence can become a plan, agent assignments, build, gallery item, and live preview after human approval. Evidence is the promotional video transcript, not an independent test. 00:55, 04:06
  • Documented by project materials: Hermes Kanban supports multi-profile workers, task lifecycle, handoffs, heartbeats, and board tools. v0.13.0 release notes, Kanban docs
  • Independent/third-party: Hermex review reports mobile access to self-hosted Hermes sessions, workspaces, dispatcher runs, and Kanban boards, but does not validate “Super Kanban” build quality. Hermex review
  • No independent corroboration found: public sources do not provide benchmarked completion rates, defect rates, security posture for this packaged Agent OS, or reproducible demo artifacts.

Context and prior work

The story fits a broader shift from chat-based agents to stateful orchestration: durable boards, task graphs, worker identities, memory, and human review gates. The “Super Kanban” pitch appears to package those primitives for business workflows—especially SEO/content/app prototypes—rather than introduce a new foundation model.

Read the full section

The story fits a broader shift from chat-based agents to stateful orchestration: durable boards, task graphs, worker identities, memory, and human review gates. Hermes’ public docs position it as an autonomous agent with persistent memory, skills, tool gateway, messaging, MCP, and deployment options. The “Super Kanban” pitch appears to package those primitives for business workflows—especially SEO/content/app prototypes—rather than introduce a new foundation model. Hermes docs, Hermes Kanban Setup

Limitations, safety and contested findings

The strongest limitation is evidentiary: the story is commentary plus a promotional transcript. Safety concerns are concrete: CSA summarized an independent audit alleging severe default-configuration risks, and a GitHub issue reported Kanban/Docker behavior where a worker searched for kanban.db, consumed many tokens, and allegedly followed unrelated markdown content.

Read the full section

The strongest limitation is evidentiary: the story is commentary plus a promotional transcript. Public docs support the architecture, not the claimed outcome quality. Safety concerns are concrete: CSA summarized an independent audit alleging severe default-configuration risks, and a GitHub issue reported Kanban/Docker behavior where a worker searched for kanban.db, consumed many tokens, and allegedly followed unrelated markdown content. These claims argue for sandboxing, scoped workspaces, approval gates, logging, and prompt-injection testing before business deployment. CSA note, GitHub issue #71456

Business and practitioner implications

For leaders, the practical opportunity is compressing messy handoffs: idea capture, task decomposition, assignment, progress visibility, and artifact storage in one workflow. For developers, the value is less “robot worker” than auditable orchestration around fallible agents.

Read the full section

For leaders, the practical opportunity is compressing messy handoffs: idea capture, task decomposition, assignment, progress visibility, and artifact storage in one workflow. For developers, the value is less “robot worker” than auditable orchestration around fallible agents. Treat it as an internal automation framework: start with low-risk repeatable workflows, require human approval before execution and release, instrument every worker action, isolate credentials, pin workspaces, and compare outputs against human baselines before using it in client-facing or production operations.

Sources

Read the full section
FOLLOW THE EVIDENCE

The source trail.

Sources (11)
A LITTLE LESS NOISE. A LOT MORE CONTEXT.

Stay curious.
Follow the evidence.

Independent perspectives, the original sources, and room for the questions that don't have easy answers.

How we build the brief