podium

Maestro
vs QM from YC

A source-linked comparison across the complete Podium ADE rubric. 5 criteria have sourced findings for both products; 5 were also cross-checked by a second agent for both. 19 rows have at least one sourced finding, including 19 with an agent-cross-checked finding. No score or winner is inferred.

Each finding carries its own verification status, scope and dates. Humans approve the rubric, templates and prototype; they do not review every finding. An agent source cross-check is not a hands-on product test. A missing sourced finding is not evidence that a capability is absent.

Rubric
0.3 · 51 criteria
Last source check
2026-09-23
Next source check due
2026-10-07
Agent cross-checked for both
5 criteria
Maestro source checks
2026-09-23
QM from YC source checks
2026-09-23

Access and agent choice

ACC-01

Client access

Judgment rule. List desktop OS and minimum versions, iOS, Android, browser, mobile browser, CLI and editor extensions separately. For each client, record start, steer, approve, review and resume actions, notifications and plan restrictions. Test moving the same task from desktop to phone; distinguish an installable mobile app from a mobile website and from status-only monitoring.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

Agent cross-checked partial

QM documents web and Slack workspaces; dedicated mobile apps and cross-device task action parity remain unverified.

Scope
Surfaces: QM web workspace, QM Slack plugin, QM core and deployment CLI · Version: ec86d8b603a6cc07b21aa4d55359b88c425b61ef · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
unknown
Delivery
native
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-10-23
  • This is a source pass; unfilled atomic questions remain unknown and no shared scenario was run.
9 atomic answers
  • Desktop os [unknown] Desktop os: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Mobile apps [unknown] Mobile apps: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Browser [supported] Web workspace is documented. Sources: qm:readme.
  • Mobile browser [unknown] Mobile browser: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Cli [supported] Deployment CLI is documented; agent-operation coverage is separate. Sources: qm:readme.
  • Editor extensions [unknown] Editor extensions: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Task actions [unknown] Task actions: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Notifications [unknown] Notifications: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Cross device continuity [unknown] Cross device continuity: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.

Sources: qm:readme

ACC-02

Agent and model choice

Judgment rule. List supported harnesses and model providers separately, including custom adapters. Record whether the product embeds a harness, controls its sessions through an API, or only launches its terminal command. Verify at least two different harnesses before describing mixed-harness operation as observed.

Maestro

Agent cross-checked partial

Multiple agent CLIs are documented.

Scope
Surfaces: desktop, ssh-remote · Version: Official sources retrieved 2026-09-16 · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
unknown
Delivery
unknown
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-10-23
  • Other atomic questions remain unverified; no shared hands-on scenario was run.
5 atomic answers
  • Harnesses [supported] Claude Code, Codex, OpenCode, Droid; Copilot beta. Sources: maestro:intro.
  • Model providers [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Custom adapters [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Session integration [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Mixed harness operation [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.

Sources: maestro:intro

QM from YC

Agent cross-checked partial

QM documents selectable Pi, OpenCode, Codex and Claude Code runtimes.

Scope
Surfaces: QM web workspace, QM Slack plugin, QM core and deployment CLI · Version: ec86d8b603a6cc07b21aa4d55359b88c425b61ef · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
unknown
Delivery
native
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-10-23
  • This is a source pass; unfilled atomic questions remain unknown and no shared scenario was run.
5 atomic answers
  • Harnesses [supported] Pi, OpenCode, Codex and Claude Code. Sources: qm:readme.
  • Model providers [unknown] Model providers: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Custom adapters [unknown] Custom adapters: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Session integration [supported] Selected runtime operates through QM core. Sources: qm:readme.
  • Mixed harness operation [unknown] Mixed harness operation: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.

Sources: qm:readme

ACC-03

Subscriptions, API keys and vendor credits

Judgment rule. Record bring-your-own agent subscription, bring-your-own API account, vendor credits and local models as separate routes. List supported plans, models/modes, multiple-account switching and what works without a vendor account or token purchase. Existing subscription support cannot be inferred from API-key support. Record required extra payments and documented provider restrictions.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

ACC-04

Onboarding and project prerequisites

Judgment rule. Measure fresh install to first successful task, with manual steps and failures. Test a local directory without Git, a local Git repo without a remote, and a GitHub-backed repo separately. Record whether account, project, host, credential and permission prerequisites are explained before accepting an unusable prompt. Evaluate managed-cloud setup and first-class BYO VPS installer/onboarding independently of hidden settings or a DIY shell recipe.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

Agent cross-checked partial

Deployment instructions describe an agent-guided installer targeting the operator cloud account.

Scope
Surfaces: QM core and deployment CLI · Version: ec86d8b603a6cc07b21aa4d55359b88c425b61ef · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
unknown
Delivery
native
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-10-23
  • This is a source pass; unfilled atomic questions remain unknown and no shared scenario was run.
7 atomic answers
  • First task setup [unknown] First task setup: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • No git [unknown] No git: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Local git without remote [unknown] Local git without remote: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Hosted git [unknown] Hosted git: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Prerequisite guidance [supported] Deployment instructions name infrastructure, sign-in and credentials. Sources: qm:readme.
  • Managed cloud setup [unknown] Managed cloud setup: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Byo vps onboarding [supported] qm init materializes deployment guidance for Fly or AWS; installation effort was not measured. Sources: qm:readme.

Sources: qm:readme

ACC-05

Programmable access

Judgment rule. Inventory actual CLI, API, SDK, MCP and webhook operations for create, inspect, steer, stop and review. Record authentication and API stability. Having an MCP client that uses external tools is different from exposing the product to other agents over MCP.

Maestro

Agent cross-checked partial

Source package exposes CLI executables.

Scope
Surfaces: desktop, ssh-remote · Version: Official sources retrieved 2026-09-16 · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
source-only
Delivery
unknown
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-10-23
  • Command behavior not audited.
8 atomic answers
  • Cli [supported] maestro-cli and maestro-p. Sources: maestro:code.
  • Api [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Sdk [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Mcp server [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Webhooks [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Operations [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Authentication [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Stability [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.

Sources: maestro:code

QM from YC

Agent cross-checked partial

Authenticated project routes and swarm APIs exist, with narrower purposes than a full issue API.

Scope
Surfaces: QM core and deployment CLI · Version: ec86d8b603a6cc07b21aa4d55359b88c425b61ef · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
unknown
Delivery
native
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-10-23
  • This is a source pass; unfilled atomic questions remain unknown and no shared scenario was run.
8 atomic answers
  • Cli [unknown] Cli: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Api [supported] Projects and swarm HTTP operations are documented or implemented. Sources: qm:project-routes, qm:swarms.
  • Sdk [unknown] Sdk: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Mcp server [unknown] Mcp server: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Webhooks [unknown] Webhooks: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Operations [unknown] Operations: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Authentication [supported] Project routes bind capability principal to actor; swarm calls require scoped agent tokens. Sources: qm:project-routes, qm:swarms.
  • Stability [unknown] Stability: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.

Sources: qm:project-routes, qm:swarms

ACC-06

External tools and extensions

Judgment rule. Record MCP client support, plugins, custom tools, skill/instruction import, integration authentication and permission scope. Name which component provides each connection and whether setup is product-managed or inherited from the underlying harness. Verify one external-tool call and its access boundary where available.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

ACC-07

Client implementation

Judgment rule. For each desktop/mobile client, record native platform UI, Electron, Tauri, React Native, web/PWA or unknown, plus hybrids. Require official technical docs or pinned source/build evidence. Tauri is not synonymous with fully native UI. Keep implementation choice separate from measured responsiveness and usability.

Maestro

Agent cross-checked partial

Desktop implementation uses Electron.

Scope
Surfaces: desktop · Version: Official sources retrieved 2026-09-16 · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
source-only
Delivery
unknown
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-12-22
  • Pinned package manifest; packaged behavior untested.
4 atomic answers
  • Desktop implementation [supported] Electron with React/Vite build. Sources: maestro:code.
  • Mobile implementation [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Hybrids [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Implementation evidence [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.

Sources: maestro:code

QM from YC

Agent cross-checked partial

The web workspace uses Lit according to its package manifest.

Scope
Surfaces: QM web workspace · Version: ec86d8b603a6cc07b21aa4d55359b88c425b61ef · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
source-only
Delivery
native
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-12-22
  • This is a source pass; unfilled atomic questions remain unknown and no shared scenario was run.
4 atomic answers
  • Desktop implementation [unknown] Desktop implementation: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Mobile implementation [unknown] Mobile implementation: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Hybrids [supported] Browser-rendered Lit UI; this does not establish a native mobile application. Sources: qm:web-package.
  • Implementation evidence [supported] Pinned web UI package declares Lit. Sources: qm:web-package.

Sources: qm:web-package

Planning and context

PLAN-01

Intake and external tracker relationship

Judgment rule. Record local creation, prompt intake, issue import and chat/Git triggers. Identify the authoritative tracker, stable issue IDs, one-way or two-way sync, supported edits and operations requiring a trip back to Linear/GitHub/Jira. Importing or attaching a ticket is not native issue management. Answer PLAN-06 and PLAN-07 separately.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

PLAN-02

Durable task plans

Judgment rule. Record task plans and acceptance criteria, editors, revision history and links to implementation. Test amending a task while work is running, restarting and revisiting it later. A task-specific starting brief, reusable memory and a persistent product specification are different objects; product specs are judged in PLAN-08.

Maestro

Agent cross-checked partial

Plans are editable Markdown documents.

Scope
Surfaces: desktop, ssh-remote · Version: Official sources retrieved 2026-09-16 · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
unknown
Delivery
unknown
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-10-23
  • Issue tracker semantics and persistent product specs not established.
7 atomic answers
  • Plans [supported] Checkbox tasks in .md files. Sources: maestro:autorun.
  • Acceptance criteria [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Editors [supported] Manual checkbox edits outside active runs. Sources: maestro:autorun.
  • Revisions [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Amendment [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Restart persistence [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.

Sources: maestro:autorun

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

PLAN-03

Hierarchy, dependencies and evolving scope

Judgment rule. Record epics, arbitrary parent/child depth, blockers, readiness, dependency enforcement and state propagation. Start with a simple interactive task, grow it into an epic and retain its identity, discussion, history and outputs. Verify blockers actually prevent dispatch. Record whether a rigid preselected workflow is required, and whether changing scope requires recreating work.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

PLAN-04

Shared context and memory

Judgment rule. List repository instructions, skills, project docs and cross-session/cross-harness memory. Record scope, provenance, citation, correction/deletion and stale-context controls. Test what survives a new session or colleague takeover. A notes page or shared file is not automatically a maintained spec or managed memory.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

Agent cross-checked partial

Memory and skills have personal/shared scope in the documented QM model.

Scope
Surfaces: QM web workspace, QM Slack plugin, QM core and deployment CLI · Version: ec86d8b603a6cc07b21aa4d55359b88c425b61ef · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
unknown
Delivery
native
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-10-23
  • This is a source pass; unfilled atomic questions remain unknown and no shared scenario was run.
10 atomic answers
  • Instructions [unknown] Instructions: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Skills [supported] Skills can be shared by grant and promoted by admins. Sources: qm:readme.
  • Project docs [unknown] Project docs: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Memory [supported] Scoped memory is documented. Sources: qm:readme.
  • Scope [supported] Personal and room scopes. Sources: qm:readme.
  • Provenance [unknown] Provenance: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Correction [unknown] Correction: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Deletion [unknown] Deletion: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Freshness [unknown] Freshness: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Takeover continuity [unknown] Takeover continuity: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.

Sources: qm:readme

PLAN-05

Architecture and API awareness

Judgment rule. Record code search, repository indexing, external knowledge and multimodal inputs, with source freshness and access scope. Ask an agent to use an existing API under a documented architecture rule, then propose a conflicting change. Check retrieval, citation, conflict detection and review-time enforcement separately. A large context window or an instruction file alone does not prove architectural consistency.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

PLAN-06

Native issue management and idea backlog

Judgment rule. Enumerate create, edit, discuss, prioritize, label, assign, search, filter, archive, reopen and project/epic views within the product. Save an idea without creating a running agent, refine it with another person, and explicitly mark it ready or dispatch it later. Record separate task/session entities, metadata depth, multiple sessions per task, durable state transitions and whether an external tracker remains necessary.

Maestro

Agent cross-checked partial

Documents can exist before dispatch.

Scope
Surfaces: desktop, ssh-remote · Version: Official sources retrieved 2026-09-16 · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
unknown
Delivery
unknown
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-10-23
  • Document parking is not proof of native issue management.
17 atomic answers
  • Create [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Edit [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Discuss [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Prioritize [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Label [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Assign [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Filter [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Archive [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Reopen [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Project epic views [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Idea parking [supported] Saved Markdown checklist. Sources: maestro:autorun.
  • Explicit dispatch [supported] Select document then Run/Go. Sources: maestro:autorun.
  • Task session separation [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Metadata [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Durable states [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • External tracker requirement [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.

Sources: maestro:autorun

QM from YC

Agent cross-checked partial

Inspected task storage is session/run-oriented. It is evidence of task metadata, not proof of a human-managed idea backlog.

Scope
Surfaces: QM core and deployment CLI · Version: ec86d8b603a6cc07b21aa4d55359b88c425b61ef · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
source-only
Delivery
native
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-10-23
  • Session/run task records cannot establish native issue management without human UI and shared-operation evidence.
17 atomic answers
  • Create [unknown] Create: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Edit [unknown] Edit: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Discuss [unknown] Discuss: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Prioritize [unknown] Prioritize: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Label [unknown] Label: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Assign [unknown] Assign: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Filter [unknown] Filter: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Archive [unknown] Archive: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Reopen [unknown] Reopen: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Project epic views [unknown] Project epic views: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Idea parking [unknown] Idea parking: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Explicit dispatch [unknown] Explicit dispatch: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Task session separation [unknown] Task session separation: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Metadata [partial] Task records contain title, sessionId, originRunId and state; issue labels/priorities/backlog are unverified. Sources: qm:task-model.
  • Durable states [partial] Status/event storage interface exists; production persistence and UI parity were not exercised. Sources: qm:task-model.
  • External tracker requirement [unknown] External tracker requirement: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.

Sources: qm:task-model

PLAN-07

Agent-native issue operations

Judgment rule. Verify an agent can read/create/update the same durable issue IDs and fields humans use, claim/assign work, change state, create dependencies/children and report blockers/results through documented CLI/API/MCP. Check that human views reflect actual writes with attribution, permissions and conflict handling. Distinguish shared authoritative records, synchronized external records, agent-private metadata and prompt-only conventions.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

PLAN-08

Living product specifications

Judgment rule. Record persistent requirements across tasks/releases, editable by humans and agents with attributable revisions and approval rules. Verify requirement-to-task/implementation links, retrieval during later tasks, drift detection and consideration during review/quality control. Amend a requirement after one task ships, then implement another. A one-time spec-to-plan workflow, per-task brief, memory store or unreferenced document does not establish a living product spec.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

Coordination and automation

COORD-01

Parallel work

Judgment rule. Record independently running sessions and tasks, documented quotas and observed concurrency. Describe how a user sees running, waiting, failed and finished work. Several chat tabs are not proof that runs execute concurrently.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

COORD-02

Ownership and dispatch

Judgment rule. Record assignment to humans or agents, claims, queues, routing and duplicate-work prevention. Test two workers attempting the same task. Distinguish a display label, an advisory lock and enforced exclusive ownership.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

COORD-03

Inspectable delegation and communication

Judgment rule. Record parent/child roles, model/harness/effort choice, messages, result return and cross-harness teams. Distinguish product orchestration, provider-native subagents, installed orchestration skills and scripts. Record activation, auth and experimental flags. Verify actual child identity, live progress, retained results and ability to reopen or follow up after completion. A final sentence claiming two reviewers ran is not proof. Test the documented invocation as well as the shared natural-language request.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

Agent cross-checked partial

Swarms document inspectable worker identities and messaging; human follow-up is limited in worker web viewers.

Scope
Surfaces: QM web workspace, QM core and deployment CLI · Version: ec86d8b603a6cc07b21aa4d55359b88c425b61ef · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
unknown
Delivery
native
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-10-23
  • This is a source pass; unfilled atomic questions remain unknown and no shared scenario was run.
12 atomic answers
  • Roles [supported] Arbitrary worker role/context metadata. Sources: qm:swarms.
  • Model harness effort [unknown] Model harness effort: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Messages [supported] Durable addressed swarm messages. Sources: qm:swarms.
  • Results return [unknown] Results return: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Cross harness [unknown] Cross harness: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Orchestration owner [supported] QM swarm API coordinates ordinary QM sessions. Sources: qm:swarms.
  • Activation [unknown] Activation: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Authentication [unknown] Authentication: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Live progress [unknown] Live progress: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Retained results [supported] Worker transcripts and run records persist when production stores are configured. Sources: qm:swarms.
  • Followup [partial] Ordinary web messages to worker transcripts are read-only; pending approvals can be allowed or denied. Sources: qm:swarms.
  • Invocation [unknown] Invocation: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.

Sources: qm:swarms

COORD-04

Unattended triggers

Judgment rule. Record schedules, event triggers, recurring tasks and completion conditions. Name where the scheduler runs, overlap handling, retries and whether closing the client stops it. A desktop cron recipe is a manual integration unless the product owns it.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

Agent cross-checked partial

QM documents background work through crons, watches and webhooks.

Scope
Surfaces: QM core and deployment CLI · Version: ec86d8b603a6cc07b21aa4d55359b88c425b61ef · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
unknown
Delivery
native
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-10-23
  • This is a source pass; unfilled atomic questions remain unknown and no shared scenario was run.
8 atomic answers
  • Schedules [supported] Crons. Sources: qm:readme.
  • Events [supported] Watches and inbound webhooks. Sources: qm:readme.
  • Recurring tasks [unknown] Recurring tasks: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Completion conditions [unknown] Completion conditions: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Scheduler location [unknown] Scheduler location: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Overlaps [unknown] Overlaps: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Retries [unknown] Retries: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Client independence [unknown] Client independence: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.

Sources: qm:readme

COORD-05

Intervention and attention

Judgment rule. Record pause, stop, resume, steering, notifications, pending decisions and escalation. Test an in-run amendment and a harmless approval prompt in both chat and terminal modes. Record silent waits, hidden CLI-only questions and whether stop terminates execution while preserving history. Require discoverable errors rather than inferring success from silence.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

COORD-06

Fleet and epic management UI

Judgment rule. Use a disclosed fixture of 12 task/agent states across several epics, including blocked, failed, awaiting input, reviewing and completed work. Locate parent/child relationships, dependencies, attention needs, history and workload without reading every transcript. Record navigation steps and errors. Boards, graphs, waterfalls, timelines or another interface can qualify; no required visual form. State whether the fixture was simulated or used live agents.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

COORD-07

Workflow flexibility and orchestration visibility

Judgment rule. Compare a quick interactive change with a growing epic. Record mandatory workflow/preset selection, natural-language configuration, per-agent model/effort control, changes during execution and whether the UI exposes what actually ran. Test a coordinator with two reviewer children; capture roles, relationships and results during work and after completion. Document discoverability and auth friction without judging an unsupported magic prompt as proof of missing capability.

Maestro

Agent cross-checked partial

Runs can follow checklists or goals.

Scope
Surfaces: desktop, ssh-remote · Version: Official sources retrieved 2026-09-16 · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
unknown
Delivery
unknown
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-10-23
  • Other atomic questions remain unverified; no shared hands-on scenario was run.
10 atomic answers
  • Quick task [supported] Goal-driven iterations without preset checklist. Sources: maestro:autorun.
  • Growing epic [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Presets [supported] Reusable Playbooks. Sources: maestro:autorun.
  • Natural language [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Agent configuration [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • In run changes [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Live roles [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Retained relationships [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Discoverability [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Auth friction [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.

Sources: maestro:autorun

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

Execution and reliability

RUN-01

Laptop-independent execution

Judgment rule. List local, user-owned VPS/remote host, self-hosted service, vendor cloud and hybrid modes. Record provisioning, payer, compute resources, region, idle/max lifetime and disconnection rules. Close/disconnect the originating client and continue from another device. A hosted UI or keep-awake setting is not evidence of remote execution. Record managed-cloud convenience separately from BYO setup effort in ACC-04.

Maestro

Agent cross-checked partial

User-provided SSH execution is documented.

Scope
Surfaces: ssh-remote · Version: Official sources retrieved 2026-09-16 · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
unknown
Delivery
unknown
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-10-23
  • Local app remains control center; laptop-off autonomy not established.
12 atomic answers
  • Local [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Byo remote [supported] SSH host, key and working directory settings. Sources: maestro:ssh.
  • Self hosted [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Managed cloud [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Hybrid [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Provisioning [supported] Import SSH config; test connection. Sources: maestro:ssh.
  • Payer [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Resources [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Region [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Expiry [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Client disconnect [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Other device continuation [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.

Sources: maestro:ssh

QM from YC

Agent cross-checked partial

QM is documented for operator-owned cloud deployment; third-party hosting is a separate offering.

Scope
Surfaces: QM core and deployment CLI · Version: ec86d8b603a6cc07b21aa4d55359b88c425b61ef · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
unknown
Delivery
native
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-10-23
  • This is a source pass; unfilled atomic questions remain unknown and no shared scenario was run.
12 atomic answers
  • Local [unknown] Local: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Byo remote [unknown] Byo remote: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Self hosted [supported] Operator-owned deployment. Sources: qm:readme.
  • Managed cloud [unknown] Managed cloud: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Hybrid [unknown] Hybrid: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Provisioning [supported] Deployment CLI targets Fly or AWS. Sources: qm:readme.
  • Payer [supported] Operator supplies its cloud account. Sources: qm:readme.
  • Resources [unknown] Resources: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Region [unknown] Region: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Expiry [unknown] Expiry: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Client disconnect [unknown] Client disconnect: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Other device continuation [unknown] Other device continuation: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.

Sources: qm:readme

RUN-02

Workspace and process isolation

Judgment rule. Record shared checkout, branch, worktree, container, VM and OS sandbox separately. Name filesystem, process, network and secret boundaries. Validate code separation with two changes; do not claim hostile-code containment without separate security evidence.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

Agent cross-checked partial

Swarm workers receive separate blank computers; optional shared forums do not copy parent files.

Scope
Surfaces: QM core and deployment CLI · Version: ec86d8b603a6cc07b21aa4d55359b88c425b61ef · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
unknown
Delivery
native
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-10-23
  • This is a source pass; unfilled atomic questions remain unknown and no shared scenario was run.
12 atomic answers
  • Checkout [unknown] Checkout: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Branch [unknown] Branch: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Worktree [unknown] Worktree: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Container [unknown] Container: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Vm [unknown] Vm: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Os sandbox [unknown] Os sandbox: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Filesystem [partial] Private blank worker computers; an explicit forum can share a disk. Sources: qm:swarms.
  • Process [unknown] Process: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Network [unknown] Network: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Secrets [unknown] Secrets: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Code separation [partial] Parent files are not automatically copied; shared forums require write coordination. Sources: qm:swarms.
  • Security boundary [unknown] Security boundary: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.

Sources: qm:swarms

RUN-03

Recovery, expiry and continuity

Judgment rule. Test client disconnect, worker crash, service restart and documented sandbox expiry safely. Separate transcript reload, process reconnection, task resumption and state reconstruction. Record lost context, duplicated actions, manual repair and ability to continue a multi-day epic after compute replacement. Short sandbox lifetimes are reported with their continuity behavior, not automatically treated as failure.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

Agent cross-checked partial

Production durability requires configured persistence; defaults must not be described as durable automatically.

Scope
Surfaces: QM core and deployment CLI · Version: ec86d8b603a6cc07b21aa4d55359b88c425b61ef · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
unknown
Delivery
native
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-10-07
  • This is a source pass; unfilled atomic questions remain unknown and no shared scenario was run.
12 atomic answers
  • Client disconnect [unknown] Client disconnect: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Worker crash [unknown] Worker crash: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Service restart [partial] README states in-memory sessions vanish on restart unless Postgres session storage is configured. Sources: qm:readme.
  • Expiry [unknown] Expiry: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Transcript reload [unknown] Transcript reload: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Process reconnect [unknown] Process reconnect: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Task resume [unknown] Task resume: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • State reconstruction [unknown] State reconstruction: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Context loss [unknown] Context loss: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Duplicate actions [unknown] Duplicate actions: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Manual repair [unknown] Manual repair: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Multi day continuity [unknown] Multi day continuity: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.

Sources: qm:readme

RUN-04

Development environment reliability

Judgment rule. Record dependency setup, node_modules/cache handling, snapshots, services, port allocation and secret injection. Start two workspaces with the same service requirements; note installation duplication, port collisions and repair steps. Repeat on a clean remote host. Record which problems the product handles automatically and which the user must diagnose.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

RUN-05

Limits and failure handling

Judgment rule. Record time, turn, token and spend limits; timeout, retry and provider-failure behavior. Test one controlled failure and a small safe budget. Record whether limits are warnings, enforced stops, or external billing controls.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

RUN-06

Truthful status and operational statistics

Judgment rule. Record logs, traces, failure reasons, usage exports and statistics by task, agent, project and time. Induce a worker failure and an approval wait: check agent, task and board state agree and remain correct after reconnect. Distinguish working, blocked, failed, idle, ready for review and delivered. For speed/success claims disclose fixture, model, hardware, attempts and failures. UI responsiveness needs observed interaction timings; never infer a framework from lag.

Maestro

Agent cross-checked partial

Statistics include per-agent and run views.

Scope
Surfaces: desktop, ssh-remote · Version: Official sources retrieved 2026-09-16 · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
unknown
Delivery
unknown
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-10-23
  • Only recorded Maestro activity; historic completeness varies.
13 atomic answers
  • Logs [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Traces [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Failure reasons [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Usage export [supported] JSON or CSV export. Sources: maestro:usage.
  • Task stats [supported] Auto Run attempts/completions. Sources: maestro:usage.
  • Agent stats [supported] Query/time charts and drilldowns. Sources: maestro:usage.
  • Project stats [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Time stats [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Failure status [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Approval status [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Reconnect consistency [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Performance method [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.
  • Responsiveness [unknown] Initial official-source pass did not establish this question; no hands-on scenario was run.

Sources: maestro:usage

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

RUN-07

Per-task API and compute cost

Judgment rule. Record visible model/API spend per task, session, child and epic, including aggregation across harness changes. Distinguish billed charges, estimates, tokens, subscription allowance and unavailable amounts; never render unknown subscription marginal cost as zero. Reconcile totals with provider evidence where available. Record compute fees separately, refresh delay, budget alerts, export and drill-down from summary statistics.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

RUN-08

Compute-aware scheduling and verification

Judgment rule. Record concurrency limits, queues, CPU/RAM limits, backpressure, per-host resource awareness and reuse/deduplication of checks. Queue multiple browser end-to-end tests under a declared safe capacity; verify excess work queues or throttles rather than overloading the host. Inspect cancellation, fairness, recovery and whether overlapping epics share limits. Do not run an uncontrolled load test.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

Human collaboration and control

TEAM-01

Live human collaboration

Judgment rule. Record shared visibility, issue discussion, comments, presence and simultaneous co-prompting as independent capabilities. Use two named accounts to steer the same agent and record input ordering/conflict handling, attribution and permission controls. Read-only links and shared issues do not establish co-steering. Record team/free-tier restrictions and whether the original owner must stay online.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

Agent cross-checked partial

QM documents people collaborating with agents in shared rooms; simultaneous prompt ordering remains untested.

Scope
Surfaces: QM web workspace, QM Slack plugin · Version: ec86d8b603a6cc07b21aa4d55359b88c425b61ef · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
unknown
Delivery
native
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-10-23
  • This is a source pass; unfilled atomic questions remain unknown and no shared scenario was run.
11 atomic answers
  • Visibility [supported] Personal and shared scopes are documented. Sources: qm:readme.
  • Discussion [supported] Shared channels, group messages and projects. Sources: qm:readme.
  • Comments [unknown] Comments: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Presence [unknown] Presence: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Coprompting [unknown] Coprompting: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Ordering [unknown] Ordering: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Conflicts [unknown] Conflicts: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Attribution [unknown] Attribution: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Permissions [unknown] Permissions: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Plan restrictions [unknown] Plan restrictions: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Owner online [unknown] Owner online: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.

Sources: qm:readme

TEAM-02

Roles and access boundaries

Judgment rule. Record organization/project/task roles, membership lifecycle, SSO and provisioning where offered. Test unauthorized access and revocation with consenting test accounts. Describe documented versus observed enforcement, including which plan provides it.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

Agent cross-checked partial

Project membership controls appear in inspected routes.

Scope
Surfaces: QM core and deployment CLI · Version: ec86d8b603a6cc07b21aa4d55359b88c425b61ef · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
source-only
Delivery
native
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-10-23
  • This is a source pass; unfilled atomic questions remain unknown and no shared scenario was run.
9 atomic answers
  • Organization roles [unknown] Organization roles: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Project roles [partial] Mutations can return forbidden; complete role matrix and revocation tests are pending. Sources: qm:project-routes.
  • Task roles [unknown] Task roles: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Membership [supported] Add/remove project-member routes. Sources: qm:project-routes.
  • Sso [unknown] Sso: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Provisioning [unknown] Provisioning: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Unauthorized access [unknown] Unauthorized access: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Revocation [unknown] Revocation: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Plan scope [unknown] Plan scope: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.

Sources: qm:project-routes

TEAM-03

Approvals and temporary access

Judgment rule. Record command, tool, network, secret and delivery approvals, approver identity and scope lifetime. Demonstrate allow, deny and revocation, including temporary tunnels or shared access after its thread/session ends. Separate default full access, available safer modes and product-enforced policy from instructions the model is expected to follow.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

TEAM-04

Decision and activity history

Judgment rule. Record attributed actions, approvals, edits, discussion and decision links. Check whether history persists across users and sessions, can be exported, and can be altered. Use the word audit only with a defined retention and integrity model.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

TEAM-05

Review experience and guidance

Judgment rule. Evaluate locating pending reviews, understanding a change, navigating evidence, commenting, requesting changes and finding the next shipping action. Measure steps/time and missed evidence on a shared fixture. Record keyboard/narrow-screen usability, context switching, clarity of agent roles and preservation of review context. Generic diff/PR access and a guided review workflow are different capabilities. Link high-volume review findings to SHIP-06.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

TEAM-06

Colleague takeover and continuity

Judgment rule. Have a second authorized person take over ongoing work with the same task, session history, decisions, artifacts, permissions and runtime context. Test after the first person disconnects. Record ownership transfer, attribution and lost context. Separately record transferring to another model/harness; a forked transcript or shared link alone is not full human takeover.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

Verification and delivery

SHIP-01

Retained artifacts and outputs

Judgment rule. Record diffs, documents, screenshots, recordings, previews and test/reviewer outputs. Check for a searchable/browsable artifact gallery or equivalent retrieval, linked to task and exact revision, usable after sessions end. Reopen a completed child reviewer and retrieve its full findings. Check whether internal prompts or tool metadata unintentionally appear in the produced application or other user-facing output. In-chat attachments, persistent review evidence and a cross-task gallery get distinct answers.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

SHIP-02

Verification and spec gates

Judgment rule. Record tests, lint, types, CI, independent review and persistent-spec conformance checks as required versus optional. Cause a failed test and a conflict with an existing product requirement. Verify review catches the conflict and delivery blocks where claimed. An agent saying checks passed is not a passing gate; record configured and enforced policy separately.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

SHIP-03

Git and pull request workflow

Judgment rule. Record branches/worktrees, commits, PR creation, review comments, CI response, rebase and conflict handling. Complete one fixture change through a draft PR. Distinguish an integrated workflow from the agent running Git commands without product state.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

SHIP-04

Merge and deployment authority

Judgment rule. Record who can approve, merge, release and deploy, plus checks against the exact revision. Verify the documented authority boundary without publishing real changes. PR creation, merge and production deployment are separate capabilities.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

SHIP-05

Completion evidence and rollback

Judgment rule. Record task-to-commit/PR/build/deployment links, delivery status, retained evidence and rollback or revert support. Change the candidate revision after review and examine whether approval/evidence is invalidated. Agent exit alone does not prove delivery.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

SHIP-06

Batch review and high-volume shipping

Judgment rule. Return to a queue of at least six completed fixture tasks and review a subset. Record summaries, grouping, risk/size filters, retained reviewer findings, artifact navigation, request-changes flow, approval of exact revisions, dependency-aware merge/release queues and clear next actions. Check stale approvals invalidate after edits and duplicate checks respect RUN-08. Distinguish basic diff/PR actions from a coordinated review-and-delivery queue; do not require Podium terminology.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

Ownership, privacy and economics

OWN-01

Source and license

Judgment rule. Record license and revision per component, including client, server and hosted extensions. Distinguish open source, source-available and closed components. Repository visibility is not a license classification or proof of hosted parity.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

Agent cross-checked partial

Public repository carries MIT; hosted extensions and parity remain unverified.

Scope
Surfaces: QM web workspace, QM Slack plugin, QM core and deployment CLI · Version: ec86d8b603a6cc07b21aa4d55359b88c425b61ef · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
unknown
Delivery
native
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-12-22
  • This is a source pass; unfilled atomic questions remain unknown and no shared scenario was run.
6 atomic answers
  • Client license [supported] Root MIT license, subject to component notices. Sources: qm:license.
  • Server license [unknown] Server license: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Hosted extensions [unknown] Hosted extensions: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Source visibility [supported] Public repository. Sources: qm:license.
  • Source revision [supported] ec86d8b603a6cc07b21aa4d55359b88c425b61ef Sources: qm:license.
  • Hosted parity [unknown] Hosted parity: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.

Sources: qm:license

OWN-02

Hosting and exit

Judgment rule. Record self-hosting scope, required vendor services, data export/import, standard formats and account cancellation behavior. Verify what remains usable without the vendor account. Exported code alone does not mean tasks, context and history are portable.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

Agent cross-checked partial

QM supports organization-owned deployments and forks.

Scope
Surfaces: QM core and deployment CLI · Version: ec86d8b603a6cc07b21aa4d55359b88c425b61ef · Plan: Not specified · OS: Not specified · Execution modes: Not specified · Region: Not specified · Harness: Not specified
Availability
unknown
Delivery
native
Checked
2026-09-23
Agent cross-check
2026-09-23
Check again by
2026-10-23
  • This is a source pass; unfilled atomic questions remain unknown and no shared scenario was run.
8 atomic answers
  • Self hosting [supported] Documented deployment into the operator cloud account. Sources: qm:readme.
  • Vendor dependencies [unknown] Vendor dependencies: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Export [unknown] Export: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Import [unknown] Import: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Formats [unknown] Formats: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Cancellation [unknown] Cancellation: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • Account independence [unknown] Account independence: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.
  • History portability [unknown] History portability: Initial primary-source pass did not establish this question. No installed-product or shared-scenario test was run.

Sources: qm:readme

OWN-03

Data handling and telemetry

Judgment rule. Record where code, prompts, credentials, logs and artifacts go, retention/deletion, training use, residency and subprocessors. Separately record telemetry disclosure, collection defaults, opt-in/opt-out controls and observed first-run prompts. No consent dialog is an observation, not proof of no telemetry, secret telemetry or unlawful processing. Separate orchestration-vendor and model-provider policies.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

OWN-04

Price and included usage

Judgment rule. Record currency, billing interval, seats/seat minimums, workspace charges, included usage, credits, trial conditions, overages, taxes treatment and separate model/compute costs. Map billing routes to ACC-03. Cite public prices or dated quotes. Per-seat versus pooled billing is a fact, not a universal quality judgment.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

OWN-05

Scenario cost and controls

Judgment rule. Price stated workloads for one, two and five developers, with fixed tasks, concurrency, tokens and compute-hours. The two-person scenario models Till and Michael collaborating. Include subscription, platform, model and compute charges; separate fixed, variable, included and unknown costs. No invented average task cost. Link cost visibility and reconciliation to RUN-07.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

OWN-06

Availability and support

Judgment rule. Record GA/beta/access restrictions, support channels, documented response commitments, release activity, status history and deprecation/export options. Funding, stars and popularity can be context but are not reliability scores. An archived repository does not prove the hosted service is discontinued.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

OWN-07

Enterprise assurance

Judgment rule. Record each claimed assurance separately: ISO standard/certificate scope, issuer and validity; SOC 2 report type, covered service and audit period; report/certificate access and verification date. Distinguish public documentation, vendor assertion, independently accessible evidence and unavailable evidence. A logo, enterprise pricing tier, SSO or a provider certification does not prove this product is certified or audited. Report controls in TEAM-02 separately; no legal-compliance verdict.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

OWN-08

Useful free tier and evaluation access

Judgment rule. Distinguish an ongoing free tier, time-limited trial, credits-only trial, open-source self-hosting costs, paid-only access and gated early access. Map which shared scenarios work for free, especially two-person collaboration, BYO subscriptions, remote compute, mobile, issue management and review. Record card requirement, usage caps and whether meaningful evaluation needs payment or sales approval.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

OWN-09

Pricing transparency

Judgment rule. Record publicly stated units, included limits, overage rates, compute/model pass-through or markup, minimum commitment, cancellation and sales-only unknowns. Use transparent, partially specified, quote-only or unknown, with named missing fields. A quote-only enterprise price is a disclosed procurement constraint, not proof of dishonesty.

Maestro

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

QM from YC

No sourced public finding not-published

No publishable sourced finding is available for this criterion. This does not mean the product lacks the capability.

Sources

qm:readme

QM from YC: README.md

Documents shared channels/projects, scope-owned memory, web and Slack, selectable agent runtimes, self-host deployment and Postgres configuration.

Product
qm
Kind
documented
Locator
README.md
Source date
Not stated
Retrieved
2026-09-23
maestro:intro

Maestro overview

Parallel coding agents and Auto Run docs; runtime-duration anecdotes are not benchmarks.

Product
maestro
Kind
documented
Locator
Introduction
Source date
Not stated
Retrieved
2026-09-23
maestro:code

Maestro source: package.json

Electron scripts package macOS/Windows/Linux; exposes maestro-cli and maestro-p binaries.

Product
maestro
Kind
source-inspected
Locator
Named definitions and implementation in pinned file
Source date
Not stated
Retrieved
2026-09-23
Commit
b3186828921680431f7179f202e0099998a321b5
Path
package.json
qm:project-routes

QM from YC: src/api/routes/projects.ts

Project HTTP routes support list/create/rename and membership mutations; capability callers are bound to their actor identity. Project APIs are not issue CRUD.

Product
qm
Kind
source-inspected
Locator
src/api/routes/projects.ts
Source date
Not stated
Retrieved
2026-09-23
Commit
ec86d8b603a6cc07b21aa4d55359b88c425b61ef
Path
src/api/routes/projects.ts
qm:swarms

QM from YC: docs/swarms.md

Documents durable worker sessions, scoped messaging, human and agent APIs, provisioning retries and finite budgets. Worker web transcripts are read-only for normal messages, with approval controls.

Product
qm
Kind
documented
Locator
docs/swarms.md
Source date
Not stated
Retrieved
2026-09-23
qm:web-package

QM from YC: plugins/web-ui/package.json

Web UI package declares Lit; this supports the browser implementation classification, not mobile-client parity.

Product
qm
Kind
source-inspected
Locator
plugins/web-ui/package.json
Source date
Not stated
Retrieved
2026-09-23
Commit
ec86d8b603a6cc07b21aa4d55359b88c425b61ef
Path
plugins/web-ui/package.json
maestro:autorun

Auto Run + Playbooks

Markdown checklists or iterative goals drive fresh agent sessions.

Product
maestro
Kind
documented
Locator
Spec/Goal toggle; Creating Tasks; Multiple Documents
Source date
Not stated
Retrieved
2026-09-23
qm:task-model

QM from YC: src/tasks/task-store.ts

Task interface ties IDs to sessions and runs, with pending/in_progress/completed/skipped/failed states and status-event APIs. It does not establish an independent human issue backlog.

Product
qm
Kind
source-inspected
Locator
src/tasks/task-store.ts
Source date
Not stated
Retrieved
2026-09-23
Commit
ec86d8b603a6cc07b21aa4d55359b88c425b61ef
Path
src/tasks/task-store.ts
maestro:ssh

SSH Remote Execution

Commands run on configured SSH hosts while local Maestro remains control center.

Product
maestro
Kind
documented
Locator
Overview; Configuring SSH remotes
Source date
Not stated
Retrieved
2026-09-23
maestro:usage

Usage Dashboard

Dashboard describes token/cost tables but later says tokens/costs are not aggregated.

Product
maestro
Kind
documented
Locator
Groups; Tokens; What is NOT tracked
Source date
Not stated
Retrieved
2026-09-23
qm:license

QM from YC: LICENSE

Repository root license is MIT; component-specific exceptions and hosted parity were not audited.

Product
qm
Kind
documented
Locator
LICENSE
Source date
Not stated
Retrieved
2026-09-23