Skip to the AI comparison desk

AI Lab // Six-Agent Evaluation Bench

AI Lab: Choose the workflow before you choose the AI.

Hyde Workshop’s AI Lab helps visitors compare current AI models, workspace agents, coding platforms, and open standards by practical workflow value, technical limitations, security risk, and required human oversight. Every evaluation uses our A.M.A.N.D.A. governance framework — where Amy defines the mission, Maya verifies the evidence, Aria audits quality, Nancy maps the integration route, Dawn tests defenses, and Astra prepares the final human-owned decision.

  • Editorial review:
  • Audience: creators, students, developers, WordPress users, and small businesses
  • Indicators: Hyde Workshop editorial signals—not independent benchmarks or certifications
  • Availability, pricing, plans, regions, and product capabilities can change
  • Mission before model
  • Primary-source evidence
  • Minimum permissions
  • Observable handoffs
  • Human approval
  • Rollback before scale
AI Aria wearing the Hyde Workshop laboratory uniform as the lead analytical auditor in the AI Lab
AI Aria // Capability is tested against evidence, accessibility, compatibility, performance, security, and the user’s actual workflow.

A.M.A.N.D.A. Diagnostic Gateway // Discovery vs. Calibration

Discover systems here. Calibrate workflows in A.M.A.N.D.A.

AI Lab (Multi-System Comparison Desk)

Discover and compare AI systems, tools, standards, workflow value, limitations, and governance considerations.

A.M.A.N.D.A. Workflow Diagnostic

Analyze the visitor’s specific workflow friction and organize a governed route before tools, permissions, data, or automation are connected.

Human Approval Checkpoint: A.M.A.N.D.A. evaluation framework roles provide structured decision support, security analysis, and route mapping. They do not independently publish code, deploy software, make purchases, or execute system changes. Mister Hyde retains final human approving authority for all operational deployments.

A.M.A.N.D.A. // Named Evaluation Responsibilities

Every comparison passes through six accountable roles.

The six agents are Hyde Workshop editorial and workflow characters. They make the evaluation method visible; they do not claim that six autonomous systems publish, deploy, purchase, delete, or make final decisions on this website.

AI Amy wearing the Hyde Workshop black laboratory uniform

Adaptive Mission Agent

AI Amy

Defines the user’s exact mission, expected result, data boundary, human owner, and approval checkpoint.

Primary output: Mission brief and acceptance criteria.
AI Maya wearing the Hyde Workshop black laboratory uniform

Multi-Source Mapping Agent

AI Maya

Maps official documentation, release dates, availability limits, conflicting claims, and missing evidence.

Primary output: Evidence map and source trail.
AI Aria wearing the Hyde Workshop black laboratory uniform

Analytical Audit Agent

AI Aria

Audits capability, accuracy, accessibility, SEO, performance, compatibility, and practical workflow quality.

Primary output: Technical findings and quality checks.
AI Nancy wearing the Hyde Workshop black laboratory uniform

Neural Navigation Agent

AI Nancy

Maps apps, tools, identities, connectors, handoffs, escalation points, and the correct route through the site.

Primary output: Controlled integration and routing plan.
AI Dawn wearing the Hyde Workshop black laboratory uniform

Defense and Diagnostics Agent

AI Dawn

Tests prompt-injection exposure, privacy, credentials, permissions, unsafe actions, monitoring, and rollback.

Primary output: Cain-risk register and defense controls.
AI Astra wearing the Hyde Workshop black laboratory uniform

Accountability and Approval Agent

AI Astra

Combines the findings, records uncertainty, defines conditions, and prepares the final human-owned route.

Primary output: Adopt, pilot, monitor, restrict, or reject decision brief.

Current Innovation Watch // From Assistants to Operations

The market is converging on agents that run longer, touch more systems, and require stronger control.

These reports are current editorial starting points. Each one links to an official source and is paired with the agent responsible for asking the next control question.

AI AmyRepeatable work

Workspace agents are turning one-off prompts into shared operating procedures.

OpenAI workspace agents in Research Preview for workspace plans can run longer workflows, use approved tools, follow schedules, and ask for approval. Amy’s question is whether the job, owner, inputs, and acceptance criteria are clear before the agent is published.

Open official OpenAI source (opens in a new tab)
AI MayaEvidence workbenches

Research systems are producing auditable artifacts, not only answers.

Anthropic’s current direction spans Claude Code, Claude Tag, and research workbenches. Maya checks which sources, channels, files, tools, and dates shaped the result—and which evidence remains unavailable.

Open official Anthropic source (opens in a new tab)
AI AriaProduction evaluation

Agent quality is becoming a continuous testing and observability problem.

AWS AgentCore uses production traces, recommendations, evaluations, and controlled tests to find recurring or silent failures. Aria requires defined test cases and human approval before changes are promoted.

Open official AWS source (opens in a new tab)
AI NancyBackground execution

Managed agents are adding remote tools, asynchronous work, and durable state.

Google Managed Agents in developer preview and the Interactions API add background execution, remote MCP connections, custom functions, and credential refresh. Nancy maps every connection, handoff, timeout, and escalation path.

Open official Google source (opens in a new tab)
AI DawnIdentity + authorization

Agent security is shifting toward named identities and explicit authority.

NIST’s AI Agent Standards Initiative focuses on interoperability, identity, authentication, authorization, and secure operation. Dawn treats every tool connection and agent identity as a new attack and accountability surface.

Open official NIST source (opens in a new tab)
AI AstraCreative orchestration

Creative agents are coordinating multi-step work while preserving editable outcomes.

Adobe is extending its creative agent across Firefly and Creative Cloud applications. Astra checks ownership, provenance, factual accuracy, brand constraints, accessibility, export settings, and final human approval.

Open official Adobe source (opens in a new tab)

Interactive Comparison Desk // Static Evidence + Governed Live Review

Run a six-agent review before installation, connection, automation, or deployment.

Product descriptions and official links remain in static HTML for visitors and crawlers. Search, filters, sorting, and the editorial review stay lightweight. When a logged-in Hyde Workshop administrator starts a live review, the same panel securely routes the mission through the server-side A.M.A.N.D.A. bridge without exposing API credentials to the browser.

14 current systems and standardsShowing all 14 entries.
Shared workflow agents Controlled team pilot

OpenAI Workspace Agents

Shared, cloud-powered agents in Research Preview for ChatGPT Business, Enterprise, Edu, and Teachers plans, operating across tools, schedules, ChatGPT, and Slack with admin enablement.

Lead: AI AmyCain exposure: 47
Coding + knowledge work Adopt with review

OpenAI Codex

Agentic coding and knowledge-work environment integrated across ChatGPT, IDEs, and CLI, using repository context, `Agents.md` rules, and automated verification loops.

Lead: AI AriaCain exposure: 52
Team context + coding Controlled team pilot

Claude Sonnet 5 + Claude Code + Claude Tag

Composite suite combining Claude Sonnet 5 reasoning, Claude Code agentic CLI/IDE tools, and Claude Tag for Slack channel collaboration in Enterprise/Team beta.

Lead: AI MayaCain exposure: 51
Managed agent runtime Governance-first pilot

Google Managed Agents + Interactions API

Managed cloud agents in developer preview with isolated Linux sandboxes, background task execution, remote MCP support, and unified Interactions API state management.

Lead: AI NancyCain exposure: 58
Enterprise control plane Adopt as control layer

Microsoft Agent 365

Enterprise control plane providing centralized registry, identity, security, data protection, and governance for AI agents across platforms.

Lead: AI DawnCain exposure: 37
Device + developer platform Pilot on supported devices

Apple Intelligence + Siri AI + Xcode Developer Tools

Device-integrated intelligence in developer beta with on-device models, Private Cloud Compute, App Intents, and agentic coding tools in Xcode for supported Apple platforms.

Lead: AI AstraCain exposure: 46
Production optimization Adopt for observability

Amazon Bedrock AgentCore Optimization

Production trace analysis, failure insight recommendations, versioned configuration bundles, and controlled A/B testing for AWS AI agents.

Lead: AI AriaCain exposure: 41
Creative orchestration Creative pilot

Adobe Firefly Creative Agent

Conversational creative agent orchestrating multi-step workflows across Creative Cloud apps and chat platforms with Content Credentials provenance under human review.

Lead: AI AstraCain exposure: 40
Repository agents Adopt with repository controls

GitHub Copilot Agents

Autonomous repository agents for issue delegation, background execution, pull request creation, and MCP integration under human code review.

Lead: AI AriaCain exposure: 45
WordPress-native AI Staging-only pilot

WordPress 7 AI Client + Connectors

Core PHP developer infrastructure and centralized connector settings in WordPress 7.0 for provider-agnostic model integration; Hyde Workshop recommends staging-first testing and API key controls.

Lead: AI AmyCain exposure: 60
Tool connection standard Guardrail before connection

Model Context Protocol — 2026 Specification

Open standard for AI tool and data integration featuring a stateless core architecture, extensions framework, user elicitation, and AAIF governance.

Lead: AI DawnCain exposure: 72
Agent handoff standard Guardrail before handoff

Agent2Agent Protocol

Open Linux Foundation standard for inter-agent discovery, capability cards, stateful task handoffs, and cross-framework collaboration.

Lead: AI NancyCain exposure: 69
Identity + security reference Adopt as governance reference

NIST AI Agent Standards Initiative

U.S. government research initiative establishing voluntary frameworks for AI agent identity, non-human authorization, interoperability, and threat defense.

Lead: AI DawnCain exposure: 27
Digital provenance Adopt for provenance

C2PA Content Credentials 2.4

Open provenance specification for certifying digital asset origin, edit history, ingredients, and cryptographic assertions.

Lead: AI MayaCain exposure: 29

Hyde Workshop Route Map // Signal to Controlled Action

Use the smallest page that answers the reader’s next question.

The redesigned site separates news, comparison, implementation controls, verdicts, tutorials, and project intake so visitors do not have to decode one oversized page.

What changed?

Forge News

Current product releases, standards, risks, and official-source signals.

Open Forge News
Which system fits?

AI Lab

Compare products, platforms, standards, mission fit, evidence, control, and Cain exposure.

Use this comparison desk
How should it be controlled?

AI Forge

Define identity, permissions, evidence, monitoring, approval, cost, recovery, and rollback.

Enter AI Forge
What is the final route?

A.M.A.N.D.A. Analysis

Convert Cain risk and Abel value into an accountable human-owned recommendation.

Open A.M.A.N.D.A.
Where is the full tutorial?

Blog Archive

Read long-form WordPress, security, SEO, AI, and workflow implementation guides.

Browse the Blog Archive
How do I get help?

Contact Hyde Workshop

Submit a workflow audit, correction, technical question, or project brief through the official form.

Open Contact

AI Lab frequently asked questions

Are the editorial indicators independent benchmarks?

No. They are Hyde Workshop comparison signals for mission fit, evidence maturity, control readiness, and Cain exposure. They are not certifications, guarantees, permanent rankings, or substitutes for testing the current product in your own environment.

Are the six agents autonomous publishers on HydeWorkshop.com?

No. Amy, Maya, Aria, Nancy, Dawn, and Astra are branded workflow roles that explain the review method. Human editors remain responsible for reporting, testing, corrections, publication, deployment, purchases, privacy decisions, and other consequential actions.

Why are protocols and standards listed beside commercial products?

Products increasingly depend on protocols, identity, authorization, tool connections, agent handoffs, and provenance. A useful comparison must examine the operating environment around the model—not only the visible assistant.

Should WordPress 7 AI connectors be enabled directly on a live site?

No initial experiment should begin on production. Use staging, verify the exact plugin feature, provider, credentials, capability checks, REST permissions, cost, logs, privacy impact, performance, backup, and rollback before enabling a narrow production use case.

Which actions should remain behind a human approval gate?

Publishing, deployment, deletion, purchasing, spending, customer contact, legal or medical claims, account and permission changes, access to sensitive information, privacy decisions, and irreversible actions should remain behind an authorized human checkpoint.

AI Agent enlarged portrait preview

A.M.A.N.D.A. Agent Portrait

AI Agent

Hyde Workshop agent role