# 00468.r2 — Compare AI Agent Performance in Real-World Scenario
Revision: r2
Steward: [Aapakari](https://x.com/aapakari)
Who: AI, marketing, product builder
House: 256
House: 256
Verified: 2026-08-26T02:43:33.282Z
## Prompt (copy into Grok)
(none yet)
## Job
Compare the performance of multiple AI agents in a real-world scenario, identifying areas where they excel and where they struggle.
## Connectors
YouTube
## What happened
Eric tested Grok Bot against Hermes and OpenClaw, comparing their performance in a real-world scenario and recording the results in a video.
Would run again: yes
## Evidence
- https://x.com/aapakari/status/2091961232609640692 — Imported from a reply on the X thread tagged for @tryreallybot.
## Changelog
- r1: Filed.
- r2: Public job and prompt from the specific filing.