Reddit User Tests Hermes AI Agent's Self-Learning Feature, Finds Critical Flaws

Hermes vs OpenClaw: A Practical Comparison
A Reddit user who has been using OpenClaw since the January 29 build tested Hermes AI agent to evaluate its self-learning capabilities. The user earns money using OpenClaw and considers it their primary tool.
What Hermes Actually Does
Hermes markets "self-learning" as its core differentiator from OpenClaw, but according to the user's testing:
- Hermes is not "self-learning" in the machine learning sense
- It uses markdown files as memory, similar to OpenClaw
- The "self-learning" refers to automatically creating skills without manual writing
- Skills = markdown files that are automatically generated
The Critical Problem: Self-Evaluation Loop
The user identified a major issue with Hermes' implementation:
- Hermes operates in a closed learning loop where it evaluates its own results
- It always thinks it did a good job, regardless of actual performance
- In a test pulling water test results from the Indiana DNR site, Hermes "jumbled up everything" but still thought it "kicked ass"
- When users manually edit skills to fix errors, Hermes' self-improvement feature overwrites those edits
Stability Claims Questioned
The user addresses stability comparisons between the two tools:
- Hermes has had 6 releases total
- OpenClaw has had 82 releases
- 3 of Hermes' releases "didn't even work"
- The user advises against claims of Hermes being more stable due to limited release history
Current State and Future
The Reddit user concludes that Hermes is currently "unusable to someone who knows how to use OpenClaw." However, they acknowledge the project could "turn out amazing" and plan to continue watching its development.
📖 Read the full source: r/openclaw
👀 See Also

SideX: A Tauri-Based Port of Visual Studio Code
SideX is a port of Visual Studio Code that replaces Electron with Tauri, using a Rust backend and the OS's native webview. The project claims the same architecture with 96% smaller size, with core editing and terminal functionality currently working.

Hawkeye Update Adds Swarm Orchestration, Remote Tasks, and Local Model Support
Hawkeye v1.0+ now supports multi-agent swarm orchestration, remote task queuing, and improved Ollama/LM Studio integration. The local-first AI agent flight recorder helps developers track what happens when agents work in repositories.

Claude Code Skill Delegates Coding to Mistral/DeepSeek: 57M Tokens Saved, 90-100% Cost Reduction
A Claude Code skill called vibe-skill delegates low-level coding to cheap models like Mistral or DeepSeek while keeping Claude's planning. After 254 runs over 10 days, it saved 57M tokens and achieved 90-100% cost savings with 98% success rate.

MCP Memory Gateway: An MCP Server for Persistent Memory in Claude Code
A developer built an MCP server called MCP Memory Gateway using Claude Code as the primary development tool. It provides Claude Code with persistent memory across sessions through feedback capture, prevention rules, and context injection.