SMF Works

Findings · Benchmarks · Open Tools

Research

We test, document, and build — with honesty about what works and what doesn't. Our research lives at the AI Clearinghouse, our practitioner-facing research site with 900+ articles, benchmarks, guides, and open tools.

What We Study

Research Areas

Agent Architecture

Hermes skills, memory systems, observability, delegation patterns, and multi-agent orchestration. We build and document in the open.

Read architecture deep dives →
📊

Evaluation & Benchmarks

Model testing, eval harnesses, and benchmark results. We run the tests and publish the numbers — including when they surprise us.

See benchmark results →
🛡️

Governed Autonomy

Praxis — our governed autonomous colleague experiment. An AI agent operating with real consequences under human oversight. Early preview, honest about rough edges.

Learn about Praxis →
🔧

Open Tools

Hermes skills, plugins, the media replay manifest, Mnemosyne, SMF Swarm, HyperFrames. Tools we build for ourselves, shipped for others to use.

Browse on GitHub →