Start from zero
Similar tasks are decomposed, explored, and debugged again—without access to prior execution lessons.
Agent Experience Intelligence
AEG captures what agents tried, what failed, what recovered, and what worked—then turns those runs into evidence-backed guidance for future tasks.
The problem
Execution experience disappears across sessions: useful failures, recovery steps, cost, and validation are rarely captured in a form the next agent can trust.
Similar tasks are decomposed, explored, and debugged again—without access to prior execution lessons.
Attempts, dead ends, recovery steps, latency, and token cost vanish even when the final patch survives.
A plausible tip is not a verified experience. Agents need provenance, constraints, outcomes, and reproducibility.
The operating loop
AEG focuses first on coding agents. It is an evidence layer—not a generic agent marketplace, an orchestrator, or a replacement for observability platforms.
Intent, context, steps, skills, artifacts, failures, recovery, outcome, and cost.
Sanitize and structure the reusable lesson without raw private work.
Attach objective checks, limitations, and reproducible provenance.
Retrieve relevant lessons and show why they apply—or abstain.
Offer a guarded capsule, then capture the new outcome and feedback.
Evidence, not theater
AEG’s public artifacts are useful precisely because their boundaries are visible.
Test whether relevant prior experience reduces retries, token usage, and execution time without lowering task success.
Source: public repair lab results. This does not establish success-rate improvement, generalized transfer, or product-market fit.
What AEG preserves
The reusable unit contains enough evidence to judge a recommendation without retaining raw conversations or proprietary code.
Intent, context, decomposition, attempted skills and tools.
Failures, rejected paths, guardrails, and the steps that recovered.
Outcome, objective oracle, provenance, reproducibility, and limitations.
Attempts, commands, tests, latency, and tokens when reliably observable.
Next validation milestone