Project File
Autonomous R&D Lab
A governed research loop: objective → plan → agents → experiments → evidence → critique → revision, with claim gates and immutable audit logs.
Overview
OBSERVEDThe lab engine turns human objectives into governed agent loops: nine specialist roles with scoped permissions, deterministic task planning, sandboxed experiments, tool budgets, and evidence-backed memory consolidation. Every claim passes Skeptic, Statistical Evaluator, and Auditor review gates before it counts.
Honest Status
EXPERIMENTALThe MVP is feature-complete and self-validates with a toy experiment playbook (dcsil demo), with a React mission-control dashboard and CLI. The default provider is still a mock: live-LLM research quality is the open frontier, and container-grade isolation is not done.
The Governed Research Loop
Every stage writes to an immutable audit log; claims must pass review gates to become results.
Current Limitations
- −No end-to-end autonomous discovery on a real research question yet.
- −Mock provider by default; live providers are wired but lightly exercised.
Next Steps
- +First supervised closed-loop run on a real (small) lab question.
- +Use Remy as the execution arm for experiment stages.