ADARD loop: build the recursive self-improvement research controller #15
Labels
No labels
bug
design
documentation
duplicate
enhancement
good first issue
help wanted
invalid
question
research
wontfix
No milestone
No project
No assignees
1 participant
Notifications
Due date
No due date set.
Dependencies
No dependencies set.
Reference
nsaspy/starintel-auto-research#15
Loading…
Add table
Add a link
Reference in a new issue
No description provided.
Delete branch "%!s()"
Deleting a branch is permanent. Although the deleted branch may continue to exist for a short time before it actually gets removed, it CANNOT be undone in most cases. Continue?
Objective
Make the ADARD loop the primary experimental focus of
starintel-auto-researchbecause it has the strongest recursive self-improvement potential.ADARD must convert the current durable Org-roam research workflow into an executable loop that can propose, test, compare, retain, and reuse improvements to its own research and development process.
Core loop
RSI target
The loop should improve its own:
Every claimed self-improvement must be tied to a reproducible run, a baseline, metrics, artifacts, and an evidence trail.
Initial deliverable
Create
STAR-RESEARCH-PIPELINE-001 Adaptive ADARD Loop and Evaluation Harnesscovering only:Semantic retrieval, full prototype execution, and learned planning policies should remain dependent designs unless required for the minimal loop.
Candidate lifecycle
QUEUED -> EXPLORING -> PROPOSING -> VIRTUAL_REVIEW -> PROTOTYPING -> EVALUATING -> PROMOTED | MERGED | PRUNED | BLOCKEDMultiple research candidates may exist simultaneously. The existing single promoted implementation-design slot remains unchanged.
Required metrics
At minimum, track:
Promotion rules
A candidate may be promoted only when:
Unsupported central claims are a veto. Failed and pruned candidates remain searchable negative evidence.
First experiments
Acceptance criteria
Constraint
Org files remain authoritative for human-maintained research and design. Ledgers, graphs, indexes, embeddings, evaluations, and run state are derived or append-only reproducible state. ADARD may propose and evaluate changes to itself, but promotion into the canonical workflow remains gated, attributable, reversible, and measurable.