Research-grade DonkeyCar RL autoresearch and sweep system.

Go to file

Paul Huliganga fc01057c14 docs: ADR-017 — always save best model, never just latest Documents the root cause of losing the mountain_track model that was doing 20-second laps at step 30k but crashed at step 90k final eval. Phase 2 (13k steps, simple track): final = best. Assumption carried forward incorrectly into Wave 4 (90k steps, policy can drift). Mandatory rule: every training script uses train_multitrack() best_model tracking OR SB3 EvalCallback. No exceptions. Agent: pi Tests: 102 passed Tests-Added: 0 TypeScript: N/A		2026-04-17 16:03:59 -04:00
.harness	feat: Wave 1 complete — real PPO training, model save, GP+UCB autoresearch, 37 tests passing	2026-04-13 10:03:15 -04:00
agent	fix: always save and return the BEST model, not the last one	2026-04-17 14:45:37 -04:00
docs	docs: ARCHITECTURE.md — complete system architecture guide	2026-04-17 14:06:38 -04:00
tests	feat: v5 reward — speed × CTE-quality, drop efficiency term	2026-04-17 13:25:38 -04:00
.gitignore	feat: Wave 1 complete — real PPO training, model save, GP+UCB autoresearch, 37 tests passing	2026-04-13 10:03:15 -04:00
AGENT.md	feat: Wave 1 complete — real PPO training, model save, GP+UCB autoresearch, 37 tests passing	2026-04-13 10:03:15 -04:00
DECISIONS.md	docs: ADR-017 — always save best model, never just latest	2026-04-17 16:03:59 -04:00
IMPLEMENTATION_PLAN.md	feat: Phase 3 — behavioral control, enhanced evaluator, 53 tests	2026-04-14 09:28:43 -04:00
PROJECT-KICKOFF.md	feat: Wave 1 complete — real PPO training, model save, GP+UCB autoresearch, 37 tests passing	2026-04-13 10:03:15 -04:00
PROJECT-SPEC.md	feat: Wave 1 complete — real PPO training, model save, GP+UCB autoresearch, 37 tests passing	2026-04-13 10:03:15 -04:00
README.md	feat: Wave 1 complete — real PPO training, model save, GP+UCB autoresearch, 37 tests passing	2026-04-13 10:03:15 -04:00
create_gitea_repo.py	Initial commit	2026-04-12 23:44:36 -04:00
ralph-loop.sh	feat: Wave 1 complete — real PPO training, model save, GP+UCB autoresearch, 37 tests passing	2026-04-13 10:03:15 -04:00

README.md

donkeycar-rl-autoresearch

Purpose

Status

Scaffolded with the agent harness
Spec not filled yet

Runbook

Fill PROJECT-SPEC.md
Create IMPLEMENTATION_PLAN.md from the spec
Start the implementation loop