π FINAL INTEGRATION: Connected trained Llama-1B-GRPO model and added Analysis UI 304fe37 joshua400 commited on Apr 26
ποΈ FINAL: Critical bug fixes, reward tuning, and judge-ready README db78ce2 joshua400 commited on Apr 26
π FEAT: Integrated Live HF LLM (Llama-3) & Training Data Logging b512de5 joshua400 commited on Apr 26
π§ FIX: Fixed simulation progression loop, added phase-aware policies, and restored action history tracking 39b98d6 joshua400 commited on Apr 26
π STRUCTURAL REBUILD: Fully aligned FairRecovery++ with OpenEnv reference patterns and stateful grading 9cfc074 joshua400 commited on Apr 26
Initial commit: FairRecovery++ complete multi-agent RL environment ce75dcf joshua400 commited on Apr 25