Fix: Resolved name 'torch' is not defined and fixed Gradio image paths 0ed200e joshua400 commited on Apr 26
Fix: Enhanced model loading with trust_remote_code and added UI error reporting 9fceabb joshua400 commited on Apr 26
Fix: Added PeftModel loading logic for LoRA adapters and updated requirements 5cb03a6 joshua400 commited on Apr 26
Final submission: Dual-model support (Llama-1B & Qwen-7B), consolidated asset_final, and README updates 8084d88 joshua400 commited on Apr 26
π FINAL INTEGRATION: Connected trained Llama-1B-GRPO model and added Analysis UI 304fe37 joshua400 commited on Apr 26
ποΈ FINAL: Critical bug fixes, reward tuning, and judge-ready README db78ce2 joshua400 commited on Apr 26
π FEAT: Integrated Live HF LLM (Llama-3) & Training Data Logging b512de5 joshua400 commited on Apr 26
π§ FIX: Fixed simulation progression loop, added phase-aware policies, and restored action history tracking 39b98d6 joshua400 commited on Apr 26
π STRUCTURAL REBUILD: Fully aligned FairRecovery++ with OpenEnv reference patterns and stateful grading 9cfc074 joshua400 commited on Apr 26
Initial commit: FairRecovery++ complete multi-agent RL environment ce75dcf joshua400 commited on Apr 25