Docs: Integrated final results ASCII table into Key Results section
6959dc8
joshua400commited on
Docs: Added explicit Qwen-1.5B training details to satisfy non-negotiable evidence requirement
ccb5303
joshua400commited on
Docs: Added Qwen-1.5B Lite Agent to top badges and Trained Models section
29a6ebf
joshua400commited on
Fix: Explicitly tracked training_metrics_extra.png and updated .gitignore
868eefc
joshua400commited on
Evidence: Added extended training metrics plot (Figure 5) to the dashboard
39625d6
joshua400commited on
Docs: Added Kaggle 1.5B Quick Testing notebook link to README
ceaa9f2
joshua400commited on
Docs: Updated README to link to root-level blog.md
29d760e
joshua400commited on
Docs: Added blog.md to root for easier judge discovery
eb259c0
joshua400commited on
Fix: Synced Llama model ID with README for consistent judge discovery
623522c
joshua400commited on
Final Winner Polish: Applied 9 surgical enhancements to opening, results, novelty, and storytelling sections
d9780ed
joshua400commited on
Fix: Update .gitignore to allow evidence/plots while ignoring root plots
916953f
joshua400commited on
Docs: Added official Colab training link to the Evidence section to satisfy the reproducible script requirement
1acc92f
joshua400commited on
Docs: Removed internal meta-scoring commentary to ensure a professional, judge-facing presentation
f047faf
joshua400commited on
Final Polish: Realistic UI equity caps, storytelling UI insights, clean model naming, and Top/Bottom README narrative impacts
e6ec862
joshua400commited on
Docs: Added explicit sections for Innovation, Fairness Metric, Before/After storytelling, and mandatory Training Evidence to optimize for hackathon judging criteria
8303ab6
joshua400commited on
UI: Refactored run_simulation to use yield for real-time UI updates during model load and simulation
24b4080
joshua400commited on
UI: Added missing reward_vs_episode and step-wise plots to the Gradio dashboard
98421fe
joshua400commited on
Fix: Removed .gitignore block and reverted app.py to local paths to fix 404 crash
9fb3d07
joshua400commited on
Fix: Replaced all relative image paths with absolute Hugging Face URLs to guarantee visibility in Markdown and Gradio
f040529
joshua400commited on
Fix: Resolved name 'torch' is not defined and fixed Gradio image paths
0ed200e
joshua400commited on
Final: Replaced README intro with refined HF Blog Post narrative and restored badges
8548d74
joshua400commited on
Fix: Final attempt at image visibility - renamed directory to evidence and updated all links
6f853ab
joshua400commited on
Fix: Renamed asset_final to evidence to resolve path issues on HF
d48ec1f
joshua400commited on
Final: Fixed broken images and added high-visibility Visual Evidence Dashboard
0b95fec
joshua400commited on
Fix: Enhanced model loading with trust_remote_code and added UI error reporting
9fceabb
joshua400commited on
Docs: Added explicit Models section and step-level analysis plots to README
92b8902
joshua400commited on
Final: Regenerated plots with correct Qwen-7B-GRPO labels and synced asset_final
624e662
joshua400commited on
Docs: Unified project structure, reordered training plots, and updated image paths to asset_final
2c6602e
joshua400commited on
Fix: Added PeftModel loading logic for LoRA adapters and updated requirements
5cb03a6
joshua400commited on
Final submission: Dual-model support (Llama-1B & Qwen-7B), consolidated asset_final, and README updates
8084d88
joshua400commited on
π FINAL SYNC: Integrated final plot assets from asset_final into main UI and README
dbe1840
joshua400commited on
π FINAL POLISH: Fixed image visibility in UI and added captioned plots to README
b8afba2
joshua400commited on
π OPTIMIZE: Model loading hardware-aware logic and eval mode
2521148
joshua400commited on
π DOCS: Added Model badge to README for easy access to trained agent
74e0a5e
joshua400commited on
π FINAL INTEGRATION: Connected trained Llama-1B-GRPO model and added Analysis UI
304fe37
joshua400commited on
π§ FIX: Standardized project links and renamed docs folder for professional submission
bee85a1
joshua400commited on
π§ FIX: Assets visibility in Gradio UI and YAML cleanup in README tab
15e53d2
joshua400commited on
π§ FINAL FIX: Eliminated all remaining 'service_level' references in environment and reward logic
8248fc6
joshua400commited on
π§ FIX: Resolved AttributeError by completing service_level to service rename in all reward helpers
3e32a79
joshua400commited on
π§ FIX: Restored missing TaskGrader class to rewards.py to resolve ImportError
aaf71de
joshua400commited on
π§ FIX: Corrected Hugging Face YAML metadata colors
5768533
joshua400commited on
π§ FIX: Added missing Hugging Face Space metadata YAML header
5002afe
joshua400commited on
ποΈ FINAL: Critical bug fixes, reward tuning, and judge-ready README