Commit History

Fix: Resolved NameError for reward_img in app.py causing startup crash
c9f3323

joshua400 commited on

Fix: Reverted to absolute HF resolve URLs and removed path replacement in app for better image reliability
5bd6696

joshua400 commited on

Docs & UI Sync: Synchronized all photos and links across App UI, README, and Blog
a2a9a1b

joshua400 commited on

Final Winner Polish: Added signature line, fairness realism disclaimer, and optimized model strategy
d094bdb

joshua400 commited on

Docs: Updated blog.md with the latest calibrated Key Numbers
e5da0a3

joshua400 commited on

Evidence: Added advanced stability analysis plot (Figure 6) and re-indexed figures
6da87f7

joshua400 commited on

Docs: Integrated final results ASCII table into Key Results section
6959dc8

joshua400 commited on

Docs: Added explicit Qwen-1.5B training details to satisfy non-negotiable evidence requirement
ccb5303

joshua400 commited on

Docs: Added Qwen-1.5B Lite Agent to top badges and Trained Models section
29a6ebf

joshua400 commited on

Fix: Explicitly tracked training_metrics_extra.png and updated .gitignore
868eefc

joshua400 commited on

Evidence: Added extended training metrics plot (Figure 5) to the dashboard
39625d6

joshua400 commited on

Docs: Added Kaggle 1.5B Quick Testing notebook link to README
ceaa9f2

joshua400 commited on

Docs: Updated README to link to root-level blog.md
29d760e

joshua400 commited on

Docs: Added blog.md to root for easier judge discovery
eb259c0

joshua400 commited on

Fix: Synced Llama model ID with README for consistent judge discovery
623522c

joshua400 commited on

Final Winner Polish: Applied 9 surgical enhancements to opening, results, novelty, and storytelling sections
d9780ed

joshua400 commited on

Fix: Update .gitignore to allow evidence/plots while ignoring root plots
916953f

joshua400 commited on

Docs: Added official Colab training link to the Evidence section to satisfy the reproducible script requirement
1acc92f

joshua400 commited on

Docs: Removed internal meta-scoring commentary to ensure a professional, judge-facing presentation
f047faf

joshua400 commited on

Final Polish: Realistic UI equity caps, storytelling UI insights, clean model naming, and Top/Bottom README narrative impacts
e6ec862

joshua400 commited on

Docs: Added explicit sections for Innovation, Fairness Metric, Before/After storytelling, and mandatory Training Evidence to optimize for hackathon judging criteria
8303ab6

joshua400 commited on

UI: Refactored run_simulation to use yield for real-time UI updates during model load and simulation
24b4080

joshua400 commited on

UI: Added missing reward_vs_episode and step-wise plots to the Gradio dashboard
98421fe

joshua400 commited on

Fix: Removed .gitignore block and reverted app.py to local paths to fix 404 crash
9fb3d07

joshua400 commited on

Fix: Replaced all relative image paths with absolute Hugging Face URLs to guarantee visibility in Markdown and Gradio
f040529

joshua400 commited on

Fix: Resolved name 'torch' is not defined and fixed Gradio image paths
0ed200e

joshua400 commited on

Final: Replaced README intro with refined HF Blog Post narrative and restored badges
8548d74

joshua400 commited on

Fix: Final attempt at image visibility - renamed directory to evidence and updated all links
6f853ab

joshua400 commited on

Fix: Renamed asset_final to evidence to resolve path issues on HF
d48ec1f

joshua400 commited on

Final: Fixed broken images and added high-visibility Visual Evidence Dashboard
0b95fec

joshua400 commited on

Fix: Enhanced model loading with trust_remote_code and added UI error reporting
9fceabb

joshua400 commited on

Docs: Added explicit Models section and step-level analysis plots to README
92b8902

joshua400 commited on

Final: Regenerated plots with correct Qwen-7B-GRPO labels and synced asset_final
624e662

joshua400 commited on

Docs: Unified project structure, reordered training plots, and updated image paths to asset_final
2c6602e

joshua400 commited on

Fix: Added PeftModel loading logic for LoRA adapters and updated requirements
5cb03a6

joshua400 commited on

Final submission: Dual-model support (Llama-1B & Qwen-7B), consolidated asset_final, and README updates
8084d88

joshua400 commited on

πŸš€ FINAL SYNC: Integrated final plot assets from asset_final into main UI and README
dbe1840

joshua400 commited on

πŸš€ FINAL POLISH: Fixed image visibility in UI and added captioned plots to README
b8afba2

joshua400 commited on

πŸš€ OPTIMIZE: Model loading hardware-aware logic and eval mode
2521148

joshua400 commited on

πŸš€ DOCS: Added Model badge to README for easy access to trained agent
74e0a5e

joshua400 commited on

πŸš€ FINAL INTEGRATION: Connected trained Llama-1B-GRPO model and added Analysis UI
304fe37

joshua400 commited on

πŸ”§ FIX: Standardized project links and renamed docs folder for professional submission
bee85a1

joshua400 commited on

πŸ”§ FIX: Assets visibility in Gradio UI and YAML cleanup in README tab
15e53d2

joshua400 commited on

πŸ”§ FINAL FIX: Eliminated all remaining 'service_level' references in environment and reward logic
8248fc6

joshua400 commited on

πŸ”§ FIX: Resolved AttributeError by completing service_level to service rename in all reward helpers
3e32a79

joshua400 commited on

πŸ”§ FIX: Restored missing TaskGrader class to rewards.py to resolve ImportError
aaf71de

joshua400 commited on

πŸ”§ FIX: Corrected Hugging Face YAML metadata colors
5768533

joshua400 commited on

πŸ”§ FIX: Added missing Hugging Face Space metadata YAML header
5002afe

joshua400 commited on

πŸ™οΈ FINAL: Critical bug fixes, reward tuning, and judge-ready README
db78ce2

joshua400 commited on

πŸ”§ FIX: Resolved Gradio 'missing input' error and improved UI layout
0f9d249

joshua400 commited on