Davi Ferreira
davifs
AI & ML interests
None yet
Recent Activity
upvoted a paper about 9 hours ago
AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning upvoted a paper 4 days ago
DAPD: Dual-Anchored Policy Distillation