Pengyu Cheng
Linear95
AI & ML interests
None yet
Recent Activity
upvoted a paper about 12 hours ago
Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills upvoted a paper about 1 month ago
GD^2PO: Mitigating Multi-Reward Conflicts via Group-Dynamic reward-Decoupled Policy Optimization upvoted a paper about 2 months ago
MARCH: Multi-Agent Reinforced Self-Check for LLM Hallucination