Pengyu Cheng
Linear95
AI & ML interests
None yet
Recent Activity
upvoted a paper 2 days ago
MARCH: Multi-Agent Reinforced Self-Check for LLM Hallucination upvoted a paper 2 days ago
CLIPO: Contrastive Learning in Policy Optimization Generalizes RLVR authored a paper 2 days ago
On Diversified Preferences of Large Language Model Alignment