arxiv:2608.08975
Tianyi Zhou
zhoutianyi
AI & ML interests
ML, NLP, RL, Multi-modality
Recent Activity
authored a paper about 5 hours ago
How Can Rhetoric Reward-Hack AI Reviewers? Dissecting Rhetorical Sensitivity in AI-Based Peer Review authored a paper about 5 hours ago
Reinforcement Learning with Evolving Rubrics as Rewards for Audio Reasoning upvoted a paper about 6 hours ago
How Can Rhetoric Reward-Hack AI Reviewers? Dissecting Rhetorical Sensitivity in AI-Based Peer Review