arxiv:2608.08975
Tianyi Zhou
zhoutianyi
AI & ML interests
ML, NLP, RL, Multi-modality
Recent Activity
authored a paper 6 days ago
How Can Rhetoric Reward-Hack AI Reviewers? Dissecting Rhetorical Sensitivity in AI-Based Peer Review authored a paper 6 days ago
Reinforcement Learning with Evolving Rubrics as Rewards for Audio Reasoning