RUT-Bench Collection Benchmark data in "Beyond Ideal Instruction: A Comprehensive Framework for Evaluating LLMs in Realistic Interactions". • 2 items • Updated Jun 4 • 1
GigaChat Audio: Time-aware Large Audio Language Model Paper • 2607.10387 • Published 28 days ago • 37
👤 Implicit Personalization in Language Models Collection Works on detecting, attributing and controlling implicit personalization in language models • 29 items • Updated Mar 20 • 4
Finetuning LLMs for Human Behavior Prediction in Social Science Experiments Paper • 2509.05830 • Published Sep 6, 2025 • 1
Mirroring Users: Towards Building Preference-aligned User Simulator with User Feedback in Recommendation Paper • 2508.18142 • Published Aug 25, 2025 • 1
Learning from Language Feedback via Variational Policy Distillation Paper • 2605.15113 • Published May 18 • 13
Flipping the Dialogue: Training and Evaluating User Language Models Paper • 2510.06552 • Published Oct 8, 2025 • 2
GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Paper • 2604.26752 • Published Apr 29 • 114
Beyond Mode Collapse: Distribution Matching for Diverse Reasoning Paper • 2605.19461 • Published May 19 • 2
HDPO: Hybrid Distillation Policy Optimization via Privileged Self-Distillation Paper • 2603.23871 • Published Mar 25 • 1