πŸ”₯ News

Aug 28, 2026 Tencent Hy4-preview is released now.
Aug 20, 2026 Our paper β€œMedACE: A Time-Aware Asynchronous Clinical Environment for Long-Horizon Medical Treatment Decision-Making” is accepted at EMNLP Findings 2026!
Aug 20, 2026 Our paper β€œTRACE: A Unified Rollout Budget Allocation Framework for Efficient Agentic Reinforcement Learning” is accepted at EMNLP 2026!
Jun 30, 2026 Honored to receive the Outstanding Graduate award of Zhejiang University!
May 21, 2026 Tencent Hy3 is released now.
Apr 23, 2026 Tencent Hy3-preview is released now.
Apr 15, 2026 Our paper β€œDataXman: Selecting and Mixing Pretraining Data via Bilingual Mixture-of-Experts Data Manager” is accepted at SCIENCE CHINA Information Sciences 2026!
Apr 07, 2026 Our paper β€œHSS-Synth: Humanities and Social Sciences Data Synthesis for LLMs” is accepted at ACL Findings 2026!
Mar 02, 2026 Our paper β€œW2S: Weak-to-Strong Prompt Correction for Large Language Models” is accepted at Machine Learning 2026!
Jan 26, 2026 Our paper β€œOptimSyn: Influence-Guided Rubrics Optimization for Synthetic Data Generation” is accepted at ICLR 2026!
Aug 18, 2025 Ant RL technical report β€œReinforcement learning with rubric anchorsβ€œ(extending RLVR with 10k+ Rubric rewards) is now released.
Feb 20, 2025 Gave an invited talk on β€œDataMan: Data Manager for Pre-training Large Language Models” at JIQIZHIXIN (ζœΊε™¨δΉ‹εΏƒ)!
Feb 11, 2025 Our paper β€œLLM-Enhanced Query Generation and Retrieval Preservation for Task-Oriented Dialogue” is accepted at Findings of ACL 2025!
Feb 11, 2025 Our paper β€œDataMan: Data Manager for Pre-training Large Language Models” is accepted at ICLR 2025!
Dec 19, 2024 Qwen2.5 technical report are released now.
Sep 20, 2024 One paper β€œInference-Time Decontamination: Reusing Leaked Benchmarks for Large Language Model Evaluation” is accepted at Findings of EMNLP 2024 and two paper β€œPredicting Rewards Alongside Tokens: Non-disruptive Parameter Insertion for Efficient Inference Intervention in Large Language Model”, β€œEmbedding and Gradient Say Wrong: A White-Box Method for Hallucination Detection” are accepted at EMNLP 2024!
Sep 19, 2024 Qwen2.5 series foundation models are released now.
Jul 15, 2024 Qwen2 technical report are released now.
Jul 04, 2024 Release the paper of β€œDotamath” for mathematical reasoning.
Jun 17, 2024 Qwen2 series foundation models are released now.
May 16, 2024 Our paper β€œDORY: Deliberative Prompt Recovery for LLM” is accepted at Findings of ACL 2024!
Feb 04, 2024 Qwen1.5 series foundation models are released now.
Jan 16, 2024 Our paper β€œEnergy-based Automated Model Evaluation” is accepted at ICLR 2024!
Oct 23, 2023 I started my internship at Alibaba Qwen Team! Ping me if you want to meet up in HangZhou :)
Jul 15, 2023 Our paper β€œCAME: Contrastive Automated Model Evaluation” is accepted at ICCV 2023!
Oct 06, 2022 Our paper β€œDistill The Image to Nowhere: Inversion Knowledge Distillation for Multimodal Machine Translation” is accepted at EMNLP 2022 (Oral)!
Sep 10, 2022 Started my PhD’s degree at College of Computer Science and Technology of Zhejiang University!
Apr 06, 2022 Our paper β€œHybridVocab: Towards Multi-Modal Machine Translation via Multi-Aspect Alignment” is accpeted at ICMR 2022 (Oral)!

Β