π₯ News
| Aug 28, 2026 | Tencent Hy4-preview is released now. |
|---|---|
| Aug 20, 2026 | Our paper βMedACE: A Time-Aware Asynchronous Clinical Environment for Long-Horizon Medical Treatment Decision-Makingβ is accepted at EMNLP Findings 2026! |
| Aug 20, 2026 | Our paper βTRACE: A Unified Rollout Budget Allocation Framework for Efficient Agentic Reinforcement Learningβ is accepted at EMNLP 2026! |
| Jun 30, 2026 | Honored to receive the Outstanding Graduate award of Zhejiang University! |
| May 21, 2026 | Tencent Hy3 is released now. |
| Apr 23, 2026 | Tencent Hy3-preview is released now. |
| Apr 15, 2026 | Our paper βDataXman: Selecting and Mixing Pretraining Data via Bilingual Mixture-of-Experts Data Managerβ is accepted at SCIENCE CHINA Information Sciences 2026! |
| Apr 07, 2026 | Our paper βHSS-Synth: Humanities and Social Sciences Data Synthesis for LLMsβ is accepted at ACL Findings 2026! |
| Mar 02, 2026 | Our paper βW2S: Weak-to-Strong Prompt Correction for Large Language Modelsβ is accepted at Machine Learning 2026! |
| Jan 26, 2026 | Our paper βOptimSyn: Influence-Guided Rubrics Optimization for Synthetic Data Generationβ is accepted at ICLR 2026! |
| Aug 18, 2025 | Ant RL technical report βReinforcement learning with rubric anchorsβ(extending RLVR with 10k+ Rubric rewards) is now released. |
| Feb 20, 2025 | Gave an invited talk on βDataMan: Data Manager for Pre-training Large Language Modelsβ at JIQIZHIXIN (ζΊε¨δΉεΏ)! |
| Feb 11, 2025 | Our paper βLLM-Enhanced Query Generation and Retrieval Preservation for Task-Oriented Dialogueβ is accepted at Findings of ACL 2025! |
| Feb 11, 2025 | Our paper βDataMan: Data Manager for Pre-training Large Language Modelsβ is accepted at ICLR 2025! |
| Dec 19, 2024 | Qwen2.5 technical report are released now. |
| Sep 20, 2024 | One paper βInference-Time Decontamination: Reusing Leaked Benchmarks for Large Language Model Evaluationβ is accepted at Findings of EMNLP 2024 and two paper βPredicting Rewards Alongside Tokens: Non-disruptive Parameter Insertion for Efficient Inference Intervention in Large Language Modelβ, βEmbedding and Gradient Say Wrong: A White-Box Method for Hallucination Detectionβ are accepted at EMNLP 2024! |
| Sep 19, 2024 | Qwen2.5 series foundation models are released now. |
| Jul 15, 2024 | Qwen2 technical report are released now. |
| Jul 04, 2024 | Release the paper of βDotamathβ for mathematical reasoning. |
| Jun 17, 2024 | Qwen2 series foundation models are released now. |
| May 16, 2024 | Our paper βDORY: Deliberative Prompt Recovery for LLMβ is accepted at Findings of ACL 2024! |
| Feb 04, 2024 | Qwen1.5 series foundation models are released now. |
| Jan 16, 2024 | Our paper βEnergy-based Automated Model Evaluationβ is accepted at ICLR 2024! |
| Oct 23, 2023 | I started my internship at Alibaba Qwen Team! Ping me if you want to meet up in HangZhou :) |
| Jul 15, 2023 | Our paper βCAME: Contrastive Automated Model Evaluationβ is accepted at ICCV 2023! |
| Oct 06, 2022 | Our paper βDistill The Image to Nowhere: Inversion Knowledge Distillation for Multimodal Machine Translationβ is accepted at EMNLP 2022 (Oral)! |
| Sep 10, 2022 | Started my PhDβs degree at College of Computer Science and Technology of Zhejiang University! |
| Apr 06, 2022 | Our paper βHybridVocab: Towards Multi-Modal Machine Translation via Multi-Aspect Alignmentβ is accpeted at ICMR 2022 (Oral)! |
Β