publications

Selected and recent work. The complete and always-current list lives on Google Scholar.

Full list and citation counts: Google Scholar.

2026

  1. ModularRSI: Modular and Generalizable Recursive Harness Self-Improvement
    Siwei Wu, Jincheng Ren, Yizhi Li, and 11 more authors
    arXiv preprint arXiv:2609.14857, 2026
  2. Harbor Adapters and Harbor-Index: Infrastructure and a Curated Meta-Dataset for Large-Scale Agentic Evaluation
    Lei Shi, Hanfeng Lin, Ziheng Zhu, and 9 more authors
    arXiv preprint arXiv:2609.04298, 2026
  3. Large-Scale Terminal Agentic Trajectory Generation from Dockerized Environments
    Siwei Wu, Yizhi Li, Yuyang Song, and 8 more authors
    In International Conference on Machine Learning (ICML), 2026
  4. IQuest-Coder-V1 Technical Report
    Jian Yang, Wei Zhang, Shawn Guo, and 9 more authors
    arXiv preprint arXiv:2603.16733, 2026
  5. InCoder-32B: Code Foundation Model for Industrial Scenarios
    Jian Yang, Wei Zhang, Jiajun Wu, and 9 more authors
    arXiv preprint arXiv:2603.16790, 2026
  6. A Self-Evolving Framework for Efficient Terminal Agents via Observational Context Compression
    Jincheng Ren, Siwei Wu, Yizhi Li, and 8 more authors
    arXiv preprint arXiv:2604.19572, 2026
  7. ACL
    HiRAS: A Hierarchical Multi-Agent Framework for Paper-to-Code Generation and Execution
    Hanhua Hong, Yizhi Li, Jiaxin Chen, and 4 more authors
    In Findings of the Association for Computational Linguistics: ACL 2026, 2026
  8. ACL
    Context as a Tool: Context Management for Long-Horizon SWE-Agents
    Shanchao Liu, Bingchang Jiang, Jian Yang, and 4 more authors
    In Findings of the Association for Computational Linguistics: ACL 2026, 2026
  9. AcadReason: Exploring the Limits of Reasoning Models with Academic Research Problems
    Xin Gui, Jincheng Ren, Qi Chen, and 7 more authors
    In International Conference on Learning Representations (ICLR), 2026
  10. ScaleLong: A Multi-Timescale Benchmark for Long Video Understanding
    David Ma, Huaqing Yuan, Xingjian Wang, and 8 more authors
    In International Conference on Learning Representations (ICLR), 2026
  11. OmniVideoBench: Towards Audio-Visual Understanding Evaluation for Omni MLLMs
    Caorui Li, Yu Chen, Yiyan Ji, and 8 more authors
    In International Conference on Learning Representations (ICLR), 2026
  12. YuE: Scaling Open Foundation Models for Long-Form Music Generation
    Ruibin Yuan, Hanfeng Lin, Shuyue Guo, and 18 more authors
    In International Conference on Learning Representations (ICLR), 2026
  13. MMRA: A Benchmark for Evaluating Multi-Granularity and Multi-Image Relational Association Capabilities in Large Visual Language Models
    Siwei Wu, Kang Zhu, Yuelin Bai, and 8 more authors
    In Findings of the Association for Computational Linguistics: EACL 2026, 2026
  14. Dynamics within Latent Chain-of-Thought: An Empirical Study of Causal Structure
    Zhaoqun Li, Xuanle Bai, Kai Chen, and 4 more authors
    arXiv preprint arXiv:2602.08783, 2026

2025

  1. Encyclo-K: Evaluating LLMs with Dynamically Composed Knowledge Statements
    Yiming Liang, Yizhi Li, Yantao Du, and 14 more authors
    arXiv preprint arXiv:2512.24867, 2025
  2. Dynamic Large Concept Models: Latent Reasoning in an Adaptive Semantic Space
    Xingwei Qu, Shaowen Wang, Zihao Huang, and 8 more authors
    arXiv preprint arXiv:2512.24617, 2025
  3. CodeSimpleQA: Scaling Factuality in Code Large Language Models
    Jian Yang, Wei Zhang, Yizhi Li, and 8 more authors
    arXiv preprint arXiv:2512.19424, 2025
  4. From Code Foundation Models to Agents and Applications: A Comprehensive Survey and Practical Guide to Code Intelligence
    Jian Yang, Xianglong Liu, Weifeng Lv, and 5 more authors
    arXiv preprint arXiv:2511.18538, 2025
  5. OmniBench: Towards the Future of Universal Omni-Language Models
    Yizhi Li, Ge Zhang, Yinghao Ma, and 8 more authors
    In Advances in Neural Information Processing Systems (NeurIPS), 2025
  6. SuperGPQA: Scaling LLM Evaluation across 285 Graduate Disciplines
    M-A-P Team, Xinrun Du, Yifan Yao, and 19 more authors
    In Advances in Neural Information Processing Systems (NeurIPS), 2025
  7. TreePO: Bridging the Gap of Policy Optimization and Efficacy and Inference Efficiency with Heuristic Tree-based Modeling
    Yizhi Li, Qingshui Gu, Zhoufutu Wen, and 14 more authors
    arXiv preprint arXiv:2508.17445, 2025
  8. First Return, Entropy-Eliciting Explore
    Tianyu Zheng, Tianshun Xing, Qingshui Gu, and 10 more authors
    arXiv preprint arXiv:2507.07017, 2025
  9. Overview of the NLPCC 2025 Shared Task: Gender Bias Mitigation Challenge
    Yizhi Li, Ge Zhang, Hanhua Hong, and 2 more authors
    In CCF International Conference on Natural Language Processing and Chinese Computing (NLPCC), 2025
  10. ACL
    MAmmoTH-VL: Eliciting Multimodal Reasoning with Instruction Tuning at Scale
    Jiawei Guo, Tianyu Zheng, Yizhi Li, and 7 more authors
    In Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (ACL), 2025
  11. ACL
    LIME: Less Is More for MLLM Evaluation
    Kang Zhu, Qianbo Zang, Shian Jia, and 8 more authors
    In Findings of the Association for Computational Linguistics: ACL 2025, 2025
  12. MIO: A Foundation Model on Multimodal Tokens
    Zekun Moore Wang, Kang Zhu, Chunpu Xu, and 8 more authors
    In Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing (EMNLP), 2025
  13. DocMMIR: A Framework for Document Multi-modal Information Retrieval
    Zirui Li, Siwei Wu, Xingyu Wang, and 3 more authors
    In Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing (EMNLP), 2025
  14. Re:Form – Reducing Human Priors in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny
    Chuanhao Yan, Fengdi Che, Xuhan Huang, and 8 more authors
    Transactions on Machine Learning Research (TMLR), 2025
  15. LongEval: A Comprehensive Analysis of Long-Text Generation Through a Plan-based Paradigm
    Siwei Wu, Yizhi Li, Xingwei Qu, and 7 more authors
    arXiv preprint arXiv:2502.19103, 2025
  16. MuPT: A Generative Symbolic Music Pretrained Transformer
    Xingwei Qu, Yinghao Ma, Chaoyang Zhou, and 8 more authors
    In International Conference on Learning Representations (ICLR), 2025
  17. Bridging the Data Provenance Gap across Text, Speech, and Video
    Shayne Longpre, Nikhil Singh, Manuel Cherep, and 4 more authors
    In International Conference on Learning Representations (ICLR), 2025

2024

  1. Consent in Crisis: The Rapid Decline of the AI Data Commons
    Shayne Longpre, Robert Mahari, Ariel Lee, and 4 more authors
    In Advances in Neural Information Processing Systems (NeurIPS), 2024
  2. Foundation Models for Music: A Survey
    Yinghao Ma, Anders Øland, Anton Ragni, and 8 more authors
    arXiv preprint arXiv:2408.14340, 2024
  3. ACL
    SciMMIR: Benchmarking Scientific Multi-modal Information Retrieval
    Siwei Wu, Yizhi Li, Kang Zhu, and 7 more authors
    In Findings of the Association for Computational Linguistics: ACL 2024, 2024
  4. ACL
    CIF-Bench: A Chinese Instruction-Following Benchmark for Evaluating the Generalizability of Large Language Models
    Yizhi Li, Ge Zhang, Xingwei Qu, and 8 more authors
    In Findings of the Association for Computational Linguistics: ACL 2024, 2024
  5. MAP-Neo: Highly Capable and Transparent Bilingual Large Language Model Series
    Ge Zhang, Scott Qu, Jiaheng Liu, and 9 more authors
    arXiv preprint arXiv:2405.19327, 2024
  6. ACL
    ChatMusician: Understanding and Generating Music Intrinsically with LLM
    Ruibin Yuan, Hanfeng Lin, Yi Wang, and 7 more authors
    In Findings of the Association for Computational Linguistics: ACL 2024, 2024
  7. MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
    Yizhi Li, Ruibin Yuan, Ge Zhang, and 15 more authors
    In International Conference on Learning Representations (ICLR), 2024

2023

  1. MARBLE: Music Audio Representation Benchmark for Universal Evaluation
    Ruibin Yuan, Yinghao Ma, Yizhi Li, and 8 more authors
    In Advances in Neural Information Processing Systems (NeurIPS), 2023
  2. On the Effectiveness of Speech Self-supervised Learning for Music
    Yinghao Ma, Ruibin Yuan, Yizhi Li, and 7 more authors
    In International Society for Music Information Retrieval Conference (ISMIR), 2023
  3. LyricWhiz: Robust Multilingual Zero-shot Lyrics Transcription by Whispering to ChatGPT
    Le Zhuo, Ruibin Yuan, Jiahao Pan, and 7 more authors
    In International Society for Music Information Retrieval Conference (ISMIR), 2023
  4. Length is a Curse and a Blessing for Document-level Semantics
    Chenghao Xiao, Yizhi Li, G Thomas Hudson, and 2 more authors
    In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing (EMNLP), 2023
  5. Interactive Natural Language Processing
    Zekun Wang, Ge Zhang, Kexin Yang, and 8 more authors
    arXiv preprint arXiv:2305.13246, 2023
  6. Chinese Open Instruction Generalist: A Preliminary Release
    Ge Zhang, Yemin Shi, Ruibo Liu, and 8 more authors
    arXiv preprint arXiv:2304.07987, 2023
  7. CORGI-PM: A Chinese Corpus for Gender Bias Probing and Mitigation
    Ge Zhang, Yizhi Li, Yaoyao Wu, and 5 more authors
    arXiv preprint arXiv:2301.00395, 2023

2022

  1. MAP-Music2Vec: A Simple and Effective Baseline for Self-Supervised Music Audio Representation Learning
    Yizhi Li, Ruibin Yuan, Ge Zhang, and 8 more authors
    arXiv preprint arXiv:2212.02508, 2022
  2. HERB: Measuring Hierarchical Regional Bias in Pre-trained Language Models
    Yizhi Li, Ge Zhang, Bohao Yang, and 4 more authors
    In Findings of AACL-IJCNLP 2022, 2022
  3. TranSHER: Translating Knowledge Graph Embedding with Hyper-Ellipsoidal Restriction
    Yizhi Li, Wei Fan, Chao Liu, and 2 more authors
    In Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing (EMNLP), 2022

2021

  1. More Robust Dense Retrieval with Contrastive Dual Learning
    Yizhi Li, Zhenghao Liu, Chenyan Xiong, and 1 more author
    In Proceedings of the 2021 ACM SIGIR International Conference on Theory of Information Retrieval, 2021
  2. From Rumor to Genetic Mutation Detection with Explanations: A GAN Approach
    Mingxi Cheng, Yizhi Li, Shahin Nazarian, and 1 more author
    Scientific Reports, 2021