Publications 论文 論文

Complete list of Yanjun Chen's papers: peer-reviewed publications, preprints, and works in submission on training environments, reward models, credit assignment, and reinforcement learning.

Peer-reviewed

2025

  1. EMNLP
    PricingLogic: Evaluating LLMs Reasoning on Complex Tourism Pricing Tasks
    Yunuo Liu, Dawei Zhu, Zena Al-Khalili, and 5 more authors
    In Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing (EMNLP), 2025
  2. ACL Findings
    Unveiling the Key Factors for Distilling Chain-of-Thought Reasoning
    Xinghao Chen, Zhixin Sun, Wenjin Guo, and 6 more authors
    In Findings of the Association for Computational Linguistics (ACL Findings), 2025
  3. NAACL
    Fine-Grained and Multi-Dimensional Metrics for Document-Level Machine Translation
    Yirong Sun, Dawei Zhu, Yanjun Chen, and 3 more authors
    In Proceedings of the 2025 Conference of the North American Chapter of the Association for Computational Linguistics (NAACL), 2025

2024

  1. EMNLP
    The Accuracy Paradox in RLHF: When Better Reward Models Don’t Yield Better Language Models
    Yanjun Chen, Dawei Zhu, Yirong Sun, and 3 more authors
    In Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing (EMNLP), 2024

Preprints & under review

2026

  1. arXiv
    Exact Is Easier: Credit Assignment for Cooperative LLM Agents
    Yanjun Chen, Yirong Sun, Hanlin Wang, and 5 more authors
    arXiv preprint arXiv:2603.06859, 2026
    In submission.
  2. arXiv
    FastDSAC: Unlocking the Potential of Maximum Entropy RL in High-Dimensional Humanoid Control
    Jun Xue, Junze Wang, Shanze Wang, and 3 more authors
    arXiv preprint arXiv:2603.12612, 2026
  3. arXiv
    SonicBench: Dissecting the Physical Perception Bottleneck in Large Audio Language Models
    Yirong Sun, Yanjun Chen, Xin Qiu, and 8 more authors
    arXiv preprint arXiv:2601.11039, 2026

2025

  1. arXiv
    Integrating Chain-of-Thought for Multimodal Alignment: A Study on 3D Vision-Language Learning
    Yanjun Chen, Yirong Sun, Xinghao Chen, and 4 more authors
    arXiv preprint arXiv:2503.06232, 2025
  2. arXiv
    LLaSO: A Foundational Framework for Reproducible Research in Large Language and Speech Model
    Yirong Sun, Yizhong Geng, Peidong Wei, and 5 more authors
    arXiv preprint arXiv:2508.15418, 2025
  3. arXiv
    MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos
    Xianghui Wang, Xinming Zhang, Yanjun Chen, and 2 more authors
    arXiv preprint arXiv:2505.08367, 2025
  4. arXiv
    Breaking the Pre-Planning Barrier: Adaptive Real-Time Coordination of Heterogeneous UAVs
    Yuhan Hu, Yirong Sun, Yanjun Chen, and 3 more authors
    arXiv preprint arXiv:2501.14488, 2025
  5. arXiv
    Reasoning Beyond Language: A Comprehensive Survey on Latent Chain-of-Thought Reasoning
    Xinghao Chen, Anhao Zhao, Heming Xia, and 7 more authors
    arXiv preprint arXiv:2505.16782, 2025

2024

  1. arXiv
    Rethinking Soft Actor-Critic in High-Dimensional Action Spaces: The Cost of Ignoring Distribution Shift
    Yanjun Chen, Xinming Zhang, Xianghui Wang, and 3 more authors
    arXiv preprint arXiv:2410.16739, 2024