publications

In reverse chronological order. * denotes equal contribution.

2026

  1. Randomized YaRN Improves Length Generalization for Long-Context Reasoning
    Manas Mehta, Fangcong Yin, and Greg Durrett
    In Findings of the Association for Computational Linguistics: EMNLP 2026, 2026
  2. Detecting and Suppressing Reward Hacking with Gradient Fingerprints
    Songtao Wang, Quang Hieu Pham, Fangcong Yin, Xinpeng Wang, Jocelyn Qiaochu Chen, Greg Durrett, and Xi Ye
    In Proceedings of the Conference on Language Modeling (COLM), 2026
  3. Visually Grounded Self-Reflection for Vision-Language Models via Reinforcement Learning
    Liyan Tang*, Fangcong Yin*, and Greg Durrett
    Preprint, 2026
  4. DySCO: Dynamic Attention-Scaling Decoding for Long-Context LMs
    Xi Ye*, Wuwei Zhang*, Fangcong Yin, Howard Yen, and Danqi Chen
    arXiv preprint arXiv:2602.22175, 2026
  5. When Many Answers Are Valid, Voting Fails: Symbolic Verification for Best-of-K Causal Reasoning in LLMs
    Omatharv Bharat Vaidya, Connor Thomas Jerzak, Zayne Sprague, Fangcong Yin, and Nhat Ho
    arXiv preprint arXiv:2608.03506, 2026

2025

  1. Learning Composable Chains-of-Thought
    Fangcong Yin, Zeyu Leo Liu, Liu Leqi, Xi Ye, and Greg Durrett
    In Findings of the Association for Computational Linguistics: EMNLP 2026. Also presented at the Workshop on Foundations of Reasoning in Language Models (Oral), NeurIPS 2025 , 2025
  2. Query-Focused Retrieval Heads Improve Long-Context Reasoning and Re-ranking
    Wuwei Zhang, Fangcong Yin, Howard Yen, Danqi Chen, and Xi Ye
    In Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing (EMNLP), 2025
  3. ChartMuseum: Testing Visual Reasoning Capabilities of Large Vision-Language Models
    Liyan Tang, Grace Kim, Xinyu Zhao, Thom Lake, Wenxuan Ding, Fangcong Yin, Prasann Singhal, Manya Wadhwa, Zeyu Leo Liu, Zayne Sprague, Ramya Namuduri, Bodun Hu, Juan Diego Rodriguez, Puyuan Peng, and Greg Durrett
    In Proceedings of the 39th Conference on Neural Information Processing Systems (NeurIPS), Datasets and Benchmarks Track, 2025
  4. Linear Representation Transferability Hypothesis: Leveraging Small Models to Steer Large Models
    Femi Bello, Anubrata Das, Fanzhi Zeng, Fangcong Yin, and Liu Leqi
    arXiv preprint arXiv:2506.00653, 2025
  5. LongProc: Benchmarking Long-Context Language Models on Long Procedural Generation
    Xi Ye, Fangcong Yin*, Yinghui He*, Joie Zhang*, Howard Yen*, Tianyu Gao, Greg Durrett, and Danqi Chen
    In Proceedings of the Conference on Language Modeling (COLM), 2025
  6. Understanding Synthetic Context Extension via Retrieval Heads
    Xinyu Zhao, Fangcong Yin, and Greg Durrett
    In Proceedings of the 42nd International Conference on Machine Learning (ICML), 2025
  7. To CoT or not to CoT? Chain-of-thought Helps Mainly on Math and Symbolic Reasoning
    Zayne Sprague, Fangcong Yin, Juan Diego Rodriguez, Dongwei Jiang, Manya Wadhwa, Prasann Singhal, Xinyu Zhao, Xi Ye, Kyle Mahowald, and Greg Durrett
    In Proceedings of the 13th International Conference on Learning Representations (ICLR), 2025

2024

  1. LoFiT: Localized Fine-tuning on LLM Representations
    Fangcong Yin, Xi Ye, and Greg Durrett
    In Proceedings of the 38th Conference on Neural Information Processing Systems (NeurIPS). Also presented at the Workshop on Foundation Model Interventions (Oral), NeurIPS 2024 , 2024

2023

  1. Linguistic Compression in Single-Sentence Human-Written Summaries
    Fangcong Yin and Marten van Schijndel
    In Findings of the Association for Computational Linguistics: EMNLP 2023, 2023