Jiayi Zhang
Other people with similar names: Jiayi Zhang
Unverified author pages with similar names: Jiayi Zhang
2026
Concise Math Reasoning via Difficulty-Aware Distillation
Yifan Wu | Jingze Shi | Bingheng Wu | Jiayi Zhang | Xiaotian Lin | Yizhang Zhu | Zhaoyang Yu | Bang Liu | Chenglin Wu | Nan Tang | Yuyu Luo
Findings of the Association for Computational Linguistics: ACL 2026
Yifan Wu | Jingze Shi | Bingheng Wu | Jiayi Zhang | Xiaotian Lin | Yizhang Zhu | Zhaoyang Yu | Bang Liu | Chenglin Wu | Nan Tang | Yuyu Luo
Findings of the Association for Computational Linguistics: ACL 2026
Human experts tackle difficult math problems by identifying and executing a few pivotal steps rather than listing every intermediate thought. In contrast, standard Chain-of-Thought (CoT) distillation trains small models on lengthy reasoning traces, encouraging a uniform overthinking style across easy and hard items alike. The result is rigid, slow solutions that sacrifice adaptivity. This approach stands in sharp contrast to human intuition. Humans naturally adapt their problem-solving strategy, dedicating significant effort to difficult problems while finding quick, simple solutions for easier ones. We argue that the root cause lies in the training data: it contains excess information and reasoning steps organized in ways misaligned with human practice. We address this with Difficulty-Aware Distillation(DAD), a procedure for producing training data that mirrors concise human reasoning. A large teacher model first assesses a problem’s difficulty and then rewrites the solution to retain only the essential steps. Using this process, we constructed LiteCoT, a 100,000-example corpus of short, clear rationales, and used it to train our Liter models. With 100k LiteCoT, we outperform models trained on 800k long CoT and cut both training and inference costs. The advantage is consistent across standard math benchmarks, showing that concise, human-aligned data delivers equal or better accuracy with much less compute. For example, on the challenging AIME24 exam, our approach reaches 74.2% Pass@1 using only about 5K inference tokens, surpassing other methods that consume many more tokens.
2025
Self-Supervised Prompt Optimization
Jinyu Xiang | Jiayi Zhang | Zhaoyang Yu | Xinbing Liang | Fengwei Teng | Jinhao Tu | Fashen Ren | Xiangru Tang | Sirui Hong | Chenglin Wu | Yuyu Luo
Findings of the Association for Computational Linguistics: EMNLP 2025
Jinyu Xiang | Jiayi Zhang | Zhaoyang Yu | Xinbing Liang | Fengwei Teng | Jinhao Tu | Fashen Ren | Xiangru Tang | Sirui Hong | Chenglin Wu | Yuyu Luo
Findings of the Association for Computational Linguistics: EMNLP 2025
Well-designed prompts are crucial for enhancing Large language models’ (LLMs) reasoning capabilities while aligning their outputs with task requirements across diverse domains. However, manually designed prompts require expertise and iterative experimentation. While existing prompt optimization methods aim to automate this process, they rely heavily on external references such as ground truth or by humans, limiting their applicability in real-world scenarios where such data is unavailable or costly to obtain. To address this, we propose Self-Supervised Prompt Optimization (SPO), a cost-efficient framework that discovers effective prompts for both closed and open-ended tasks without requiring external reference. Motivated by the observations that prompt quality manifests directly in LLM outputs and LLMs can effectively assess adherence to task requirements, we derive evaluation and optimization signals purely from output comparisons. Specifically, SPO selects superior prompts through pairwise output comparisons evaluated by an LLM evaluator, followed by an LLM optimizer that aligns outputs with task requirements. Extensive experiments demonstrate that SPO outperforms state-of-the-art prompt optimization methods, achieving comparable or superior results with significantly lower costs (e.g., 1.1% to 5.6% of existing methods) and fewer samples (e.g., three samples).
Data Interpreter: An LLM Agent for Data Science
Sirui Hong | Yizhang Lin | Bang Liu | Bangbang Liu | Binhao Wu | Ceyao Zhang | Danyang Li | Jiaqi Chen | Jiayi Zhang | Jinlin Wang | Li Zhang | Lingyao Zhang | Min Yang | Mingchen Zhuge | Taicheng Guo | Tuo Zhou | Wei Tao | Robert Tang | Xiangtao Lu | Xiawu Zheng | Xinbing Liang | Yaying Fei | Yuheng Cheng | Yongxin Ni | Zhibin Gou | Zongze Xu | Yuyu Luo | Chenglin Wu
Findings of the Association for Computational Linguistics: ACL 2025
Sirui Hong | Yizhang Lin | Bang Liu | Bangbang Liu | Binhao Wu | Ceyao Zhang | Danyang Li | Jiaqi Chen | Jiayi Zhang | Jinlin Wang | Li Zhang | Lingyao Zhang | Min Yang | Mingchen Zhuge | Taicheng Guo | Tuo Zhou | Wei Tao | Robert Tang | Xiangtao Lu | Xiawu Zheng | Xinbing Liang | Yaying Fei | Yuheng Cheng | Yongxin Ni | Zhibin Gou | Zongze Xu | Yuyu Luo | Chenglin Wu
Findings of the Association for Computational Linguistics: ACL 2025
Large Language Model (LLM)-based agents have excelled in various domains but face significant challenges when applied to data science workflows due to their complex, multi-stage nature. Current LLM-based agents struggle with non-linear relationships, recursive dependencies, implicit data- and logic-dependent reasoning, and managing extensive context. In this paper, we introduce Data Interpreter, an LLM-based agent that addresses these challenges through hierarchical graph-based modeling to represent the complexity and a progressive strategy for step-by-step verification, refinement, and consistent context management. Extensive experiments confirm the effectiveness of Data Interpreter. On InfiAgent-DABench, it boosts performance by 25% (from 75.9% to 94.9%), and on machine learning and open-ended tasks, it lifts accuracy from 88% to 95% and from 60% to 97%, respectively. Moreover, our method surpasses state-of-the-art baselines by 26% on the MATH dataset. We will release the code upon publication.
Search
Fix author
Co-authors
- Yuyu Luo 3
- Chenglin Wu 3
- Sirui Hong 2
- Xinbing Liang 2
- Bang Liu 2
- Zhaoyang Yu 2
- Jiaqi Chen 1
- Yuheng Cheng 1
- Yaying Fei 1
- Zhibin Gou 1
- Taicheng Guo 1
- Danyang Li 1
- Xiaotian Lin 1
- Yizhang Lin 1
- Bangbang Liu 1
- Xiangtao Lu 1
- Yongxin Ni 1
- Fashen Ren 1
- Jingze Shi 1
- Nan Tang 1
- Robert Tang 1
- Xiangru Tang 1
- Wei Tao 1
- Fengwei Teng 1
- Jinhao Tu 1
- Jinlin Wang 1
- Bingheng Wu 1
- Binhao Wu 1
- Yifan Wu 1
- Jinyu Xiang 1
- Zongze Xu 1
- Min Yang 1
- Ceyao Zhang 1
- Li Zhang 1
- Lingyao Zhang 1
- Xiawu Zheng 1
- Tuo Zhou 1
- Yizhang Zhu 1
- Mingchen Zhuge 1