Analyzing LLM Instruction Optimization for Tabular Fact Verification

Xiaotang Du; Giwon Hong; Wai Chung Kwan; Rohit Saxena; Ivan Titov; Pasquale Minervini; Emily Allaway

Analyzing LLM Instruction Optimization for Tabular Fact Verification

Xiaotang Du, Giwon Hong, Wai-Chung Kwan, Rohit Saxena, Ivan Titov, Pasquale Minervini, Emily Allaway

Abstract

Instruction optimization provides a lightweight, model-agnostic approach to enhancing the reasoning performance of large language models (LLMs). This paper presents the first systematic comparison of instruction optimization, based on the DSPy optimization framework, for tabular fact verification. We evaluate four out-of-the-box prompting techniques that cover both text-only prompting and code use: direct prediction, Chain-of-Thought (CoT), ReAct with SQL tools, and CodeAct with Python execution. We study three optimizers from the DSPy framework—COPRO, MiPROv2, and SIMBA—across four benchmarks and three model families. We find that instruction optimization consistently improves verification accuracy, with MiPROv2 yielding the most stable gains for CoT, and SIMBA providing the largest benefits for ReAct agents, particularly at larger model scales. Behavioral analyses reveal that SIMBA encourages more direct reasoning paths by applying heuristics, thereby improving numerical comparison abilities in CoT reasoning and helping avoid unnecessary tool calls in ReAct agents. Across different prompting techniques, CoT remains effective for tabular fact checking, especially with smaller models. Although ReAct agents built with larger models can achieve competitive performance, they require careful instruction optimization.

Anthology ID:: 2026.findings-eacl.161
Volume:: Findings of the Association for Computational Linguistics: EACL 2026
Month:: March
Year:: 2026
Address:: Rabat, Morocco
Editors:: Vera Demberg, Kentaro Inui, Lluís Marquez
Venue:: Findings
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 3078–3108
Language:
URL:: https://aclanthology.org/2026.findings-eacl.161/
DOI:
Bibkey:
Cite (ACL):: Xiaotang Du, Giwon Hong, Wai-Chung Kwan, Rohit Saxena, Ivan Titov, Pasquale Minervini, and Emily Allaway. 2026. Analyzing LLM Instruction Optimization for Tabular Fact Verification. In Findings of the Association for Computational Linguistics: EACL 2026, pages 3078–3108, Rabat, Morocco. Association for Computational Linguistics.
Cite (Informal):: Analyzing LLM Instruction Optimization for Tabular Fact Verification (Du et al., Findings 2026)
Copy Citation:
PDF:: https://aclanthology.org/2026.findings-eacl.161.pdf
Checklist:: 2026.findings-eacl.161.checklist.pdf

PDF Cite Search Checklist Fix data