OPINE: A Prior-calibrated Scoring Framework for LLM-based Multi-label Scientific Opinion Classification

Mengting Zhang; Gaofeng Pan; Zhixiong Zhang; Yang Li; Guangyin Zhang

OPINE: A Prior-calibrated Scoring Framework for LLM-based Multi-label Scientific Opinion Classification

Mengting Zhang, Gaofeng Pan, Zhixiong Zhang, Yang Li, Guangyin Zhang

Abstract

Scientific opinion classification based on discourse functions provides a structured semantic basis for analytical tasks such as gap identification and hypothesis generation. However, this task is uniquely challenged by the multi-label nature of scientific expressions and AIMRaD structural constraints. Existing LLM-based methods typically rely on direct label generation, which obscures decision logic, or treat discourse information as passive context rather than a structural prior. We propose OPINE, a multi-stage framework that reformulates classification as a controllable *scoring-calibration-refinement* pipeline. By decoupling textual evidence from decision logic, OPINE generates independent label-wise affinity scores calibrated by AIMRaD priors. To resolve the multi-label challenge, we introduce a quantile-based decoding rule to naturally capture co-existing roles, alongside a pairwise refinement mechanism to mitigate confusion between similar categories. We contribute a new benchmark of 18 discourse functions across diverse sections. Experimental results show that OPINE generally outperforms strong baselines, reaching F1 scores of 63.20%, 53.68%, and 63.22% under Micro, Macro, and Example settings, respectively. Our analysis reveals that integrating discourse structures as explicit priors is superior to conventional passive context integration, while pairwise refinement successfully mitigates confusion between functionally similar categories. The code and dataset are available at https://github.com/znoodle63/OPINE.

Anthology ID:: 2026.findings-acl.1617
Volume:: Findings of the Association for Computational Linguistics: ACL 2026
Month:: July
Year:: 2026
Address:: San Diego, California, United States
Editors:: Maria Liakata, Viviane P. Moreira, Jiajun Zhang, David Jurgens
Venue:: Findings
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 32313–32333
Language:
URL:: https://aclanthology.org/2026.findings-acl.1617/
DOI:
Bibkey:
Cite (ACL):: Mengting Zhang, Gaofeng Pan, Zhixiong Zhang, Yang Li, and Guangyin Zhang. 2026. OPINE: A Prior-calibrated Scoring Framework for LLM-based Multi-label Scientific Opinion Classification. In Findings of the Association for Computational Linguistics: ACL 2026, pages 32313–32333, San Diego, California, United States. Association for Computational Linguistics.
Cite (Informal):: OPINE: A Prior-calibrated Scoring Framework for LLM-based Multi-label Scientific Opinion Classification (Zhang et al., Findings 2026)
Copy Citation:
PDF:: https://aclanthology.org/2026.findings-acl.1617.pdf
Checklist:: 2026.findings-acl.1617.checklist.pdf

PDF Cite Search Checklist Fix data