Yuchang Cheng

2025

Investigating Neurons and Heads in Transformer-based LLMs for Typographical Errors
Kohei Tsuji | Tatsuya Hiraoka | Yuchang Cheng | Eiji Aramaki | Tomoya Iwakura
Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing

This paper investigates how LLMs encode inputs with typos. We hypothesize that specific neurons and attention heads recognize typos and fix them internally using local and global contexts. We introduce a method to identify typo neurons and typo heads that work actively when inputs contain typos. Our experimental results suggest the following: 1) LLMs can fix typos with local contexts when the typo neurons in either the early or late layers are activated, even if those in the other are not. 2) Typo neurons in the middle layers are the core of typo-fixing with global contexts. 3) Typo heads fix typos by widely considering the context not focusing on specific tokens. 4) Typo neurons and typo heads work not only for typo-fixing but also for understanding general contexts.

pdf bib abs

SubRegWeigh: Effective and Efficient Annotation Weighing with Subword Regularization
Kohei Tsuji | Tatsuya Hiraoka | Yuchang Cheng | Tomoya Iwakura
Proceedings of the 31st International Conference on Computational Linguistics

NLP datasets may still contain annotation errors, even when they are manually annotated. Researchers have attempted to develop methods to automatically reduce the adverse effect of errors in datasets. However, existing methods are time-consuming because they require many trained models to detect errors. This paper proposes a time-saving method that utilizes a tokenization technique called subword regularization to simulate multiple error detection models for detecting errors. Our proposed method, SubRegWeigh, can perform annotation weighting four to five times faster than the existing method. Additionally, SubRegWeigh improved performance in document classification and named entity recognition tasks. In experiments with pseudo-incorrect labels, SubRegWeigh clearly identifies pseudo-incorrect labels as annotation errors. Our code is available at https://github.com/4ldk/SubRegWeigh.

2022

pdf bib abs

Sharing Parameter by Conjugation for Knowledge Graph Embeddings in Complex Space
Xincan Feng | Zhi Qu | Yuchang Cheng | Taro Watanabe | Nobuhiro Yugami
Proceedings of TextGraphs-16: Graph-based Methods for Natural Language Processing

A Knowledge Graph (KG) is the directed graphical representation of entities and relations in the real world. KG can be applied in diverse Natural Language Processing (NLP) tasks where knowledge is required. The need to scale up and complete KG automatically yields Knowledge Graph Embedding (KGE), a shallow machine learning model that is suffering from memory and training time consumption issues. To mitigate the computational load, we propose a parameter-sharing method, i.e., using conjugate parameters for complex numbers employed in KGE models. Our method improves memory efficiency by 2x in relation embedding while achieving comparable performance to the state-of-the-art non-conjugate models, with faster, or at least comparable, training time. We demonstrated the generalizability of our method on two best-performing KGE models 5^★E (CITATION) and ComplEx (CITATION) on five benchmark datasets.

2014

pdf bib

Detecting the Untranslatable Colloquial Expressions of Japanese Verbs in Cross-Language Instant Messaging
Yuchang Cheng | Masaru Fuji | Tomoki Nagase | Minoru Uegaki | Isaac Okada
Proceedings of the 28th Pacific Asia Conference on Language, Information and Computing

2012

pdf bib

An Example-Based Japanese Proofreading System for Offshore Development
Yuchang Cheng | Tomoki Nagase
Proceedings of COLING 2012: Demonstration Papers

Yuchang Cheng

2025

2022

2014

2012

2008

2007

2006

2005

Co-authors

Venues