Computational Linguistics Journal
@complingjournal
Computational Linguistics, established in 1974, is the official flagship journal of the Association for Computational Linguistics (ACL). Tags: #CLJournal #NLProc
Mis- and disinformation continue to propagate online & cause significant harm due to the rapid evolution of such content. Out-of-Distribution generalisation datasets can be very difficult to build & this is the major focus of this paper: doi.org/10.1162/COLI... #CLJournal @ioverho.eurosky.social
Have you noticed that preference datasets in reinforcement learning/DPO are not fully utilised? This article proposes a framework called Vote-based Preference Optimisation that leverages user voting data to better align language models with diverse subjective preferences: doi.org/10.1162/COLI...
Do you know what Multimodal OXYmorons are? If you're curious, read this research, which introduces this concept by constructing and communicating meaning through the interplay of multiple modalities: doi.org/10.1162/COLI... #CLjournal #NLProc @elidp.bsky.social
This paper introduces a wide range of encodings for constituent parsing as sequence adaptations of parsers from other paradigms that had not previously been cast under the sequence-labeling framework. Code: github.com/Polifack/CoD... and paper:https://doi.org/10.1162/COLI.a.603 @carlosg.bsky.social
How do prompt language and cultural framing influence model responses and their alignment with human values in different countries? Answer to this and more insights at: doi.org/10.1162/COLI... #NLProc #CLjournal @aylart.bsky.social
🏆The Computational Linguistics High Impact Paper Award recognises papers published in Computational Linguistics within the past 3–5 years that have already demonstrated exceptional scholarly impact. Congratulations to the authors of doi.org/10.1162/coli... #NLProc #CLjournal @chrmanning.bsky.social
🏆 2026 Computational Linguistics Distinguished Reviewer Awards: It aims to highlight and promote high standards in peer review and to acknowledge reviewers whose contributions have had a meaningful impact on the quality of our journal. Congratulations to our 2026 awardees! #NLProc #CLjournal
The ACL Computational Linguistics Doctoral Dissertation Award (ACL Dissertation Award) recognises three honourable mentions: Dennis Ulmer, Shen-Chieh Lin, and Haoran Xu. Congratulations to all three for this achievement! #NLproc #CLJournal @dnnslmr.bsky.social
We are pleased to announce the winner of the ACL Computational Linguistics Doctoral Dissertation Award (ACL Dissertation Award)! Congratulations, Verna Dankers! @vernadankers.bsky.social #NLproc #CLJournal
She shared invaluable research strategies that can help lead to success, and she encouraged every researcher never to take no for an answer. Her speech is published here: doi.org/10.1162/COLI... #CLJournal #NLProc
The second issue of Computational Linguistics Journal (Volume 52, Issue 2, June 2026) is finally out! Read it here: direct.mit.edu/coli/issue/5.... We'll introduce all the articles one by one soon. #CLJournal #NLProc
CL Journal has over 50 years of history in the field. While the field is moving, some of the articles from decades ago provide perspective and still relevant content. What do you think of these articles from 20 and 40 years ago? (2006 and 1986) & yes, they used to come in paper format too! #NLProc
Like humans, LLMs can be right for the wrong reasons. They can also be wrong for the wrong reasons. Anthropocentric bias has gone largely unexamined, posing a serious obstacle to the objective assessment of LLM capacities. Read more at doi.org/10.1162/COLI... #NLProc #CLJournal @raphaelmilliere.com
Interpretability provides a toolset for understanding how and why LMs behave in certain ways. This survey proposes a perspective on interpretability research grounded in causal mediation analysis: doi.org/10.1162/COLI... #NLProc #CLJournal @jannikbrinkmann.bsky.social @amuuueller.bsky.social
How Can We Effectively Expand the Vocabulary of LLMs with 0.01GB of Target Language Text? This article explores a very important question for low-resource languages by experimenting with various techniques across 10 languages. A must-read at:https://doi.org/10.1162/COLI.a.581 @gucciiiii.bsky.social
Should NLP metrics for bilingual code-switching use words as the token level or Intonation Units? Authors of this article show that intonation units will enhance comparisons between bilingual individuals, settings, and communities: doi.org/10.1162/COLI... #NLProc #NLP @pennlinguistics.bsky.social
Have you heard of soft metrics such as soft micro F1? In this article, the authors argue that for evaluating model predictions with human label variation, the standard metrics may not be sufficient. Read more at: doi.org/10.1162/COLI... @kmkurn.sigmoid.social.ap.brid.gy #NLProc #NLP
To assess whether multilingual NLP performs well across languages, we need to evaluate it on all world languages.That is not feasible.This paper implements two sampling methods from linguistic typology and provides a Python package to facilitate this: doi.org/10.1162/COLI... @andreashhp.bsky.social
Developing methods to assess the factuality of LLMs has become urgent. This paper presents LLM-OASIS for the factuality evaluation task. It turns out it significantly challenges SOTA LLMs. If you're up for improving over what's out there, start reading: doi.org/10.1162/COLI... #CLjournal #NLProc
Do current LLMs with fast-improving functional linguistic abilities exhibit distinct localization of formal (e.g., producing fluent, grammatical text) and functional (e.g., reasoning and consistent fact retrieval) linguistic mechanisms? Answer in: doi.org/10.1162/COLI... @michaelwhanna.bsky.social
If your research is around metaphor detection and interpretation, this article is for you! It introduces Meta4XNLI, the first parallel dataset for Natural Language Inference (NLI), in both English and Spanish: doi.org/10.1162/COLI... #NLProc @ragerri.bsky.social
What % of the NLP papers measure their impact in the real world? This paper proposes an impact evaluation of NLP models or systems for real-world usage, changing the research culture of NLP to focus more on real-world impact and less on SOTA-chasing: doi.org/10.1162/COLI... @ehudreiter.bsky.social
Hallucinations pose a substantial challenge to the reliability of LLMs in real-world scenarios. Zhang et al. survey methods of detection, explanation, & mitigation of hallucination, & provide a taxonomy & list of benchmarks for evaluation in this paper: doi.org/10.1162/COLI... @fredashi.bsky.social
Generative AI has advanced. But, for complex problems, it's lagging. In this position paper, the authors propose Human–AI Co-Construction (HAI-Co2), a framework for human–AI cooperative problem solving that facilitates such interaction. Read more at: doi.org/10.1162/COLI... @igurevych.bsky.social
Tasks such as summarization, QA & timeline creation need temporal expression normalization. The authors propose a novel method that deals with the known problems of temporal expression normalization on data scarcity, language and domain adaptation: doi.org/10.1162/COLI... @jmartinezromo.bsky.social
Have you heard of BLiMP for evaluating English LMs? Well, the authors of this paper introduce the Dutch version: BLiMP-NL. It's a benchmark set of linguistic minimal pairs for grammatical evaluation of Dutch LMs. Read more at: doi.org/10.1162/COLI... @stefanfrank.bsky.social @illc-uva.bsky.social
Interested in the semantic neighborhood of words? LGDE is a graph-based, local-diffusion framework for the discovery of keywords that are semantically similar to a seed dictionary. The authors release their code and data for a hate speech dataset too: doi.org/10.1162/COLI... #NLP
Tokenisation and Finite-State Transducers are two well-known concepts in computing. This article combines them for neural networks. Cognetta & Okazaki introduce a finite-state transduction framework that can encode all possible tokenizations of a regular language: doi.org/10.1162/COLI... #NLProc