*apologies for cross-postings*
Joint CODI CRAC 2026 Workshop: call for papers
�
July 2026 - ACL 2026 - San Diego, USA
�
We are pleased to announce that we are organizing the second joint CODI-CRAC workshop which will be held during ACL 2026! More information at:
�
<https://sites.google.com/view/codi-crac2026/home> https://sites.google.com/view/codi-crac2026/ �
�
CODI-CRAC is officially endorsed by SIGDial, the ACL Special Interest Group on Discourse and Dialogue.
�
Deadline for CODI CRAC papers: March 20 2026
�
The workshop will also host the CRAC shared task. More information at:
�
- CRAC shared task: <https://ufal.mff.cuni.cz/corefud/crac26> https://ufal.mff.cuni.cz/corefud/crac26
Aims and scope
�
Recent breakthroughs in NLP and Large Language Models have dramatically expanded our systems’ abilities to interpret and generate not just sentences, but whole documents and conversations. This shift has renewed interest in discourse-level challenges, driving new work on inter-sentential phenomena, coherence modeling, long-form summarization, discourse-aware representation learning, and large-scale resources for discourse understanding and parsing.
�
Discourse sits at the intersection of many NLP subfields, as it is where context, structure, and meaning come together beyond single sentences. Discourse shapes how we capture coherence, cohesion, and inference across long texts, and brings together researchers tackling the shared challenges of document structure, long-range dependencies, and the requirements of extended context.
�
In 2025, we organized the first joint CODI-CRAC workshop. The CODI workshop on Computational Approaches to Discourse has been a forum for a broad range of work at the discourse level. The CRAC workshop on Computational Models of Reference, Anaphora and Coreference has been a primary venue for researchers interested in the computational modeling of reference phenomena. Together, these workshops have catalyzed work to advance research on discourse-level problems and have served as a forum for discussing suitable datasets and reliable evaluation methods.
�
This joint edition corresponds to the 7th CODI workshop and the 9th CRAC workshop. It will welcome contributions from all the areas below, including state-of-the-art textual NLU and NLG work using LLMs, as well as classic structured work on automatic discourse analysis -- corresponding to challenging tasks such as coreference resolution or discourse parsing -- to encourage interaction between communities. The workshop is set to host the 5th edition of the CRAC shared task on Multilingual Coreference Resolution.
�
The workshop is planned as a 1-day event that brings together different subcommunities. It will feature regular papers and invited talks by Ruihong Huang (Texas A&M University) and Philippe Laban (Microsoft Research). We also accept papers accepted at other major conferences for non-archival presentation, including Findings papers.
�
Topics of interest
�
We welcome papers on symbolic and probabilistic approaches, corpus development and analysis, as well as machine and deep learning approaches to discourse. We appreciate theoretical contributions as well as practical applications, including demos of systems and tools. The goal of the workshop is to provide a forum for the community of NLP researchers working on all aspects of discourse.
�
Topics of interest include, but are not limited to:
�
- discourse structure
- discourse connectives
- discourse relations
- annotation tools and schemes for discourse phenomena
- corpora annotated with discourse phenomena
- discourse parsing
- cross-lingual discourse processing
- cross-domain discourse processing
- anaphora and coreference resolution
- event coreference
- argument mining
- coherence modeling
- discourse and semantics
- discourse in applications such as machine translation, summarization, etc.
- evaluation methodology for discourse processing
- discourse pretraining tasks
- long-text modeling and generation
�
Submissions
Double submission of papers is allowed, but this information will need to be disclosed at submission time.
�
We solicit three categories of papers: �
* (1) Regular workshop papers �
* (2) Demos
* (3) Extended abstracts
Only regular workshop papers and demos will be included in the proceedings as archival publications. Extended abstracts are non-archival and will be included in the workshop program and handbook, but will not appear in the workshop proceedings.
1- Regular papers must describe original unpublished research. �
* Long papers may consist of up to 8 pages of content, plus unlimited pages for references.
* Short papers can be up to 4 pages, plus unlimited pages for references.
2- Demo submissions may describe systems, tools, visualizations, etc., and may consist of up to 4 pages, plus unlimited pages for references.
3- Extended abstracts can describe work in progress. They may be two pages long (without references). Extended abstracts are non-archival. They will be included in the workshop program and handbook, but will not appear in the workshop proceedings.
Each submission can contain unlimited pages for Appendices, but the paper submissions need to remain fully self-contained, as these supplementary materials are completely optional, and reviewers are not even asked to review them.
Final versions of all types of papers will be given one additional page of content.
Paper accepted or rejected at one of the main conferences
�
We also invite presentations of papers accepted at another main conference. They will be included in the workshop program and handbook, but will not appear in the workshop proceedings.
We also fast-track ARR papers with existing reviews.
Submission website
�
All submissions must be anonymous and follow the ACL 2026 formatting instructions described here: <https://aclrollingreview.org/cfp> https://aclrollingreview.org/cfp �
�
Submission website:
* CODI-CRAC: <https://softconf.com/acl2026/codi-crac2026/> https://softconf.com/acl2026/codi-crac2026/ �
Schedule
Important dates for the workshop are listed below:
�
* CODI-CRAC papers due: March 20
* Pre-reviewed ARR fast-track (with reviews, can be accepted or rejected): April 5 �
* Notification of acceptance: April 28, 2026
* Grant application: May 5, 2026
* Camera-ready paper due: May 12, 2026
* Pre-recorded video due: June 4, 2026
* Workshop dates: July 3 or 4, 2026
�
�All deadlines are 11.59 pm UTC -12h ("anywhere on Earth").
Invited Speakers
�
- Ruihong Huang, Texas A&M University
- Philippe Laban, Microsoft Research
Organizers
�
- Chloé Braud, CNRS-IRIT
- Christian Hardmeier, IT University of Copenhagen
- Chuyuan (Lisa) Li, � University of British Columbia
- Jessy Li, University of Texas, Austin
- Sharid Loáiciga, University of Gothenburg
- Vincent Ng, University of Texas at Dallas
- Michal Novák, Charles University, Prague
- Maciej Ogrodniczuk, Institute of Computer Science, Polish Academy of Sciences
- Massimo Poesio, Queen Mary University of London and University of Utrecht
- Michael Strube, Heidelberg Institute for Theoretical Studies
- Amir Zeldes, Georgetown University, Washington DC
�
To contact the organizers, please send an email to: <mailto:codi-crac-workshop@googlegroups.com> codi-crac-workshop(a)googlegroups.com �
�
�
*SEM 2026: The 15th Joint Conference on Lexical and Computational Semantics [San Diego, CA (Co-located with ACL 2026)]
Website: https://starsem2026.github.io/
SEM brings together researchers interested in the semantics of natural languages and its computational modelling. The conference embraces a wide range of approaches including data-driven, neural, probabilistic and symbolic; practical applications as well as theoretical contributions are welcome. The long-term goal of *SEM is to provide a forum for NLP researchers working on any aspect of natural language semantics.
*SEM invites submissions related to the computational modelling of natural language semantics (understood broadly) and its application. Relevant areas include (but are not limited to) theoretical aspects of computational semantics, empirical and data-driven approaches, resources, evaluation, and applications/tools.
*SEM encourages authors to consider ethical aspects of their work, and to address and discuss ethical questions and implications relevant to their research. *SEM also values reproducibility and particularly welcomes submissions that adhere to the reproducibility guidelines as specified here.
Please fill out this form<https://forms.gle/634oW3yvtTkur6qL9> if you would like to volunteer as a reviewer or as an Area Chair.
Questions may be directed to: startsem-2026-pcs(a)googlegroups.com<mailto:startsem-2026-pcs@googlegroups.com>
New for *SEM 2026
1. One Day Conference: Unlike past iterations, *Sem 2026 will be a one-day conference. (ACL has informed us that this is due to venue size limitations.)
2. Centering Research Questions: Research questions in *Sem, and NLP generally, can be roughly categorized into those that address:
* new findings about language (linguistic phenomena, semantic patterns),
* new findings about people (language use, behavior, health, ethics, etc.),
* new findings about automatic language processing (advancing language understanding through ML/AI and other approaches).
Centering and explicitly articulating the research question helps authors frame and present their contribution more clearly. It also helps reviewers and Area Chairs evaluate the work within the appropriate context. For example, a paper that centers a compelling linguistic or behavioral research question and offers meaningful new insights need not also introduce methodological novelty or rely on the latest models (including LLMs). A simple and interpretable approach may make good sense.
To support this, the *SEM 2026 submission form asks authors to explicitly identify the predominant research question type for their work, as well as any additional categories that apply. There are no quotas for accepted papers of different types, and submissions will not receive preferential treatment based on category selection.
Including this information also allows *SEM to track the kinds of research questions authors pursue and how the conference’s focus evolves over time.
3. Lasting Impact
Modern NLP and ML papers have often been criticized for being overly incremental or becoming obsolete shortly after publication. To encourage work with broader scientific value and longer-term relevance, reviewers of *Sem 2026 will be asked to explicitly assess the potential lasting impact of each submission. This assessment will be included as a short-written justification and will factor into the overall recommendation.
Importantly, a healthy research ecosystem requires diversity in the time horizons of research contributions. Some papers offer immediate practical value; others generate insights or resources whose importance unfolds over years. *Sem 2026 welcomes this full spectrum. Reviewers should evaluate the potential for lasting influence—not only immediate performance gains.
Work can have a lasting impact in many ways. See our blog post<https://starsem2026.github.io/blog/> on this.
Topics of Interest (non-exhaustive)
* Compositional semantics and sentence representations
* Statistical, machine learning, and deep learning methods in semantic tasks
* Multilingual and cross-lingual semantics
* Word sense disambiguation and induction
* Sentiment Analysis, Computational Affective Science, Stylistic Analysis, and Argument Mining
* Computational Social Science, Digital Humanities, and Cultural Analytics
* Semantic parsing, and syntax-semantics interface
* Frame semantics and semantic role labeling
* Textual inference, textual entailment, and question answering
* Formal approaches to semantics
* Extraction of events and of causal and temporal relations
* Entity linking, pronouns and coreference
* Discourse, pragmatics, and dialogue
* Machine reading
* Abusive language detection, Fact verification and related tasks
* Extra-propositional aspects of meaning
* Multiword and idiomatic expressions
* Metaphor, irony, and humor processing
* Knowledge mining and acquisition
* Common sense reasoning
* Language generation
* Multidisciplinary research on semantics
* Grounding and multimodal semantics
* Psycholinguistics
* Interpretability and Explainability
* Human semantic processing
* Semantic annotation, evaluation, and resources
* NLP Applications
* Ethical aspects and bias in semantic representations
Submission Instructions
Submissions must describe unpublished work and be written in English. We solicit both long and short papers. Long papers describe original research and may consist of up to eight (8) pages of content, plus unlimited pages for references. Appendices are allowed after the references, but the paper should be self-contained, and reviewers will not be required to check the appendices, if any. Final versions of long papers will be given one additional page of content (up to 9 pages) so that reviewers' comments can be taken into account. Short papers describe original focused research and may consist of up to four (4) pages, plus unlimited pages for references. Upon acceptance, short papers will be given five (5) content pages in the proceedings. Authors are encouraged to use this additional page to address reviewers' comments in their final versions.
Limitations and Ethics Statement sections are allowed and encouraged but are not mandatory. These sections should be placed after the conclusion and will not count towards the overall page limit.
Submissions should follow the ARR formatting requirements<https://github.com/acl-org/acl-style-files>.
Submission routes and deadlines
*SEM solicits direct submissions (not through ARR). The deadline for direct submissions is Feb 13, 2026, and these submissions will be reviewed by the *SEM2025 program committee. Submissions are made through OpenReview.
Direct submission link:
https://openreview.net/group?id=aclweb.org/StarSEM/2026/Conference<https://url.au.m.mimecastprotect.com/s/vcD5CK1qwBSDAz9BmTvhvT5qRYM?domain=o…>
Multiple submission policy
*SEM does not prohibit the submission of work that is under consideration for another venue at the same time as the *SEM review period. However, authors of such papers will be asked to declare this at submission time.
Important Dates
(All deadlines are 11:59pm UTC-12h, AoE)
Direct submission deadline (long & short papers): Feb 13, 2026
Notification of acceptance: May 5, 2026
Camera-ready deadline: May 26, 2026
Conference date: July 6, 2026 (co-located with ACL 2026)
Following the ACL and ARR policies<https://www.aclweb.org/portal/content/report-acl-committee-anonymity-policy>, there is no anonymity period requirement.
Best regards,
The *SEM 2026 Program Chairs
In this newsletter:
Renew your LDC membership today
New publications:
CALLHOME Japanese Second Edition<https://catalog.ldc.upenn.edu/LDC2026S02>
CALLHOME Japanese Lexicon Second Edition<https://catalog.ldc.upenn.edu/LDC2026L01>
MATERIAL Swahili-English Language Pack<https://catalog.ldc.upenn.edu/LDC2026S01>
________________________________
Renew your LDC membership today
The importance of curated resources for language-related education, research, and technology development drives LDC's mission to create them, to accept data contributions from researchers across the globe, and to broadly share such resources through the LDC Catalog. LDC members enjoy no-cost access to new corpora released annually, as well as the ability to license legacy data sets from among our 1000 holdings at reduced fees. Ensure that your data needs continue to be met by renewing your LDC membership or by joining the Consortium today.
Now through March 2, 2026, any organization that joins the Consortium or renews their membership will receive a 10% discount off the 2026 membership fee. Membership remains the most economical way to access current and past LDC releases. Consult Join LDC<https://www.ldc.upenn.edu/members/join-ldc> for more details on membership options and benefits.
________________________________
New publications:
CALLHOME Japanese Second Edition<https://catalog.ldc.upenn.edu/LDC2026S02> was developed by LDC and contains 49 hours of speech from 120 telephone conversations between native Japanese speakers. This publication is a re-release of the original CALLHOME Japanese collection, combining CALLHOME Japanese Speech (LDC96S37)<https://catalog.ldc.upenn.edu/LDC96S37> and CALLHOME Japanese Transcripts (LDC96T18)<https://catalog.ldc.upenn.edu/LDC96T18> with additional transcription and updated directory structure, file formats, and documentation.
This corpus contains the 120 calls from CALLHOME Japanese Speech which represented training and development data and a subset of evaluation data. Participants spoke on topics of their choice in a single telephone call lasting up to 30 minutes. Calls were manually audited for language, recording quality, channel characteristics, dialect, and region. For this second edition, all audio was converted from SPHERE files to FLAC format, and the original training/development/test partitioning was removed.
This release also features revised transcripts conforming to updated LDC transcription guidelines that addressed normalization of annotation formats, standardization of speaker-produced and background noises, application of foreign-language marking, whitespace cleanup, and corrections and consistency fixes.
The CALLHOME series consists of telephone conversations and transcripts developed by LDC and Rutgers, The State University of New Jersey, in support of research in speaker identification, language identification, and related technologies. Languages in the series include American English, Egyptian Arabic, German, Japanese, Mandarin Chinese, and Spanish.
2026 members can access this corpus through their LDC accounts. Non-members may license this data for a fee.
*
CALLHOME Japanese Lexicon Second Edition<https://catalog.ldc.upenn.edu/LDC2026L01> was developed by LDC and contains 80,688 Japanese words with morphological, phonological, and stress information. This second edition updates file formats, directory structure, and documentation. The first edition is available as CALLHOME Japanese Lexicon (LDC96L17)<https://catalog.ldc.upenn.edu/LDC96L17>. The words in the lexicon were derived from 80 transcripts representing telephone conversations between native Japanese speakers contained in CALLHOME Japanese Second Edition (LDC2026S02)<https://catalog.ldc.upenn.edu/LDC2026S02>.
The lexicon contains seven tab-separated information fields: (1) headword: orthographic form in kanji or katakana or hiragana (if only written in hiragana); (2) hiragana: orthographic form in hiragana; (3) romanization: orthographic form in romaji; (4) pron: pronunciation of the headword; (5) morph: morphological analysis of the headword; (6) train freq: frequency of the headword in the transcripts; and (7) gloss: glosses of the headword. This release also includes a pronunciation dictionary derived from the lexicon in CMUdict<https://stdlib.io/docs/api/latest/@stdlib/datasets/cmudict> format and the grapheme-to-phoneme (G2P) tools used to automatically generate pronunciations for the original lexicon.
2026 members can access this corpus through their LDC accounts provided they have submitted a completed copy of the special license agreement. Non-members may license this data for a fee.
*
MATERIAL Swahili-English Language Pack<https://catalog.ldc.upenn.edu/LDC2026S01> was developed by Appen<http://www.appen.com/> for the IARPA MATERIAL<https://www.iarpa.gov/index.php/research-programs/material> program and contains 112 hours of Swahili conversational telephone speech, transcripts, English translations, annotations, and queries. Calls were made using different telephones (e.g., mobile, landline) from a variety of environments. Transcripts cover approximately 30% of the speech files, 3% of which were translated into English. This release also includes domain annotations, English queries, and their relevance annotations.
The MATERIAL program focused on underserved languages with the ultimate goal to build cross language information retrieval systems to find speech and text content using English search queries.
2026 members can access this corpus through their LDC accounts provided they have submitted a completed copy of the special license agreement. Non-members may license this data for a fee.
To unsubscribe from this newsletter, log in to your LDC account<https://catalog.ldc.upenn.edu/login> and uncheck the box next to "Receive Newsletter" under Account Options or contact LDC for assistance.
Membership Coordinator
Linguistic Data Consortium<ldc.upenn.edu>
University of Pennsylvania
T: +1-215-573-1275
E: ldc(a)ldc.upenn.edu<mailto:ldc@ldc.upenn.edu>
M: 3600 Market St. Suite 810
Philadelphia, PA 19104
========== First Call of Papers: SwissText 2026 ==========
Paper submission due: 23:59 AOE March 17, 2026
Conference date: June 10, 2026 in Zurich, Switzerland
Conference website: https://www.swisstext.org/
================================================
Dear colleagues,
We are pleased to announce the Call for Papers for SwissText 2026, the 11th edition of the Swiss Text Analytics Conference.
SwissText 2026 will take place on June 10, 2026, at the University of Zurich (Campus Oerlikon) in Zurich, Switzerland. SwissText is an established international forum for researchers and practitioners working on natural language processing, computational linguistics, and text analytics, with a strong tradition of fostering exchange between academia and industry.
We invite submissions of substantial, original, and unpublished work to the following tracks:
* Applied Track (non-archival), with a strong focus on industry and applied research.
* Scientific Track (archival), with technical research papers from the international scientific community, including corpus- and benchmark-related research papers with a focus on Swiss languages from the scientific community and industry.
* Corpus Track (archival), with Swiss-related NLP datasets.
* Demonstration Track (non-archival), with NLP systems presented live at the SwissText conference.
The special theme of SwissText 2026 is Reproducible NLP, we therefore encourage submissions working specifically in reproducible NLP research and fully open NLP.
We plan to publish the proceedings of SwissText 2026 in the ACL Anthology.
Important dates
* Submission deadline: March 17, 2026
* Notification of acceptance: April 21, 2026
* Camera-ready deadline: May 5, 2026
* Conference date: June 10, 2026
All deadlines are at 11:59PM UTC-12:00 AOE (“anywhere on Earth”).
Detailed submission guidelines and formatting instructions can be found on the conference website: https://www.swisstext.org/call-for-papers/
General Chair: Prof. Dr. Rico Sennrich, University of Zurich
Organizing Committee: Jannis Vamvas, Yingqiang Gao, Tilia Ellendorff, Michelle Wastl, Gerold Schneider, University of Zurich
For questions, please contact info(a)swisstext.org<mailto:info@swisstext.org> or the organizing committee members.
Best regards,
Dr. Yingqiang Gao (he/him)
Department of Computational Linguistics
Andreasstrasse 15, Office AND 2-20
University of Zurich, CH-8050 Zurich
There is an open part-time (50%) Faculty position at the University of Hildesheim in Germany to fill for 3 years: Pre- or Postdoc (Translation Studies, Applied Linguistics, Computational Linguistics), knowledge of both German and English is obligatory. The announcement in German provides more details:
https://bewerbung.uni-hildesheim.de/jobposting/393fc6472ac442461bd082e36807…
--
Prof. Dr. Ekaterina Lapshinova-Koltunski
Mehrsprachige technische Fachkommunikation
Geschäftsführende Direktorin
Institut für Übersetzungswissenschaft und Fachkommunikation
Fachbereich 3: Sprach und Informationswissenschaften
Stiftung Universität Hildesheim
Lübecker Straße 3
31141 Hildesheim
+49 5121 883-30934
Second Call for Papers
*****************
NooJ 2026 International Conference
Naples, Italy
June 24-26, 2026
https://nooj2026.sciencescall.org/resource/page/id/2
*******************
Important dates:
*******************
Abstract submission: 31 January 2026
Notification of acceptance: 25 March 2026
Registration: until 13 April 2026
Conference dates: 24-26 June 2026
***********************************************
University of Naples "L'Orientale" and the NooJ association organize the 20th NooJ Conference in Naples, Italy from 24-26 June, 2026.
NooJ is a linguistic development environment that allows linguists to formalize several levels of linguistic phenomena: orthography and spelling; lexicons of simple words, multiword units and frozen expressions; inflectional, derivational and agglutinative morphology; local, constituent and dependency syntax; transformational grammars and semantic analysis. For each phenomenon, NooJ provides linguists with formal tools specifically adapted to facilitate the description, using the four types of Chomsky-Schützenberger formal grammars (regular, context-free, context-sensitive and unrestricted). This approach distinguishes NooJ from most computational linguistic frameworks which provide a single formalism.
NooJ is also a corpus processing tool, used in the digital humanities (in History, Literature, Psychology and Sociolinguistics) as it allows users to apply sophisticated linguistic resources to large corpora and build indices and concordances, annotate texts automatically, perform various statistical analyses, etc.
NooJ is freely available and linguistic modules can already be freely downloaded for over 30 languages, see https://nooj.univ-fcomte.fr
A Web demo is available for English, French, Spanish and Ukrainian at: https://webnooj.univ-fcomte.fr
******************************
The conference intends to:
******************************
* give NooJ users and researchers in Linguistics, Computational Linguistics and in the Digital Humanities the opportunity to meet and share their experience as developers, researchers and teachers;
* present to NooJ users the latest linguistic resources and NLP applications developed for/with NooJ, its latest functionalities, as well as its future developments;
* offer researchers and graduate students an advanced tutorial dedicated to the automatic transformational analysis/generation of texts.
*******************
Topics of interest:
*******************
* Lexical resources
* Computational morphology
* Syntactic analysis
* Semantic analysis
* Linguistic-based NLP applications
***************
Submission:
***************
We invite the submission of abstracts in English until 31 January 2026. The abstracts should contain the title, name and email of the author(s) and their institutions. Abstracts should not exceed one page (between 400 and 600 words) and should be sent to nooj2026(a)gmail.com. All proposals will be reviewed by the members of the scientific committee; authors will be given notice of acceptance of their papers no later than 25 March 2026.
Further information about the conference can be found at https://nooj2026.sciencescall.org/resource/page/id/5. You can also contact the organizing committee at nooj2026(a)gmail.com for any additional information.
************************
Scientific Committee:
************************
Marco Angster, University of Zadar, Croatia
Anabela Barreiro, INESC-ID, Portugal
Anita Bartulović, University of Zadar, Croatia
Magali Bigey, Université de Franche-Comté, France
Xavier Blanco, Autonomous University of Barcelona, Spain
Christian Boitet, Université Joseph Fourier, Grenoble, France
Maria Pia Di Buono, Università degli Studi di Napoli l'Orientale, Italy
Héla Fehri, University of Sfax, Tunisia
Zoe Gavriilidou, Democritus University of Thrace, Greece
Yuras Hetsevich, National Academy of Sciences, Belarus
Agata Jackievicz, Université Paul Valéry, France
Agnieszka Kaliska, Poznan University, Poland
Kristina Kocijan, University of Zagreb, Croatia
Walter Koza, National, University of General Sarmiento, Argentina
Svetlana Krylosova, INALCO, France
Mathieu Lafourcade Université de Montpellier, France
Laetitia Leonarduzzi, Université d’Aix-Marseille, France
Stefania Maci, Università di Bergamo, Italy
Samir Mbarki, IbnTofail University, Morocco
Linda Mijić, University of Zadar, Croatia
Johanna Monti, Università degli Studi di Napoli l'Orientale, Italy
Kamal Naït-Zerrad, INALCO, France
Thierry Poibeau, Laboratoire Lattice, CNRS, France
Andrea Rodrigo, University of Rosario, Argentina
Olena Saint-Joanis, INALCO, France
Max Silberztein, Université de Bourgogne Franche-Comté, France
Marko Tadić, University of Zagreb, Croatia
François Trouilleux, Université Clermont Auvergne, France
**************************
Organizing Committee:
**************************
* Johanna Monti, Università di Napoli L'Orientale, Italy
* Maria Pia di Buono, Università di Napoli L'Orientale, Italy
* Max Silberztein, Université de Franche-Comté, France
-----------------------------------------------------------------------------
Call for submissions
1st International Workshop on Quality in Large Language Models and
Knowledge Graphs
In conjunction with EDBT/ICDT 2026
QuaLLM-KG @ EDBT/ICDT 2026
24 March 2026, Tampere, Finland
Website: https://quallmkg2026.github.io/
*New deadline: January 25th AoE*
-----------------------------------------------------------------------------
**** Goal ****
QuaLLM-KG aims to bring together researchers and practitioners working
on quality issues at the intersection of large language models and
knowledge graphs. The workshop focuses on theories, methods, and
applications for assessing, improving, and monitoring the quality of
LLMs and KGs.
**** Important Dates ****
- Submission deadline: January 25th, 2026
- Notification: February 8th, 2026
- Camera-ready: February 20th, 2026
**** Topics ****
* Quality in Knowledge Graphs
- Accuracy, consistency, completeness, freshness
- Schema validation, constraint checking, error detection
- Entity resolution, link prediction, ontology alignment
- Provenance, explainability, trust in KG data
- KG quality in dynamic and large-scale settings
* Quality in Large Language Models
- Hallucination reduction & factual grounding
- Bias detection and mitigation
- Metrics & benchmarks for quality assessment
- Uncertainty estimation, calibration, interpretability
* Synergies Between KGs and LLMs
- KG-based grounding and fact-checking for LLMs
- LLM-based KG enrichment, extraction, entity linking
- Quality-driven prompting and fine-tuning
- Hybrid KG–LLM architectures for quality assurance
- Evaluation frameworks for integration and consistency
* Benchmarks and Evaluation Frameworks
- Datasets and metrics for KG & LLM quality
- Tools for monitoring, validation, maintenance
- Reproducibility, transparency, responsible AI
* Applications and Case Studies
- Scientific, industrial, enterprise use cases
- Quality at scale
- Human-in-the-loop quality control
**** Submissions ****
We invite submissions of full papers (up to 8 pages, excluding
references) and short papers describing work in progress, systems,
demos/systems/applications,
or vision/innovative ideas (up to 4 pages, excluding references).
Submissions should be in the CEUR-WS proceedings template.
Accepted papers will be published in the CEUR Workshop proceedings
(CEUR-WS.org).
**** Workshop Organizers ****
- Soror Sahri, Université Paris Cité, France
- Sven Groppe, University of Lübeck, Germany
- Farah Benamara, IPAL-CNRS, Singapore & University of Toulouse
--
========================
Farah Benamara Zitoune
Professor in Computer Science, Université de Toulouse
IRIT and IPAL-CNRS Singapore
118 Route de Narbonne, 31062, Toulouse.
Tel : +33 5 61 55 77 06
http://www.irit.fr/~Farah.Benamara
==================================
**Second Call for Papers**
Gaze4NLP - The Second Workshop on Gaze Data and Natural Language Processing
12 May 2026, Palma de Mallorca, Spain (co-located with LREC 2026)
https://gaze4nlp.github.io/Gaze4NLP2026/
The Second Workshop on Gaze Data and Natural Language Processing
(Gaze4NLP), co-located with LREC 2026 in Palma de Mallorca, Spain,
invites papers of a theoretical or experimental nature describing
research methodologies by employing interdisciplinary perspectives,
including computer science and engineering perspectives and cognitive
sciences, and identifying challenges to resolve in the intersection of
the two domains: eye tracking and NLP. Gaze4NLP aims to bring together
researchers conducting research on eyes on eyes on text and NLP; and
establishing bridges between them for identifying future venues of
research.
Workshop webpage:
https://gaze4nlp.github.io/Gaze4NLP2026/
Important Dates
Workshop paper submission deadline: 16 February 2026
Workshop paper acceptance notification: 16 March 2026
Workshop paper camera-ready versions: 30 March 2026
Workshop date: 12 May 2026
All deadlines are 11:59PM UTC-12:00 (anywhere on Earth)
Topics for the workshop will include, but are not limited to:
- Investigating the pillars for bridging the gap between the research
on eyes on text and NLP. Study how to expand research methodologies
by employing interdisciplinary perspectives, including computer
science and engineering perspectives and cognitive sciences, and
identify challenges, issues to resolve.
- Exploring new areas so that both fields benefit from each other
better than the past, identifying novel domains of exploration for
further research.
- Discussing how to develop cognitively inspired models that align
human reading data with LLMs.
Submissions
We solicit regular workshop papers, which will be included in the
proceedings as archival publications. The length of the papers should
be between 4 and 8 pages (excluding references). The submissions
should not include any appendices. Accepted papers will be presented
in the form of either oral or poster presentations.
Please note that camera-ready papers are allowed an additional page of
content to address reviewer comments, and unlimited pages for
appendices. The workshop proceedings will be part of the ACL
anthology. Accepted papers will also be given an opportunity with an
extended version to be published as part of an edited book.
Submissions will be handled via the START Conference Manager.
- Submission link: https://softconf.com/lrec2026/Gaze4NLP/
All submissions should follow the LREC style guidelines. We strongly
recommend the use of the LaTeX style files, OpenDocument, or Microsoft
Word templates created for LREC: <https://lrec2026.info/authors-kit/>.
All papers must be anonymous, i.e., not reveal author(s) on the title
page or through self-references. So, e.g., “We previously showed
(Smith, 2020)”, should be avoided. Instead, use citations such as
“Smith (2020) previously showed”.
LRE-Map and Sharing Language Resources
When submitting a paper from the START page, authors will be asked to
provide essential information about resources (in a broad sense, i.e.
also technologies, standards, evaluation kits, etc.) that have been
used for the work described in the paper or are a new result of your
research. Moreover, ELRA encourages all LREC authors to share the
described LRs (data, tools, services, etc.) to enable their reuse and
replicability of experiments (including evaluation ones).
Organization Committee:
Cengiz Acarturk, Jagiellonian University, Poland
Jamal Nasir, University of Galway, Ireland
Burcu Can, University of Stirling, Scotland, UK
Cagri Coltekin, University of Tubingen, Germany
School of Computer Science at University of Leeds has
two UK DLA PhD scholarships for this year (Oct 2026)
with a deadline of Friday 30th January:
https://phd.leeds.ac.uk/project/2357-epsrc-dla-scholarship-in-the-school-of…
Please encourage any eligible students to apply, as we have not had much demand so far so the success rate is likely to be quite good. These can be for any project in the school, starting October 2026, lasting 3.5 years.
This year both are for UK home-fee rated students only; applicants can check their eligibility here:
https://www.ukcisa.org.uk/student-advice/find-a-category-for-he-england
Eric Atwell, Professor of Artificial Intelligence for Language
School of Computer Science, Uni of LEEDS, LS2 9JT, UK
http://www.comp.leeds.ac.uk/eric
English version below
Bonjour,
Dans le cadre du projet DataLens, nous proposons un stage de M2 en
Machine Learning et Web sémantique visant à améliorer la complétion et
la structuration des métadonnées de jeux de données pour faciliter la
fédération de sources hétérogènes.
Plus d’informations et candidature :
https://recrutement.inria.fr/public/classic/fr/offres/2025-09456.
------------------------------------------------------------------------
Hello,
As part of the DataLens project, we are offering a Master’s internship
in Machine Learning and the Semantic Web focused on improving the
completion and structuring of dataset metadata to support the federation
of heterogeneous sources.
More information and application:
https://recrutement.inria.fr/public/classic/fr/offres/2025-09456.
Best regards,
--
Anaïs OLLAGNIER
Assistant Professor at Université Côte d'Azur | I3S | INRIA wimmics team
--------------------------------------------
Templiers 1, Bureau 417, 930 Route des Colles, BP 145
06903 Sophia Antipolis Cedex, France
anais.ollagnier(a)inria.fr |https://aollagnier.github.io/