Venue: Federation of the Scientific and Technical Unions in Bulgaria (FNTS) – 108 Georgi S. Rakovski St

The tutorials, the workshops and the main conference will take place at the same building.
Each workshop will run in parallel with the other and with the tutorial sessions on 7 September 2026.
The registration for all events will take place on Floor 2.

View parallel sessions schedule
CLIB 2026 Schedule

PROGRAMME | 7 September 2026 | TUTORIALS

08:30 – 9:00 | Registration (Floor 2, Lobby)

TUTORIALS | Floor 3, Hall 312

09:00 – 12:30 | Ollama, No Drama: Step-by-Step Guide of Practical Local AI

Lecturer: Dimitar Hristov (Institute for Bulgarian Language, BAS)

09:00 – 10:30 | Tutorial Session

10:30 – 11:00 | Coffee Break

11:00 – 12:30 | Tutorial Session

12:30 – 13:30 | Lunch Break

13:30 – 17:00 | Composing Theoretical and Computational Linguistic Problems

Lecturer: Ivan Derzhanski (Institute of Mathematics and Informatics, BAS)

13:30 – 15:00 | Tutorial Session

15:00 – 15:30 | Coffee Break

15:30 – 17:00 | Tutorial Session

17:00 – 17:15 | Break

17:15 – 18:15 | KEYNOTE TALK: Natural Language Processing in the Era of Artificial Intelligence: The Wind of Change is Blowing

Speaker: Prof. Ruslan Mitkov (Lancaster University, University of Alicante)

18:30 – 21:00 | Welcome reception at Panorama Rooftop (Floor 6 of the FNTS building)

PROGRAMME | 8 September 2026 | MAIN CONFERENCE

Floor 2, Hall 4

08:30 – 9:00 | Registration

09:00 – 09:15 | Conference opening
Welcome Addresses
Prof. Sylvia Ilieva, Director of GATE Institute
Iliya Krastev, CEO at A Data Pro
Veselin Georgiev, CTO at Plan A

Trustworthy AI and Multimodal Modelling

Chair: Ivan Koychev

09:15 – 10:15 | KEYNOTE TALK: Factuality Challenges in the Era of Large Language Models

Speaker: Prof. Preslav Nakov (Mohamed bin Zayed University of Artificial Intelligence)

10:15 – 10:20 | Break

10:20 – 11:00 | Session Presentations

  • 10:20 – 10:40Irina Temnikova, Devora Kotseva, Ruslana Margova, Todor Kiriakov, Yolina Petrova, Nevena Grigorova, Alexander Komarov, Bogomil Katanov, Boryana Kostadinova and Stefan Minkov
    PROPer-BG: A New Publicly Accessible Package of Resources for Propaganda Detection in Bulgarian
  • 10:40 – 11:00Jordan Kralev
    Direct Projection of Visual Features into Linguistic Embeddings for Domain-Adapted Multimodal Learning

11:00 – 11:30 | Coffee Break

Specialised Methods in NLP and AI

Chair: Milena Dobreva

11:30 – 12:30 | Session Presentations

  • 11:30 – 11:50Svetla Boytcheva, Sylvia Vassileva and Pavel Boytchev
    BGSynthMedic: Synthetic Parallel English-Bulgarian Discharge Summary Corpus with Automatic Annotations
  • 11:50 – 12:10George Pashev, Silvia Gaftandzhieva
    SymLang: A Structured Symbolic Intermediate Representation for Optimized Human-AI Communication in Assistive Clinical Technologies
  • 12:10 – 12:30Yolina Petrova and Nikola Blajev
    Zero-Shot Detection of AI-Generated Bulgarian Text via Binoculars Framework

12:30 – 13:30 | LUNCH BREAK AND POSTER SESSION

POSTER SESSION

Chair: Hristina Kukova

  • Mullosharaf Arabov
    A Systematic Benchmark of Machine Transliteration Models for the Tajik-Farsi Language Pair: A Comparative Study from Rule-Based to Transformer Architectures
  • Mullosharaf Arabov, Karomatullo Habibullozoda and Nurali Shirinov
    TajikNLP: An Open-Source Toolkit for Comprehensive Text Processing of Tajik (Cyrillic Script)
  • Öner Aytaş, Mehmet Ali Bayram, Tuğçe Şen, Banu Diri and Göksel Biricik
    Mitigating Security Vulnerabilities in Open-Source Large Language Models via a Supervisor Agent Architecture
  • Abdelouahab Baflah
    A Qualitative Error Analysis of Whisper-Based Sentiment Recognition in Algerian Spoken Arabic: Prosodic and Pragmatic Failures in a Low-Resource Dialect
  • Mehmet Ali Bayram, Öner Aytaş, Tuğçe Şen
    Morphology-Aware Tokenization Improves Turkish Sentence Embeddings under Controlled Training Conditions
  • Nevena Grigorova, Stanislav Penkov, Keith Kiely, Dobromir Tsolyov, Iglika Ivanova
    Syntactic Patterns in Political Speech: Dependency-Based Ad Hominem Detection in Bulgarian Social Media
  • Maria Khokhlova
    A Comparative Study of Synonym Generation in Russian with Instruction-Tuned Large Language Models in a Zero-Shot Setting
  • Maria Khokhlova, Mikhail Koryshev
    Thematic Dominants of German-Language Catholic Hymns: Automatic Analysis using Word Embeddings
  • Ruslana Margova
    Detecting Pseudoscientific Narratives in Bulgarian Science Communication: Nationalistic Mystification, Pseudo-Expertise and Cultural Blind Spots in AI-Assisted Analysis
  • Mila Marcheva-Nash and Weiwei Sun
    PS-CHILDES-BG: A Constituency Treebank of Bulgarian Morphemically Tokenised Child-directed Speech
  • Brandon Caillahua Mendoza
    Genetic Algorithms with Normalized Fitness Functions for Low-Resource Language Identification: A Case Study on Archaic Romanian and Medieval Slavonic Liturgical Manuscripts
  • Junya Morita
    Corpus-based Study of Word Formation: Anti-agentive Nominals in English and Japanese

Language Models and Computational Methods for Low-Resource Languages

Chair: Rositsa Dekova

13:30 – 15:10 | Session Presentations

  • 13:30 – 13:50Amina Laggoun, Youness Moukafih, Ouassim Karrakchou, Kamel Smaili
    Towards Frugal BERT Models for Maghrebi Dialects: Pruning and Distillation Approaches
  • 13:50 – 14:10Ján Mačutek, Michaela Koščová, Michaela Nogolová and Radek Čech
    What to Take as the Basic Verb Form? A Pilot Study in Czech
  • 14:10 – 14:30Radu Ion, Verginica Barbu Mititelu, Elena Irimia, Maria Mitrofan and Vasile Păiș
    Benchmarking Language Models on Romanian Eighth-Grade National Exams
  • 14:30 – 14:50Silvie Cinková, Ivana Kvapilíková and Jan Černý
    Lexical Surprisal from Large Language Models: Cautionary Notes for Legal Discourse
  • 14:50 – 15:10Natalia Dankova
    Analysing French Narrative Texts Generated by ChatGPT-5

15:10 – 15:30 | Coffee Break

CORPORA, ANNOTATION SYSTEMS, AND NLP TOOLS

Chair: Mariana Damova

15:30 – 17:10 | Session Presentations

  • 15:30 – 15:50Verginica Barbu Mititelu, Elena Irimia, Simina Popa, Mihaela Cristescu, Catalin Mihaila, Lea Radu, Ioana Bucsa
    The Romanian Component of the Parallel ELEXIS Corpus: ELEXIS-Ro
  • 15:50 – 16:10Atanas Atanasov
    A Dual-Interface Syntax Tree Editor with Automated Assessment for Learning Management Systems
  • 16:10 – 16:30Maria Sampedro Mella
    Developing a Pragmatic Annotation System for a Spoken Corpus
  • 16:30 – 16:50Tsvetelina Stefanova, Petya Osenova
    Evaluating the Linguistic Properties of AI-Generated Bulgarian News Texts
  • 16:50 – 17:10Svetla Koeva
    Towards a New Grammar Checker for Bulgarian

18:00 – 20:00 | Guided walking tour (starting point: the Alexander Nevsky Cathedral)

PROGRAMME | 9 September 2026 | MAIN CONFERENCE

Floor 2, Hall 4

09:15 – 10:15 | KEYNOTE TALK: TBA

Speaker: Dr Khalid Choukri (European Language Resources Association)

10:15 – 10:20 | Break

Corpus Building and LLM Benchmarking for Language Understanding

Chair: Hristo Dolutarov

10:20 – 11:00 | Session Presentations

  • 10:20 – 10:40Karla Csuros, Madalina Chitez, Roxana Rogobete and Petya Osenova
    Building a Cross-lingual Corpus of BA Theses: Rationale and Preliminary Observations
  • 10:40 – 11:00Maria Todorova, Valentina Stefanova, Tsvetana Dimitrova, Mihaela Moskova, Hristina Kukova, Svetlozara Leseva, Ivelina Stoyanova, Dimitar Georgiev and Svetla Koeva
    Hybrid Benchmark for Long-Text Understanding and Reasoning in Bulgarian

11:00 – 11:30 | Coffee Break

SPECIAL SESSION ON SEMANTIC RELATIONS, FRAMENETS AND ONTOLOGIES

Chair: Yovka Tisheva

11:30 – 12:30 | Session Presentations

  • 11:30 – 11:50Marina Bagi, Aleksandra Marković and Ranka Stanković
    Toward a Serbian FrameNet: Resource Construction, Methodology, and Initial Annotation Results
  • 11:50 – 12:10Stanislav Penkov
    A WordNet-Aligned Taxonomy of Logical Fallacies for Low-Resource NLP
  • 12:10 – 12:30Ranka Stanković, Cvetana Krstev, Milica Ikonić Nešić and Miloš Utvić
    Relation Extraction from Text of Various Domains: A Serbian Case Study

12:30 – 13:30 | LUNCH BREAK AND POSTER SESSION

Cognitive and Corpus-Based Approaches to Language Processing

Chair: Georgi Iliev

13:30 – 15:10 | Session Presentations

  • 13:30 – 13:50Andrea Nadalini, Claudia Marzi, Marcello Ferro, Vito Pirrelli, Maria Todorova, Valentina Stefanova, Tsvetana Dimitrova, Hristina Kukova and Svetla Koeva
    Assessing Reading Development through Finger-Voice Span: a Comparative Study of Bulgarian and Italian
  • 13:50 – 14:10Ruslana Margova and Irina Temnikova
    Experiment Design of a Balanced Eye-Tracking Study for Measuring the Cognitive Impact of the Stimulus Word “Disinformation”
  • 14:10 – 14:30Tatiana Sherstinova, Irina Petrova, Aleksei Melnik, Sofiia Chepovetskaia, Karina Azarevich, Viktoria Deborina
    Everyday Speech of Russian Students: From Field Recordings to the ESC Speech Corpus (Data Processing and Representation)
  • 14:30 – 14:50Gunta Nešpore-Bērzkalne, Mikus Grasmanis and Ilze Lokmane
    Multi-word Expressions in the Electronic Dictionary Tēzaurs: Integrating Data from Different Sources
  • 14:50 – 15:10Dobromir Tsolyov, Keith Kiely, Georgi Goranov, Iglika Ivanova, Stanimir Dimitrov
    Optimizing LLM-based Discourse Analysis in Low-Resource Languages: A Case Study of Bulgarian Euro Adoption Narratives

15:10 – 15:30 | Coffee Break

Corpus-Based Approaches to Style, Register, and Meaning

Chair: Ventsislav Venkov

15:30 – 17:10 | Session Presentations

  • 15:30 – 15:50Iglika Nikolova-Stoupak, Gaël Lejeune, Eva Schaeffer-Lacroix
    Quantifying Literary Style: A Corpus-Based Study
  • 15:50 – 16:10Tomás Chismol Carbonell
    Zipfian Structure, Eurolect and Translationese: Exploratory Analysis of an EU Legal Corpus in English, French, Spanish and Bulgarian
  • 16:10 – 16:30Iryna Karamysheva, Jakob Horsch
    Imploring You to Investigate More Languages: A Corpus-based Contrastive Study of Ukrainian and English Causative Constructions
  • 16:30 – 16:50Ekaterina Tarpomanova
    Balkan Focus Particles in Bulgarian: bash and taman
  • 16:50 – 17:10Ivan Derzhanski, Olena Siruk
    When Bulgarian ‘Still’ Meets Ukrainian ‘Already’

17:10 | Conference closing and Award presentation

19:00 | Conference dinner at Hadzhidraganov’s Cellars (18 Hristo Belchev St)

PROGRAMME | 10 September 2026 | MAIN CONFERENCE

TUTORIALS | Floor 3, Hall 312

08:30 – 9:00 – Registration

09:00 – 12:30 | The Work of Linguists in Computational Linguistics

Lecturer: Irina Temnikova (Big Data for Smart Society Institute – GATE)

09:00 – 10:30 | Tutorial Session

10:30 – 11:00 | Coffee Break

11:00 – 12:30 | Tutorial Session

13:30 – 16:00 | Hack the Agent: Vulnerabilities and Security in AI Systems

Lecturer: Ivan Ivanov (Stihia.ai)

13:30 – 14:30 | Tutorial Session

14:30 – 15:00 | Coffee Break

15:00 – 16:00 | Tutorial Session