Skip to main navigation Skip to search Skip to main content

Do Language Models Understand Honorific Systems in Javanese?

Mohammad Rifqi Farhansyah, Iwan Darmawan, Adryan Kusumawardhana, Genta Indra Winata, Alham Fikri Aji, Derry Tanti Wijaya

Research output: Chapter in Book/Report/Conference proceedingConference PaperResearchpeer-review

Abstract

The Javanese language features a complex system of honorifics that vary according to the social status of the speaker, listener, and referent. Despite its cultural and linguistic significance, there has been limited progress in developing a comprehensive corpus to capture these variations for natural language processing (NLP) tasks. In this paper, we present UNGGAH-UNGGUH1, a carefully curated dataset designed to encapsulate the nuances of Unggah-Ungguh Basa, the Javanese speech etiquette framework that dictates the choice of words and phrases based on social hierarchy and context. Using UNGGAH-UNGGUH, we assess the ability of language models (LMs) to process various levels of Javanese honorifics through classification and machine translation tasks. To further evaluate cross-lingual LMs, we conduct machine translation experiments between Javanese (at specific honorific levels) and Indonesian. Additionally, we explore whether LMs can generate contextually appropriate Javanese honorifics in conversation tasks, where the honorific usage should align with the social role and contextual cues. Our findings indicate that current LMs struggle with most honorific levels, exhibiting a bias toward certain honorific tiers.

Original languageEnglish
Title of host publicationACL 2025ACL 2025ACL 2025, The 63rd Annual Meeting of the Association for Computational Linguistics (ACL 2025), Proceedings of the Conference - Volume 1: Long Papers July 27
EditorsWanxiang Che, Joyce Nabende, Ekaterina Shutova, Mohammad Taher Pilehvar
Place of PublicationKerrville TX USA
PublisherAssociation for Computational Linguistics (ACL)
Pages26732-26754
Number of pages23
ISBN (Electronic)9798891762510
DOIs
Publication statusPublished - 2025
EventAnnual Meeting of the Association for Computational Linguistics 2025 - Vienna, Austria
Duration: 27 Jul 20251 Aug 2025
Conference number: 63rd
https://2025.aclweb.org/ (Website)
https://aclanthology.org/events/acl-2025/#2025acl-long (Proceedings - Long)
https://aclanthology.org/events/acl-2025/#2025acl-short (Proceedings - Short)
https://aclanthology.org/events/acl-2025/#2025acl-demo (Proceedings - Demo)

Publication series

NameProceedings of the Annual Meeting of the Association for Computational Linguistics
PublisherAssociation for Computational Linguistics (ACL)
Volume1
ISSN (Print)0736-587X

Conference

ConferenceAnnual Meeting of the Association for Computational Linguistics 2025
Abbreviated titleACL 2025
Country/TerritoryAustria
CityVienna
Period27/07/251/08/25
Internet address

Cite this