Skip to content

Can Large Language Models Handle Discourse Particles? A Case Study of Colloquial Malay

May 2026 · arXiv.org · Vol abs/2605.28782 · 0 citations · 57 references
Computer Science

TL;DR

MalayPrag is proposed, a benchmark designed to systematically evaluate and analyze LLMs' capabilities in handling discourse particles in colloquial Malay and five attributes are introduced that provide a theoretically grounded, unified framework for interpreting pragmatic functions of discourse particles.

Abstract

Discourse particles, such as well and kind of, are crucial components that enable LLMs to"speak"more like humans. They are used to convey emotions, intentions, and interpersonal attitudes. However, existing studies have not yet built a comprehensive understanding of LLMs'capabilities in handling discourse particles. Moreover, the limited number of research focuses primarily on high-resource languages such as English, with little attention paid to Southeast Asian languages. In this paper, we (1) propose MalayPrag, a benchmark designed to systematically evaluate and analyze LLMs'capabilities in handling discourse particles in colloquial Malay; (2) introduce five attributes that provide a theoretically grounded, unified framework for interpreting pragmatic functions of discourse particles. Applying these two, we prompt ten off-the-shelf LLMs to perform three prediction tasks. The experimental results reveal substantial challenges for current LLMs to accurately connect discourse particles and their pragmatic functions in Malay. The provision of the five attributes designed in this study is found to significantly improve the connections, highlighting the need for structured scaffolding for models'pragmatic competence.

View source

Similar papers

2026

Conversational Implicatures through the Lens of LLMs

It is proposed that LLMs can serve not only as benchmarks for human-model alignment, but also as tools for investigating the nature of pragmatic phenomena and their relationship to linguistic theory.

A. Lombardi, Alessandro Lenci · 0 citations
Open access Aug 2026

A Descriptive Analysis of the Semantic and Pragmatic Functions of Chotto in Selected Japanese Literary and Media Materials

The Japanese adverb chotto is commonly introduced as a scalar expression meaning ‘a little’, ‘slightly’, or ‘for a short while’. In actual use, however, it also performs a range of pragmatic functions. Speakers may use it to attract attention, maintain a conversational turn, reduce the directness of a request, signal h...

N. Anh, N. Anh · 0 citations
Open access Sep 2026

Discourse pragmatic functions of the particles no and nu in Iźva Komi

This research focuses on the pragmatic functions of the discourse particles no and nu in the Iźva dialect of the Komi language. The study is based on contemporary spoken-language material. No and nu appear to carry the same functions, with the exception of the confirming stand-alone use, which is restricted to no. The...

Eda-Riin Tuuling · 0 citations
Open access Sep 2025

Benchmark of stylistic variation in LLM-generated texts

This study investigates the register variation in texts written by humans and comparable texts produced by large language models (LLMs). Multidimensional analysis (MDA) is applied to a sample of human-written texts and AI-generated counterpart texts to find the dimensions of variation in which LLMs differ most si...

Jiří Milička, Anna Marklová, Václav Cvrček · 10 citations
2026

Beyond Literal Meaning: How LLMs Interpret Yemeni Proverbs

Results show that instruction-tuned models like GPT-4o and Gemini 1.5 Pro outperform smaller models in both automatic and human evaluations, and LLM-as-a-Judge evaluation correlates strongly with human assessment.

Nasser Thmer, Ali Allaith, Muhammad Shoaib · 0 citations
Open access Aug 2026

The role of lexical and grammatical means of modality in realising intentionality of media discourse

This study investigates the contribution of lexical and grammatical modality devices towards conveying intentionality in media discourse through a comparative analysis of feature journalism articles from The Times of India (TOI) and The Los Angeles Times (The LA Times). Mixed-methods corpus-assisted discourse analysis...

S. Rajeswari · 0 citations

Related blog posts

MIT News · Artificial Intelligence Sep 24, 2026

Estimating suicide risk from text

A new language-processing tool could help identify the highest-risk individuals from natural language, enabling swifter interventions.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.