Skip to content

Category

small language model

425 papers

#small language model Preprint Aug 2026

'Ghaib in Translation'aka Unseen Harm: Measuring Cross-Script Safety Inconsistency with'Missed-in-Urdu'Scores in LLM Hate Speech Detection

Results indicate that current LLMs provide uneven safety assurance across Urdu's script varieties, with smaller open-weight models showing substantially higher instability and missed-harm rates than frontier closed models.

F. Kara-Isitt, Sonal Khosla, S. Swift · 0 citations
#small language model Preprint Aug 2026

Constraint-Guided Enterprise Data Mapping with Large Language Models

Constrained-guided mapping is proposed, a neuro-symbolic method with three stages: schema-grounded admissibility constraints with metadata mc =, where tau_c denotes the constraint type and delta_c provides executable relation and normalization logic, and constraint-restricted candidate generation with cascade relaxation to guarantee a nonempty feasible set under noise.

Sebastian Monka, Pramod Anantharam, Thị Minh et al. · 0 citations
#small language model Review Open access Sep 2026

Automating cost-effectiveness models with agentic artificial intelligence: Case study and implications for value assessment.

Findings support a hybrid paradigm in which AI augments, but does not replace, health economists in value assessment and formulary decision support within managed care settings.

R. Mudumba, A. Modi, Kevin Mayo · 0 citations

Stacking Approaches for Multi-Label Emotion Recognition in Persian Using Large and Small Transformer Models

A hybrid model, termed SE_LLM_ST, is introduced, which leverages recent advancements in large language models (LLMs) and transformer-based architectures to effectively capture both contextual and sequential information vital for Persian emotion recognition.

Toktam Khatibi, Elham Farahani · 0 citations

MEPO-SLM: multi-objective evolutionary prompt optimization for energy-efficient small language models on edge devices

MEPO-SLM is presented, a framework that reformulates prompt engineering for SLMs as a four-objective Pareto problem over task inaccuracy, and Phi-3-mini and Gemma-2B on English TriviaQA and Arabic medical QA, and TinyLlama-1.1B on TriviaQA only are evaluated.

Yousef K. Sanjalawe, Salam R. Al-E’mari, S. Makhadmeh · 0 citations
#small language model Preprint Aug 2026

From Authorial Mathematics to Studio Mathematics:Ecobiontic Forms of Proof after Large Language Models

Mathematics has often been organized around an authorial subject: one person, or a small group, composing proofs through language, notation, and judgment. Large language models, proof assistants, formal libraries, and repositories now make another production unit technically credible: a human-machine assemblage. This article calls that unit a studio ecobiont and asks when it is epistemically legitimate. Its governance thesis is that human participation is substantive only when the system preserves traceable provenance, reconstructible human competence, capacity to challenge the result, effective authority to stop or withdraw it, and public responsibility. These conditions distinguish a governed studio from a degenerate studio whose human oversight is ceremonial. A comparison of Polymath, the Liquid Tensor Experiment, Danus, and the Jacobian counterexample episode shows that collaboration, formalization, technical orchestration, and epistemic governance are independent dimensions. The proposed understanding audit and contribution-authority trace are governance designs, not validated measures. No causal superiority over authorial practice is claimed.

Oliver López Corona · 0 citations
#small language model Preprint Aug 2026

SchemaGUI: A Schema-Driven Benchmark for Controllable GUI Generation Evaluation

SchemaGUI, a template-based benchmark for controllable GUI generation evaluation, synthesizing paired natural language instructions and deterministic function-call references from parameterized interface schemas can generate thousands of deterministically annotated tasks in seconds without human labeling.

Jiarui Dong, Yin Cai, Zhouhong Gu et al. · 0 citations
#small language model Preprint Aug 2026

Small Reasoning Models are Instruction Followers in Function Calling

This work introduces Instruction-Followed Function Calling (IFFC), a novel framework that decouples function-calling logic from the primary LLM and delegates it to a dedicated smaller model operating within the instruction-following paradigm, establishing a new paradigm for reliable, resource-efficient function calling in edge-computing scenarios.

Yalda Taheri, Mohammad Hassan Heydari, Erfan Naaman et al. · 0 citations
#natural language process... Preprint Aug 2026

Dual-Layer Agentic Memory with Fast Write Routing and Slow Consolidation

Dual-Layer Agentic Memory is proposed, a framework that shifts memory management to the write phase through cost-aware epistemic routing and periodic parametric consolidation, allowing the router to adaptively suppress redundant writes as the model's epistemic boundaries evolve.

Wenzhi Li, Dong Nie, Ruiyi Lan et al. · 0 citations
#small language model Preprint Aug 2026

Evaluation of Small Vision-Language Models on Qualitative Mechanical Problems

Two state-of-the-art multimodal models, Gemma-3 and Qwen-VL, are assessed on their ability to interpret mechanical problem images by eliciting a step-by-step chain of thought (CoT) and a final answer, and final answers are compared to verified solutions to measure accuracy.

Henry Fordjour Ansah, Shreya Banerjee, Pranish Ghimire · 0 citations

From tech blogs

See all →
Microsoft Research Blog Aug 31, 2026

GigaPath-Flash and GigaTIME-Flash: Toward population-scale discovery with efficient pathology foundation models

What if pathology foundation models could do more with less? GigaPath-Flash and GigaTIME-Flash cut computational demands while maintaining strong performance, opening the door to larger studies and broader exploration. The post GigaPath-Flash and GigaTIME-Flash: Toward population-scale discovery with efficient pathology foundation models appeared first on Microsoft Research.