Skip to content
Open access

Enhancing Consistency in Academic English Writing Feedback Generation with a MacBERT-large Model Combining Adversarial Training and Contrastive Learning

Aug 2026 · Advanced Electromagnetics · Vol 15, pp. 8858-8862 · 0 citations · 10 references

TL;DR

An enhanced MacBERT-large encoder–decoder model integrating Fast Gradient Method adversarial training and supervised contrastive learning is proposed, providing a semantic consistency modeling framework for intelligent text generation and academic writing assistance systems.

Abstract

Academic English writing feedback generation requires robust semantic understanding, stable feedback output, and accurate discrimination among similar error types. Existing feedback generation systems often produce inconsistent suggestions for semantically equivalent inputs and show limited generalization to complex academic expressions. To improve feedback consistency, this study proposes an enhanced MacBERT-large encoder–decoder model integrating Fast Gradient Method adversarial training and supervised contrastive learning. The MacBERT-large encoder extracts contextual semantic representations of academic text, while a Transformer decoder generates feedback sequences using a dedicated academic vocabulary. FGM adversarial training introduces controlled perturbations into the embedding layer, enabling the model to maintain stable predictions under paraphrased or slightly varied inputs. A supervised contrastive learning module maps text samples into a representation space where feedback cases with the same error type are pulled closer and different error types are separated through NT-Xent loss. A multi-task learning framework jointly optimizes cross-entropy loss, adversarial loss, and contrastive loss to balance generation quality, robustness, and category discrimination. Experiments on the AEW-Feedback dataset containing 15,000 academic papers show that the proposed model achieves 67.3% BLEU-4, 71.2% ROUGE-L, 74.8% METEOR, and a feedback consistency score of 0.891, outperforming MacBERT-large and single-enhancement variants. The method provides a semantic consistency modeling framework for intelligent text generation and academic writing assistance systems.

Read PDF

Similar papers

Open access 2026

Improving Machine Translation Using an Efficient Dual-Bert Adversarial Network (DBAN) Model for User-Generated Content

An Efficient Dual-BERT Adversarial Network (DBAN) is proposed to improve the translation of noisy UGC by integrating contextual representation learning with adversarial training and significantly improves contextual understanding and cross-lingual semantic alignment while maintaining computational efficiency.

A. A. Aliero, Nasiru Muhammad Dankolo · 0 citations
#natural language process... Preprint Sep 2026

Generating Adversarial Texts for Machine Translation via GRPO

As machine translation (MT) systems continue to improve, standard benchmarks become less informative for exposing remaining weaknesses. Traditional methods for creating challenging test sets rely on expensive manual creation or curation, while automated approaches struggle to produce sets with the necessary translation...

Florian Zogaj, Jakob Hütteneder, Giovanni De Muri et al. · 1 citation
Aug 2026

Augmenting Text to Increase Translation Difficulty

This work proposes augmenting existing benchmarks to increase translation difficulty by combining adversarial optimization with a differentiable translation difficulty estimator, and uses gradients from a combined difficulty and fluency objective to iteratively replace tokens in Adversarial Translation Optimization (AT...

William Kalikman, Šimon Sukup, Michal Tesnar et al. · 2 citations
Open access Aug 2026

Automatic Assessment Model for Chinese-English Scientific Translation Quality Based on Contrastive Learning

A Contrastive Learning-based Chinese-English Scientific Translation Quality Evaluation model (C-TQE), which provides an effective solution for large-scale scientific translation quality assessment and facilitates the accurate international communication of multidisciplinary engineering research, including electromagnet...

Rui Hu · 0 citations
Open access Jul 2026

Semi-Supervised Marginal Likelihood Training with Curriculum-Guided Rewriting for Low-Resource Machine Translation

Large language models continue to face challenges in translating low-resource languages with scarce parallel data. This study investigates how to fine-tune them effectively using target-side monolingual data. Existing approaches—dominated by back-translation and recent LLM-based rewriting—remain limited by noisy synthe...

Wenjie Yu, Zhiqiang Yu, Zuo Jiang et al. · 0 citations
Preprint Aug 2026

Efficient Multilingual Neural Machine Translation via Corpus-Driven Vocabulary Pruning: An English-Arabic Case Study

This paper proposes a general optimization framework that combines a vocabulary pruning method with a targeted fine-tuning protocol for MNMT models, and reduces the vocabulary size from over 128,000 to approximately 10,000 tokens, enabling a 60% memory saving without any loss in performance.

A. A. Aliane, N. Semmar, H. Aliane · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.