Skip to content

Global Merger-Arbitrage Forecasting with Language Models

Jul 2026 · arXiv.org · Vol abs/2607.09921 · 0 citations · 22 references
Computer Science

TL;DR

The results, together with ablation studies, show that LLM-based forecasting can succeed in specialized, long-context financial workflows, with hindsight-based supervision and expert-designed context playing a critical role.

Abstract

We present a language-model forecasting system for merger arbitrage, a specialized high-stakes financial setting in which the task is to predict the outcome of announced M\&A deals. Unlike prior work on judgmental forecasting with LLMs, which has focused on broad mixed-topic benchmarks and short context such as news snippets, we study a setting that requires long-context reasoning over hundreds of pages of technical documents. Our system combines expert-guided context engineering with finetuning on hindsight-guided reasoning traces derived from historical deals. Given an announced deal, it outputs a probability distribution over three mutually exclusive outcomes: closing at announced terms, a higher bid, or deal termination. On an out-of-sample set of more than 400 large deals spanning 42 countries, our finetuned system achieves the best performance of any method we evaluate, reducing class-balanced Brier score to 0.151. This is 24\% below calibrated market-implied probabilities, 19\% below XGBoost, and 25-42\% below frontier language models. These results, together with ablation studies, show that LLM-based forecasting can succeed in specialized, long-context financial workflows, with hindsight-based supervision and expert-designed context playing a critical role.

View source

Similar papers

Preprint Aug 2026

Long-Horizon Forecasting of Complete Financial Statements with Forma

Specialist training beats generalist scale when forecasting financial statements. To our knowledge, no prior work jointly forecasts complete financial statements beyond one year, yet in a discounted-cash-flow valuation most firm value sits past that window. We release ProForma-20Q, a reproducible benchmark for forecast...

Travis L. Johnson, Jian Jiang, Soumyabrata Chaudhuri et al. · 0 citations
Aug 2026

A large language model framework for predicting significant equity index fluctuations

It is concluded that the genuinely difficult part of deployment is tail-event detection rather than directional forecasting, and that a credible evaluation must rely on event-sensitive metrics and time-aware validation; without them, performance under heavy imbalance is easily overstated.

Rishov Nag, Sankhayan Choudhury · 0 citations
#natural language process... Preprint Aug 2026

Large Language Model-Driven Small-Capitalization Trading: Integrating Financial News Sentiment, Macroeconomic Indicators, and Technical Signals

This work examines the pipeline on Russell 2000 equities under three stock-selection regimes and suggests that stock-selection regime and allocator choice matter at least as much as the sentiment model, and that separating firm-specific and macro-exposure triggers is more informative than requiring both to fire simulta...

Alireza Kargarzadeh, Nariman Khaledian, Navid Parvini et al. · 0 citations
Preprint Aug 2026

Financial Numerical Prediction and Allocation as Token Generation

The results support the feasibility of direct language-model token generation for financial numerical prediction and decision-making, while motivating broader tests across assets, regimes, and random seeds.

Xu Ouyang, Moontae Lee · 0 citations
Case report Aug 2026

An Open Benchmark for Evaluating Time Series Forecasting Methods Across Financial Markets

Holding the data fixed and evaluating roughly a dozen univariate methods without exogenous regressors, it is found that asset returns remain near unforecastable across every method family and that hybrid and machine learning methods exhibit additional forecasting power on basis spreads and bank indicators.

J. Bejarano, Viren Desai, K. Keshava et al. · 0 citations
#machine learning Preprint Aug 2026

Temporal Leakage in Financial News NLP: A Multi-Architecture Audit with a Regime-Specific M&A Signal

This work audits financial-news direction prediction dependence on a 49,799-article corpus across 16 feature-model combinations spanning TF-IDF, MiniLM, FinBERT, and fine-tuned RoBERTa-large / DeBERTa-v3-large, plus separate zero/few-shot and LoRA probes of Llama-3 and Qwen2.

Chen-Hao Xue, Raslen Guesmi, Si-Wei Feng et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.