AGT is best on all 13 KPIs against LightGBM, TimeMixer, and SOFTS in the matched seed-42 comparison, while final-architecture ablations show that relational attention, accounting topology, and the recency path each improve validation and test accuracy.
Abstract
Small businesses often have only 12-24 months of accounting history, yet planning and risk workflows require coordinated forecasts across financial statements. We study joint 12-month forecasting of 13 income-statement, balance-sheet, cash-flow, and working-capital key performance indicators (KPIs) from 71 monthly ledger series. We introduce the Accounting Graph Transformer (AGT), which represents each ledger series as a masked token, exchanges information through typed attention on a fixed accounting-relation graph, pools target-specific context, and fuses it with a gated three-month recency path. Across 11,993 forecast origins from 1,060 unseen companies, AGT achieves sample-weighted KPI-macro mean absolute error (MAE) $0.6990 \pm 0.0013$ over three independent seeds, compared with $0.7378 \pm 0.0014$ for the strongest baseline, LightGBM. At the pre-specified seed 42, a paired company-clustered bootstrap gives a LightGBM-minus-AGT difference of 0.0395 with 95% confidence interval (CI) $[0.0350,0.0439]$. AGT is best on all 13 KPIs against LightGBM, TimeMixer, and SOFTS in the matched seed-42 comparison, while final-architecture ablations show that relational attention, accounting topology, and the recency path each improve validation and test accuracy. On 7,094 additional unseen companies with origins sampled from January-May 2025, AGT obtains 0.7548 MAE versus 0.7694 for SOFTS. A single 5.3M-parameter model produces 156 aligned forecasts without company-specific fitting, providing one forecasting layer for integrated planning, liquidity, and working-capital analysis.
Modern semiconductor production relies on a globally distributed, multi-tier supply chain in which financial stress at one firm spreads with a delay and eventually affects the revenue, inventory, and profitability of the companies that design AI chips. Most firms see only their direct partners, and prior predictive res...
Specialist training beats generalist scale when forecasting financial statements. To our knowledge, no prior work jointly forecasts complete financial statements beyond one year, yet in a discounted-cash-flow valuation most firm value sits past that window. We release ProForma-20Q, a reproducible benchmark for forecast...
Travis L. Johnson, Jian Jiang, Soumyabrata Chaudhuri et al.· 0 citations
Covariate-Adjusted Residual Policy Learning (CAR-PL) is introduced, an action-wise R-learner that operates directly on multi-hot logs and regularizes selection by observational support and support objective-specific ranking of SMB financial guidance from multi-action accounting logs.
Shrutendra Harsola, Vignesh T. Subrahmaniam, Vikas Raturi et al.· 0 citations
This study evaluates whether large language models (LLMs) can convert the "Business and other risks" sections of Japanese annual securities reports into interpretable signals for long-horizon equity risk forecasting. Each extraction returns schema-constrained JSON, which is flattened into ordinal, binary, category-leve...
Nobushige Doi· IEEE Conference on Computati...· 0 citations
Japan introduced a sustainability-information section in annual securities reports for fiscal years ending on or after March 2023. We develop filing-order-constrained semantic exposure, measuring each filing’s position relative to earlier peer disclosures. Within fiscal year, industry, and disclosure layer, we retain u...
Eiji Sakihama· IEEE Conference on Computati...· 0 citations
This study develops a leakage-aware evaluation workflow for predicting shipment dispatch delays at the time of order, auditing when each predictor becomes available and quantifying how much apparent accuracy on the DataCo dataset derives from post-outcome information.
On 172,765 completed DataCo shipments, a...