RECAST is proposed, a lightweight sampling module for buffered TTA frameworks that improves estimation accuracy and trend tracking over baselines and ablations, and stays practical, adding only sub-second latency per segment on a single GPU and CPU core.
Abstract
Streaming biosignals vary across subjects and drift over time, so population-trained models lose accuracy during long-term monitoring. Test-time adaptation (TTA) enables online personalization by updating the model on incoming samples. But in a stream, a basic question is left open: \emph{which samples should drive each update?} Using all buffered samples blurs the update with irrelevant segments. Using only the latest segment makes the update noisy and unstable. The most useful samples are recent, aligned with the current physiological state, and reliable enough to learn from. We propose \textbf{RECAST} (REcent \&Context-Aware Sampling for TTA), a lightweight sampling module for buffered TTA frameworks. RECAST builds each adaptation batch from three signals: temporal recency, contextual similarity, and predictive reliability. It changes only which samples are used, leaving the model and the training objective unchanged. On two blood-pressure datasets, RECAST improves estimation accuracy and trend tracking over baselines and ablations. The per-patient gains are statistically significant on both datasets, with broad improvement on the regular benchmark and gains concentrated on the hardest patients in the emergency-department setting. RECAST stays practical, adding only sub-second latency per segment on a single GPU and CPU core.
Pre-trained visual models are increasingly deployed in streaming environments where the target distribution changes with weather, illumination, sensor aging, background composition, and time. Existing test-time adaptation (TTA) algorithms usually assume either a fixed target domain or sufficiently large independent tes...
Wen-Juan Guo, Wen-Tao Zhang, Shu-Yuan Wang et al.· 2026 6th International Confe...· 0 citations
LongVU-TTT is introduced, which inserts a convolutional Test-Time Training (TTT) resampler with causal fast-weight updates between the vision encoder and the LLM, and is stronger than attention- and fixed-state recurrent resamplers across three benchmarks.
Mahmoud Ahmed, Sameh Abdulah, Olatunji Ruwase et al.· 0 citations
Results demonstrate that STRAP delivers both high predictive accuracy and ultra-low-latency responsiveness in highly dynamic, high-concurrency streaming environments, exhibiting strong practical deployability for latency-sensitive real-time forecasting scenarios such as smart grid anomaly detection, financial risk cont...
Fangrui Yu, Xiangyu Lin· International Conference on...· 0 citations
This work proposes an inference-only streaming autoregressive framework that replaces repeated full-context recomputation with a one-time context warmup and incremental decoding, enabling efficient long-history forecasting without retraining.
Xi-Yu Meng, Yuhan Wu, Can-Ran Xiao et al.· Proceedings of the Thirty-Fi...· 0 citations
This work introduces StreamTTT, which writes long-range history into online-updated fast weights outside the attention context, leaving a short sliding key-value cache dedicated to recent evidence, mitigating attention dilution.
Autoregressive video diffusion enables scalable long-video generation by producing chunks from a bounded recent context. While recency-based caching preserves local continuity, it evicts historical cues needed when subjects, objects, scenes, or attributes reappear. Existing memory mechanisms expose models to nonlocal h...
With $2.1 million funding from Google.org, the open-source Public Transit Intelligence Hub will unify public transit monitoring, operations, and passenger communication.
Professor Sherry Turkle’s new book, “Artificial Intimacy,” offers a withering critique of chatbots and the antisocial dynamics she believes they encourage.
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.