huggingface_hub v1.0: Five Years of Building the Foundation of Open Machine Learning
More from the blog
TimesFM-3: A zero-shot foundation model for multivariate forecasting
Data Management
GigaPath-Flash and GigaTIME-Flash: Toward population-scale discovery with efficient pathology foundation models
What if pathology foundation models could do more with less? GigaPath-Flash and GigaTIME-Flash cut computational demands while maintaining strong performance, opening the door to larger studies and broader exploration. The post GigaPath-Flash and GigaTIME-Flash: Toward population-scale discovery with efficient pathology foundation models appeared first on Microsoft Research.
The Open ASR Leaderboard Adds Its First Global South Language
Looking beyond natural sequences
A new machine-learning framework aims to improve the success rate of computational protein design while moving away from results that reproduce sequences found in nature.
Related papers
Intelligent agents: architectures, learning paradigms, deployment models, and applications
Anomaly detection method for satellite networks based on adaptive federated learning driven by deep reinforcement learning
Asynchronous Federated Reinforcement Learning for Adaptive Resource Slicing and Low-Latency Task Offloading in Heterogeneous 6G Edge Computing Networks
The emerging paradigm of 6G wireless communication networks envisions ultra-reliable low-latency communication (URLLC), massive machine-type communications (mMTC), and pervasive edge computing intelligence. In heterogeneous mobile edge computing (MEC) networks, dynamically offloading compute-intensive tasks (e.g., augmented reality rendering, connected vehicular telemetry, autonomous robotic control) while orchestrating multi-tenant network slicing under time-varying channel conditions is an NP-hard stochastic optimization problem. Centralized reinforcement learning algorithms suffer from extreme communication overhead, severe backhaul congestion, and severe privacy vulnerabilities. Conversely, standard synchronous Federated Learning (FL) methods encounter severe 'straggler effects' caused by heterogeneous edge device processing capabilities. In this paper, we propose AF-EdgeRL, a novel Byzantine-resilient Asynchronous Federated Reinforcement Learning framework tailored for distributed resource allocation and dynamic task offloading. AF-EdgeRL deploys a distributed Proximal Policy Optimization (PPO) agent across edge servers and end-user devices, combined with a Staleness-Aware Adaptive Weight Aggregator (SAWA) that dynamically adjusts model update gradients based on hardware compute latency and channel state information (CSI). Furthermore, we establish theoretical convergence guarantees under non-convex reinforcement learning objectives. Evaluated on a high-fidelity 6G MEC simulator with real-world mobile mobility traces (Telecom Italia Milano dataset), AF-EdgeRL reduces end-to-end task execution latency by 41.2%, achieves 99.999% URLLC deadline compliance, and decreases edge energy consumption by 32.6% compared to state-of-the-art synchronous FedRL and centralized DRL baselines.