Skip to content

Author

Maggie Haitian Wang

We have 2 of 5 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Flux Attention: Context-Aware Hybrid Attention for Efficient LLMs Inference

Flux Attention is introduced, a context-aware framework that dynamically optimizes attention computation at the layer level by integrating a lightweight Layer Router into frozen pretrained LLMs, which adaptively routes each layer to FA or SA based on the input context.

Quantong Qiu, Zhiyi Hong, Yi Yang et al. · 0 citations

MMLongEmbed: Benchmarking Multimodal Embedding Models in Long-Context Scenarios

This work introduces MMLongEmbed, the first comprehensive benchmark for evaluating MEMs in long-context scenarios, and finds that current architectures rely heavily on superficial feature matching and struggle to capture deep semantic and structural dependencies.

Maggie Haitian Wang, Ruoxi Sun, Quantong Qiu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.