pipelining

Tag

Cards List
#pipelining

Flattening Every Memory Peak in Long-Context Mixture-of-Experts Training

Hugging Face Daily Papers · 4d ago Cached

This paper introduces techniques to manage memory peaks in training large Mixture-of-Experts models with long context lengths, including Pipelined LLEP, Ring-DTP, SCO, and OffloadStreamAdamW, which enable fixed GPU working sets and improve throughput up to 10.4x.

0 favorites 0 likes
#pipelining

Streaming Communication in Multi-Agent Reasoning

Hugging Face Daily Papers · 2026-06-03 Cached

StreamMA introduces a streaming communication paradigm for multi-agent reasoning that pipelines intermediate results to reduce latency and improve effectiveness by leveraging more reliable early steps, outperforming baselines across benchmarks and revealing a step-level scaling law.

0 favorites 0 likes
← Back to home

Submit Feedback