model-splitting

Tag

Cards List
#model-splitting

TIL Why my dual 5060 Ti setup refuses to go past 50% usage and no, it's not broken.

Reddit r/LocalLLaMA · 2026-07-22

An investigation into why dual RTX 5060 Ti GPUs max out at ~50% utilization when running large LLMs like Qwen 27B reveals that memory bandwidth is the bottleneck and layer-by-layer splitting causes idle time, making it a relay race rather than parallel computation.

0 favorites 0 likes
#model-splitting

QSplitFL: Capability Aware Deep Q-Learning for Optimal Split Point Selection in Split Federated Learning

arXiv cs.LG · 2026-06-10 Cached

QSplitFL proposes a DQN-based framework for optimal split point selection in split federated learning, using client hardware metrics to adapt to heterogeneous devices. Experiments show improved convergence and accuracy across multiple datasets and architectures.

0 favorites 0 likes
← Back to home

Submit Feedback