@tarat_211: I made a video about my first real attempt at ML research. I started with a simple question about pruning vision-langua…

X AI KOLs Timeline Papers

Summary

A researcher made a video about their first ML research attempt on pruning vision-language models, including negative results, and submitted a paper to arXiv.

I made a video about my first real attempt at ML research. I started with a simple question about pruning vision-language models, spent weeks breaking Qwen2.5-VL and SmolVLM2 in different ways, hit a bunch of negative results, and eventually submitted my first paper to arXiv. https://t.co/berWTRdcvH
Original Article
View Cached Full Text

Cached at: 07/21/26, 01:36 AM

I made a video about my first real attempt at ML research.

I started with a simple question about pruning vision-language models, spent weeks breaking Qwen2.5-VL and SmolVLM2 in different ways, hit a bunch of negative results, and eventually submitted my first paper to arXiv. https://t.co/berWTRdcvH

Similar Articles

@harshbhatt7585: https://x.com/harshbhatt7585/status/2063593933314113587

X AI KOLs Timeline

The author shares learnings from training a 160M parameter LLM from scratch, experimenting with architectures like multi-token prediction and hierarchical reasoning models. They emphasize the importance of fast iteration, simplifying ideas, and understanding why architectures work.

@swyx: full writeup and links here

X AI KOLs Timeline

A Latent Space podcast episode discusses the thesis that video models derive intelligence from LLMs, and that the next frontier is video agents. Guest Ethan He, who built Grok Imagine at xAI, shares insights on building frontier image and video systems.