Tag
The article discusses a talk by Jing Huang from Stanford NLP at the Summer of Data event, explaining that larger AI models retain rare skills better due to task interference and greater capacity.