downstream-tasks

Tag

Cards List
#downstream-tasks

@iScienceLuvr: Music-JEPA: Learning a World Model of Sound from Action "we propose to learn a world model of piano sound using JEPA by…

X AI KOLs Following · 3d ago Cached

This paper proposes Music-JEPA, a world model that learns piano sound representations by framing audio as a state and piano roll as an action. It captures action-sound relationships and enables downstream tasks like beat tracking and piano transcription via planning.

0 favorites 0 likes
← Back to home

Submit Feedback