Paper Brief
Mathieu Virbel
Add to My Podcasts
Episodes
About
Reviews
Promote
EP105 - WhisBERT: Multimodal Text-Audio Language Modeling on 100M Words
3 minutes
Posted Dec 6, 2023 at 3:22 am.
0:00
3:07
Add to My Queue
Download
MP3
Share
episode
Share at current time
Show notes
Previous
EP104 - LLaVA-Grounding: Grounded Visual Chat with Large Multimodal Models
Next
EP106 - Alchemist: Parametric Control of Material Properties with Diffusion Models
← See all 155 episodes of Paper Brief