Learning generative models that span multiple data modalities, such as vision and language, is often motivated by the desire to learn more useful, generalisable representations that faithfully capture common underlying factors between the modalities. In this work, we characterise successful learning...
Research Assistant
AI chat, annotations, notes & similar papers
No comments yet
Be the first to share your thoughts!