ML Academy · Track 3 · Large Language Models

Multimodal models: bringing image and text into one space

CLIP's core is genuinely trained here. Two of the three results do not come out as expected.

4 steps 225 XP A free account is needed
Start the lesson →

Sources

ML Academy · an interactive machine learning course that runs in your browser · All lessons