Tutorials
[Olewave's Review] CLIP (1/3): Learning Transferable Visual Models From Natural Language Supervision
Before there was BLIP, LLaVA, or any speech-LLM worth its salt, there was CLIP, and this is where the story begins.
Tutorials
[Olewave's Review] Token-level Sequence Labeling for SLU using Compositional E2E Models
End-to-end SLU is having a moment, but treating sequence labeling as sequence prediction quietly throws away decades of well-understood token-level tagging mac…
