Atlas node / artificial-intelligence

Multimodal Learning & VLMs

Aligning visual and language representations so models can ground symbols in scenes.

CHILD CONCEPTS

What belongs here.

  • Vision-Language Models
  • Contrastive Learning
  • Cross-Modal Alignment
PREREQUISITES

Planned

RELATED CONCEPTS

Planned

LEADS TO
BLOG

Planned

PAPERS

Planned

PROJECTS

Planned

← Back to Atlas