How I Study AI - Learn AI Papers & Lectures the Easy Way

Compositional Generalization Requires Linear, Orthogonal Representations in Vision Embedding Models

Intermediate

Arnas Uselis, Andrea Dittadi et al.Feb 27arXiv

The paper asks a simple question: what must a vision model’s internal pictures (embeddings) look like if it can recognize new mixes of things it already knows?

#compositional generalization#linear representation hypothesis#orthogonal representations

Papers1

Compositional Generalization Requires Linear, Orthogonal Representations in Vision Embedding Models