Glossary · Multimodal systems

Shared Embedding Space

A common vector space in which representations from different modalities can be compared with the same similarity function.

Why it matters

It enables cross-modal retrieval and matching, such as finding images from text, without requiring both items to share a raw representation.

In practice

Train paired and unpaired negatives deliberately, normalize vectors when the objective requires it, evaluate both retrieval directions, and inspect subgroup and language performance.

Common confusion

Sharing a vector dimension does not create a shared semantic space. The training objective and data must establish cross-modal comparability.

Related terms

Sources

Browse the learning paths to see this term in context — every lesson is free to read.