Glossary · Multimodal systems
Shared Embedding Space
A common vector space in which representations from different modalities can be compared with the same similarity function.
Why it matters
It enables cross-modal retrieval and matching, such as finding images from text, without requiring both items to share a raw representation.
In practice
Train paired and unpaired negatives deliberately, normalize vectors when the objective requires it, evaluate both retrieval directions, and inspect subgroup and language performance.
Common confusion
Sharing a vector dimension does not create a shared semantic space. The training objective and data must establish cross-modal comparability.
Related terms
Sources
Browse the learning paths to see this term in context — every lesson is free to read.