WeChat has launched WeMM-Embedding, a family of models that maps text, images, videos, and visual documents into a single unified representation space.
HOW THIS AFFECTS YOU
●
builderYou can use these embeddings to perform cross-modal retrieval across different media types.
●
researcherThis advances unified representation learning across heterogeneous data types.