Nomic Embed Vision:扩展潜在空间
计算机视觉与模式识别
2024-06-28 v1 人工智能
摘要
本技术报告描述了 nomic-embed-vision 模型的训练,该模型是高性能的、开源代码和开源权重的图像嵌入模型,与 nomic-embed-text 共享相同的潜在空间。nomic-embed-vision 与 nomic-embed-text 共同构成了首个在视觉、语言和多模态任务上实现高性能的统一潜在空间。
引用
@article{arxiv.2406.18587,
title = {Nomic Embed Vision: Expanding the Latent Space},
author = {Zach Nussbaum and Brandon Duderstadt and Andriy Mulyar},
journal= {arXiv preprint arXiv:2406.18587},
year = {2024}
}