LLaVA-OneVision-1.5-Mid-Training-85M mvp-lab

🚀 LLaVA-One-Vision-1.5-Mid-Training-85M Dataset is being uploaded 🚀 Upload Status All Completed: ImageNet-21k、LAIONCN、DataComp-1B、Zero250M、COYO700M、SA-1B、MINT、Obelics 📜 Cite If you find LLaVA-One-Vision-1.5-Mid-Training-85M useful in your research, please consider to cite the following related papers: @misc{an2025llavaonevision15fullyopenframework, title={LLaVA-OneVision-1.5: Fully Open Framework for Democratized Multimodal Training}… See the full description on the dataset page: https://huggingface.co/datasets/mvp-lab/LLaVA-OneVision-1.5-Mid-Training-85M.

类型
dataset
许可
apache-2.0
下载量
733,882
点赞
89
访问
public
文件
0

标签

  • 多模态数据集
  • 预训练语料
  • 图像数据集
  • 视觉语言模型
  • 多模态
  • 深度学习

摘要

该数据集是LLaVA-OneVision-1.5模型的中间训练数据集,包含约8500万样本,整合了ImageNet-21k、LAIONCN、DataComp-1B、COYO700M、SA-1B等多样化的图文数据源。主要用于大规模多模态模型的预训练与中间训练阶段,覆盖图像理解、视觉语言对齐等多种任务。适用于需要海量图文数据来训练或改进视觉语言模型的研究与工程场景。

README

--- license: apache-2.0 --- # 🚀 LLaVA-One-Vision-1.5-Mid-Training-85M 数据集正在上传中 🚀 # 上传状态 - **全部已完成**:ImageNet-21k、LAIONCN、DataComp-1B、Zero250M、COYO700M、SA-1B、MINT、Obelics # 📜 引用 如果您的研究中使用了 *LLaVA-One-Vision-1.5-Mid-Training-85M*,请考虑引用以下相关论文: ``` @misc{an2025llavaonevision15fullyopenframework, tit…

查看完整页面 · 查看原文