LLaVA-OneVision-1.5-Mid-Training-85M mvp-lab

🚀 LLaVA-One-Vision-1.5-Mid-Training-85M Dataset is being uploaded 🚀 Upload Status All Completed: ImageNet-21k、LAIONCN、DataComp-1B、Zero250M、COYO700M、SA-1B、MINT、Obelics 📜 Cite If you find LLaVA-One-Vision-1.5-Mid-Training-85M useful in your research, please consider to cite the following related papers: @misc{an2025llavaonevision15fullyopenframework, title={LLaVA-OneVision-1.5: Fully Open Framework for Democratized Multimodal Training}… See the full description on the dataset page: https://huggingface.co/datasets/mvp-lab/LLaVA-OneVision-1.5-Mid-Training-85M.

種別
dataset
ライセンス
apache-2.0
ダウンロード
733,882
いいね
89
アクセス
public
ファイル
0

タグ

  • マルチモーダル
  • 画像データセット
  • 事前学習コーパス
  • 多言語コーパス
  • 合成データ
  • コンピュータビジョン
  • 視覚モデル
  • NLP

概要

LLaVA-OneVision-1.5モデルの中間学習段階向けに構築された約85Mサンプルの大規模マルチモーダルデータセットです。ImageNet-21k、LAION、DataComp-1B、COYO、SA-1B、Obelicsなど多様な画像・テキスト・画像セグメンテーションソースを統合しています。オープンなマルチモーダル学習の民主化を目的とし、視覚と言語の統合モデルの事前学習・中間学習に活用できます。ライセンスはApache-2.0です。

README

--- license: apache-2.0 --- # 🚀 LLaVA-One-Vision-1.5-Mid-Training-85M データセットをアップロード中です 🚀 # アップロード状況 - **すべて完了**: ImageNet-21k、LAIONCN、DataComp-1B、Zero250M、COYO700M、SA-1B、MINT、Obelics # 📜 引用 研究において *LLaVA-One-Vision-1.5-Mid-Training-85M* が役立つ場合は、以下の関連論文を引用することをご検討ください: ``` @misc{an2025llavaonevis…

查看完整页面 · 查看原文