LLaVA-OneVision-1.5-Mid-Training-85M mvp-lab
🚀 LLaVA-One-Vision-1.5-Mid-Training-85M Dataset is being uploaded 🚀 Upload Status All Completed: ImageNet-21k、LAIONCN、DataComp-1B、Zero250M、COYO700M、SA-1B、MINT、Obelics 📜 Cite If you find LLaVA-One-Vision-1.5-Mid-Training-85M useful in your research, please consider to cite the following related papers: @misc{an2025llavaonevision15fullyopenframework, title={LLaVA-OneVision-1.5: Fully Open Framework for Democratized Multimodal Training}… See the full description on the dataset page: https://huggingface.co/datasets/mvp-lab/LLaVA-OneVision-1.5-Mid-Training-85M.
- 種別
- dataset
- ライセンス
- apache-2.0
- ダウンロード
- 733,882
- いいね
- 89
- アクセス
- public
- ファイル
- 0
タグ
- マルチモーダル
- 画像データセット
- 事前学習コーパス
- 多言語コーパス
- 合成データ
- コンピュータビジョン
- 視覚モデル
- NLP
概要
LLaVA-OneVision-1.5モデルの中間学習段階向けに構築された約85Mサンプルの大規模マルチモーダルデータセットです。ImageNet-21k、LAION、DataComp-1B、COYO、SA-1B、Obelicsなど多様な画像・テキスト・画像セグメンテーションソースを統合しています。オープンなマルチモーダル学習の民主化を目的とし、視覚と言語の統合モデルの事前学習・中間学習に活用できます。ライセンスはApache-2.0です。
README
--- license: apache-2.0 --- # 🚀 LLaVA-One-Vision-1.5-Mid-Training-85M データセットをアップロード中です 🚀 # アップロード状況 - **すべて完了**: ImageNet-21k、LAIONCN、DataComp-1B、Zero250M、COYO700M、SA-1B、MINT、Obelics # 📜 引用 研究において *LLaVA-One-Vision-1.5-Mid-Training-85M* が役立つ場合は、以下の関連論文を引用することをご検討ください: ``` @misc{an2025llavaonevis…