[ICLR 2025 Spotlight] OmniCorpus: A Unified Multimodal Corpus of 10 Billion-Level Images Interleaved with Text https://github.com/OpenGVLab/OmniCorpus
Qingyun Li
Qingyun
AI & ML interests
Object Detection, Remote Sensing
Organizations
lmmrotate 🎮
[IGARSS 2025 Oral] A Simple Aerial Detection Baseline of Multimodal Language Models. https://github.com/Li-Qingyun/mllm-mmrotate
-
A Simple Aerial Detection Baseline of Multimodal Language Models
Paper • 2501.09720 • Published • 2 -
Qingyun/lmmrotate-sft-data
Updated • 30 • 2 -
Qingyun/Florence-2-large-DOTA-v1.0-lmmrotate
Image-Text-to-Text • 0.9B • Updated • 56 • 3 -
Qingyun/Florence-2-models-lmmrotate
Image-Text-to-Text • Updated • 2
OmniCorpus 🐳
[ICLR 2025 Spotlight] OmniCorpus: A Unified Multimodal Corpus of 10 Billion-Level Images Interleaved with Text https://github.com/OpenGVLab/OmniCorpus
RSCoVLM 🤖
[Remote Sensing 2026] Co-Training Vision Language Models for Remote Sensing Multi-task Learning. https://github.com/VisionXLab/RSCoVLM
lmmrotate 🎮
[IGARSS 2025 Oral] A Simple Aerial Detection Baseline of Multimodal Language Models. https://github.com/Li-Qingyun/mllm-mmrotate
-
A Simple Aerial Detection Baseline of Multimodal Language Models
Paper • 2501.09720 • Published • 2 -
Qingyun/lmmrotate-sft-data
Updated • 30 • 2 -
Qingyun/Florence-2-large-DOTA-v1.0-lmmrotate
Image-Text-to-Text • 0.9B • Updated • 56 • 3 -
Qingyun/Florence-2-models-lmmrotate
Image-Text-to-Text • Updated • 2