HTR ByteDance/Sa2VA-4B Image-Text-to-Text • 4B • Updated Sep 8, 2025 • 793 • 99 Finnish-NLP/Ahma-2-4B-Instruct Text Generation • 4B • Updated Nov 25, 2025 • 8 • 4 black-forest-labs/FLUX.2-dev Image-to-Image • 32B • Updated Feb 17 • 941k • • 2k mistralai/Mistral-Large-Instruct-2407 123B • Updated Jul 28, 2025 • 4.74k • 865
Computer Vision Vision Grid Transformer for Document Layout Analysis Paper • 2308.14978 • Published Aug 29, 2023 • 4
HTR ByteDance/Sa2VA-4B Image-Text-to-Text • 4B • Updated Sep 8, 2025 • 793 • 99 Finnish-NLP/Ahma-2-4B-Instruct Text Generation • 4B • Updated Nov 25, 2025 • 8 • 4 black-forest-labs/FLUX.2-dev Image-to-Image • 32B • Updated Feb 17 • 941k • • 2k mistralai/Mistral-Large-Instruct-2407 123B • Updated Jul 28, 2025 • 4.74k • 865
Computer Vision Vision Grid Transformer for Document Layout Analysis Paper • 2308.14978 • Published Aug 29, 2023 • 4