-
-
-
-
-
-
Inference Providers
Active filters:
Sa2VA
Image-Text-to-Text
•
4B
•
Updated
•
151k
•
93
Image-Text-to-Text
•
4B
•
Updated
•
4
Dense-World/Sa2VA_InternVL2.5_4b
Image-Text-to-Text
•
4B
•
Updated
•
6
•
1
Dense-World/Sa2VA_InternVL2.5_8b
Image-Text-to-Text
•
8B
•
Updated
•
4
Dense-World/Sa2VA_InternVL2.5_26b
Image-Text-to-Text
•
26B
•
Updated
•
6
Image-Text-to-Text
•
8B
•
Updated
•
1.28k
•
65
Image-Text-to-Text
•
1B
•
Updated
•
1.11k
•
29
Image-Text-to-Text
•
26B
•
Updated
•
78
•
31
Image Segmentation
•
4B
•
Updated
•
2
Image Segmentation
•
1B
•
Updated
•
270
Image Segmentation
•
8B
•
Updated
•
3
Image Segmentation
•
26B
•
Updated
•
1
ByteDance/Sa2VA-InternVL3-2B
Image-Text-to-Text
•
2B
•
Updated
•
174
•
1
ByteDance/Sa2VA-InternVL3-8B
Image-Text-to-Text
•
8B
•
Updated
•
79
•
4
ByteDance/Sa2VA-InternVL3-14B
Image-Text-to-Text
•
15B
•
Updated
•
47
•
9
ByteDance/Sa2VA-Qwen2_5-VL-3B
Image-Text-to-Text
•
4B
•
Updated
•
134
•
2
ByteDance/Sa2VA-Qwen2_5-VL-7B
Image-Text-to-Text
•
9B
•
Updated
•
80
•
4
ByteDance/Sa2VA-Qwen3-VL-4B
Image-Text-to-Text
•
5B
•
Updated
•
1.26k
•
14
ByteDance/Sa2VA-Qwen3-VL-2B
Image-Text-to-Text
•
3B
•
Updated
•
35
•
14