my recommended vision/mm models Collection detection, segmentation, OCR, depth, pose, grounding, VLM detection • 17 items • Updated Jun 17 • 2
BennyDaBall/Qwen3-4b-Z-Image-Engineer-V4 Text Generation • 4B • Updated about 19 hours ago • 7.23k • 165