InternVL 2
Open multimodal model rivaling GPT-4V — image, video, OCR and doc understanding.
View on AIWEBTOOLS.AI