ClovisOCR
A vision and OCR model that can extract text, convert documents to Markdown, locate elements in an image, and describe images visually. Supports DeepSearch modes with grounding.
Capacités
EntréeCe que vous pouvez envoyer
ContentPrompt + one or more images
Accepted formatsPNGJPGWEBPURLBase64
SortieCe que le modèle produit
Output formatsPlain textMarkdownCoordinates
ModesOCRDescriptionLocalization
Visual groundingYes
Custom promptYes