Guides
Cortiqa models can process both text and images in a unified request, enabling visual reasoning, document OCR, chart analysis, and user interface inspection.
You can provide images as base64-encoded strings or public URLs inside the message content array.
falin-vision, 128K context) will provide state-of-the-art native document perception, fine-grained diagram OCR, and spatial UI navigation.1import base642from cortiqa import Cortiqa34client = Cortiqa()56with open("architecture_diagram.png", "rb") as f:7 base64_image = base64.b64encode(f.read()).decode("utf-8")89response = client.chat.completions.create(10 model="openai/gpt-oss-120b",11 messages=[12 {13 "role": "user",14 "content": [15 {"type": "text", "text": "Analyze this architecture diagram and highlight any single points of failure."},16 {17 "type": "image_url",18 "image_url": {19 "url": f"data:image/png;base64,{base64_image}"20 }21 }22 ]23 }24 ]25)2627print(response.choices[0].message.content)| Format | MIME Type | Max Upload Size |
|---|---|---|
| PNG | image/png | 20 MB |
| JPEG | image/jpeg | 20 MB |
| WebP | image/webp | 20 MB |
| GIF | image/gif | 20 MB |