Prompt
Which vision-language model supports image and PDF uploads with a large context window?
Latest observation
Jul 19, 2026 · OpenAI APIWeb search: off
A good fit is Gemini 1.5 Pro.
Why it matches:
- Supports image uploads
- Supports PDF uploads
- Has a very large context window compared with many other vision-language models
If you want, I can also give you a short comparison table of other VLMs that support images + PDFs + long context, like GPT-4o, Claude 3.5 Sonnet, and Gemini 1.5 Flash.
Observation history
Unlock observation history
Create a free Obsurfable account to unlock historical responses, model comparisons, and deeper brand intelligence.
How did Obsurfable measure this prompt?
Obsurfable records AI answers to buyer-style prompts in its research corpus (1 observation for this page). Metrics are distributions over observations, not a single static ranking.
Which AI systems does Obsurfable collect answers from?
OpenAI, ChatGPT, Google, Gemini, Google AI Mode, Anthropic, Claude, Perplexity, Grok, DeepSeek, Mistral, Copilot, and Meta AI.