Multimodal AI
4 items tagged with "multimodal-ai"
Models4
Amazon Nova 2 Lite
Amazon Nova 2 Lite is a lightweight multimodal Nova model referenced for cost-optimized scanned document processing, where it handles native multimodal extraction before downstream Claude processing.
Gemini 3.1 Flash Image
Google Gemini image-focused model listed on OpenRouter with a 131,072-token context window. It is positioned as a Flash-tier multimodal/image model for lower-latency image-centric workloads.
Gemini 3 Pro Image
Google Gemini Pro-tier image-focused model listed on OpenRouter with a 65,536-token context window. It targets higher-capability multimodal and image-generation use cases than Flash-tier variants.
Gemini Omni
Gemini Omni is a Google Gemini-family model announced at Google I/O 2026 and showcased in Google demos alongside Gemini 3.5. The discovery data verifies the announcement but does not provide context length, output limit, or pricing details.