Model Catalog
Approved foundation models available to OneMain teams across OpenAI, AWS Bedrock, and Google Gemini.
GPT-5.5
OpenAI
OpenAI's newest frontier reasoning model for the most complex coding and professional work, with text and image input and configurable reasoning effort.
GPT-5.5 pro
OpenAI
Highest-compute variant of GPT-5.5 that thinks harder to deliver consistently better answers on the hardest professional tasks.
GPT-5.4
OpenAI
Affordable frontier reasoning model for complex coding and professional tasks with text and image input.
GPT-5.4 mini
OpenAI
OpenAI's strongest mini model for coding, computer use, and subagents at lower cost.
GPT-5.4 nano
OpenAI
OpenAI's cheapest GPT-5-class reasoning model for simple, high-volume tasks with text and image input.
GPT-5.4 pro
OpenAI
Higher-compute variant of GPT-5.4 for more reliable answers on demanding professional tasks.
GPT Realtime 2
OpenAI
Reasoning model for realtime voice interactions with stronger instruction following and reliable tool use for voice-agent workflows.
GPT Realtime 1.5
OpenAI
Default realtime voice model for natural two-way audio conversations, voice agents, and customer support.
GPT Image 2
OpenAI
State-of-the-art image generation and editing model accepting text and high-fidelity image inputs.
Sora 2
OpenAI
Video generation model priced per second of generated video.
Sora 2 Pro
OpenAI
Higher-quality video generation model priced per second of generated video.
Claude Opus 4.8
Anthropic
Anthropic's most capable model for complex reasoning, long-horizon agentic coding, and high-autonomy work, available on Bedrock.
Claude Sonnet 4.6
Anthropic
Anthropic's balanced model offering the best combination of speed and intelligence with extended and adaptive thinking.
Claude Haiku 4.5
Anthropic
Anthropic's fastest model with near-frontier intelligence and extended thinking support.
Claude Opus 4.5
Anthropic
Opus 4.5 reasoning model with extended thinking, still available on Bedrock.
Claude Sonnet 4.5
Anthropic
Balanced Sonnet 4.5 model with extended thinking, available on Bedrock.
Claude Opus 4.1
Anthropic
Claude Opus 4.1, deprecated by Anthropic and scheduled for retirement on August 5, 2026.
Amazon Nova Pro
Amazon
Amazon's balanced multimodal model offering strong accuracy, speed, and cost across text, image, and video inputs.
Amazon Nova 2 Lite
Amazon
Amazon's cost-efficient next-gen multimodal model for automation, document processing, and customer support across text, images, and video.
Amazon Nova Lite
Amazon
Amazon's low-cost multimodal model that processes text, image, and video inputs for tasks like document analysis and visual Q&A.
Amazon Nova Sonic
Amazon
Amazon's speech-to-speech model enabling natural, real-time voice conversations with low latency and multilingual support.
Amazon Nova Multimodal Embeddings
Amazon
Amazon's embedding model that converts text, images, and video into vector representations for search and retrieval.
Amazon Nova Premier
Amazon
Amazon's multimodal Nova model for complex reasoning and model distillation; marked Legacy with end-of-life September 14, 2026.
Llama 4 Maverick 17B Instruct
Meta
Meta's 17-billion active-parameter mixture-of-experts model with 128 experts, optimized for multimodal chat and instruction following.
Llama 4 Scout 17B Instruct
Meta
Meta's 17-billion active-parameter mixture-of-experts model with 16 experts and a 10M-token context window for long-document tasks.
Llama 3.2 90B Instruct
Meta
Meta's 90-billion-parameter multimodal model that processes text and images for visual reasoning and document understanding.
Llama 3.2 11B Instruct
Meta
Meta's 11-billion-parameter multimodal model that processes both text and images with a 128K context window.
Pixtral Large
Mistral
Mistral AI's 124-billion-parameter multimodal model that processes text and images for visual reasoning and document understanding.
Cohere Embed v4
Cohere
Cohere's unified multimodal embedding model that processes text, images, and mixed content in a single model for search and RAG.
Gemini 3.1 Pro
Google's most advanced reasoning model for complex problem-solving, deep reasoning, and frontier coding tasks.
Gemini 3.5 Flash
Most intelligent flash-class model for sustained frontier performance at high volume.
Gemini 3 Flash
Frontier-class flash model delivering performance rivaling larger models at lower cost.
Gemini 3.1 Flash-Lite
Most cost-efficient multimodal model in the Gemini 3 family for high-throughput tasks.
Gemini 3.1 Flash Live
High-quality, low-latency Live API model for real-time bidirectional voice and video dialogue.
Gemini 2.5 Pro
Advanced reasoning and coding model with deep thinking capabilities for complex tasks.
Gemini 2.5 Flash
Best price-performance model for low-latency, high-volume tasks with thinking support.
Gemini 2.5 Flash-Lite
Fastest and most budget-friendly multimodal model for high-frequency, cost-sensitive tasks.
Gemini 2.5 Flash Native Audio
Native audio dialog model producing natural conversational speech from multimodal input.
Gemini 2.5 Computer Use
Specialized model that controls a browser or UI by reasoning over screenshots to take actions.
Gemini 3 Pro Image
High-quality Gemini-native image generation and editing model for detailed visual creation.
Gemini 2.5 Flash Image (Nano Banana)
Gemini-native conversational image generation and editing model known as Nano Banana.
Gemini Embedding 2
Latest multimodal embedding model generating vectors from text, image, audio, and video inputs.