Model Catalog
Approved foundation models available to OneMain teams across OpenAI, AWS Bedrock, and Google Gemini.
GPT-5.5
OpenAIOpenAI's newest frontier reasoning model for the most complex coding and professional work, with text and image input and configurable reasoning effort.
GPT-5.5 pro
OpenAIHighest-compute variant of GPT-5.5 that thinks harder to deliver consistently better answers on the hardest professional tasks.
GPT-5.4
OpenAIAffordable frontier reasoning model for complex coding and professional tasks with text and image input.
GPT-5.4 mini
OpenAIOpenAI's strongest mini model for coding, computer use, and subagents at lower cost.
GPT-5.4 nano
OpenAIOpenAI's cheapest GPT-5-class reasoning model for simple, high-volume tasks with text and image input.
GPT-5.4 pro
OpenAIHigher-compute variant of GPT-5.4 for more reliable answers on demanding professional tasks.
GPT Realtime 2
OpenAIReasoning model for realtime voice interactions with stronger instruction following and reliable tool use for voice-agent workflows.
GPT Realtime 1.5
OpenAIDefault realtime voice model for natural two-way audio conversations, voice agents, and customer support.
GPT Image 2
OpenAIState-of-the-art image generation and editing model accepting text and high-fidelity image inputs.
Claude Opus 4.8
AnthropicAnthropic's most capable model for complex reasoning, long-horizon agentic coding, and high-autonomy work, available on Bedrock.
Claude Sonnet 4.6
AnthropicAnthropic's balanced model offering the best combination of speed and intelligence with extended and adaptive thinking.
Claude Haiku 4.5
AnthropicAnthropic's fastest model with near-frontier intelligence and extended thinking support.
Claude Opus 4.5
AnthropicOpus 4.5 reasoning model with extended thinking, still available on Bedrock.
Claude Sonnet 4.5
AnthropicBalanced Sonnet 4.5 model with extended thinking, available on Bedrock.
Claude Opus 4.1
AnthropicClaude Opus 4.1, deprecated by Anthropic and scheduled for retirement on August 5, 2026.
Amazon Nova Pro
AmazonAmazon's balanced multimodal model offering strong accuracy, speed, and cost across text, image, and video inputs.
Amazon Nova 2 Lite
AmazonAmazon's cost-efficient next-gen multimodal model for automation, document processing, and customer support across text, images, and video.
Amazon Nova Lite
AmazonAmazon's low-cost multimodal model that processes text, image, and video inputs for tasks like document analysis and visual Q&A.
Amazon Nova Canvas
AmazonAmazon's image generation model that creates studio-quality images from text and image prompts with watermarking and content moderation.
Amazon Nova Premier
AmazonAmazon's multimodal Nova model for complex reasoning and model distillation; marked Legacy with end-of-life September 14, 2026.
Llama 4 Maverick 17B Instruct
MetaMeta's 17-billion active-parameter mixture-of-experts model with 128 experts, optimized for multimodal chat and instruction following.
Llama 4 Scout 17B Instruct
MetaMeta's 17-billion active-parameter mixture-of-experts model with 16 experts and a 10M-token context window for long-document tasks.
Llama 3.2 90B Instruct
MetaMeta's 90-billion-parameter multimodal model that processes text and images for visual reasoning and document understanding.
Llama 3.2 11B Instruct
MetaMeta's 11-billion-parameter multimodal model that processes both text and images with a 128K context window.
Pixtral Large
MistralMistral AI's 124-billion-parameter multimodal model that processes text and images for visual reasoning and document understanding.
Gemini 3.1 Pro
GoogleGoogle's most advanced reasoning model for complex problem-solving, deep reasoning, and frontier coding tasks.
Gemini 3.5 Flash
GoogleMost intelligent flash-class model for sustained frontier performance at high volume.
Gemini 3 Flash
GoogleFrontier-class flash model delivering performance rivaling larger models at lower cost.
Gemini 3.1 Flash-Lite
GoogleMost cost-efficient multimodal model in the Gemini 3 family for high-throughput tasks.
Gemini 3.1 Flash Live
GoogleHigh-quality, low-latency Live API model for real-time bidirectional voice and video dialogue.
Gemini 2.5 Pro
GoogleAdvanced reasoning and coding model with deep thinking capabilities for complex tasks.
Gemini 2.5 Flash
GoogleBest price-performance model for low-latency, high-volume tasks with thinking support.
Gemini 2.5 Flash-Lite
GoogleFastest and most budget-friendly multimodal model for high-frequency, cost-sensitive tasks.
Gemini 2.5 Computer Use
GoogleSpecialized model that controls a browser or UI by reasoning over screenshots to take actions.
Gemini 3 Pro Image
GoogleHigh-quality Gemini-native image generation and editing model for detailed visual creation.
Gemini 2.5 Flash Image (Nano Banana)
GoogleGemini-native conversational image generation and editing model known as Nano Banana.
Imagen 4
GoogleGoogle's dedicated text-to-image generation model offering Fast, Standard, and Ultra quality tiers.
Veo 3.1
GoogleState-of-the-art cinematic text-to-video and image-to-video generation model with native audio.
Veo 3
GoogleHigh-quality text-to-video and image-to-video generation model with synchronized audio.
Gemini Embedding 2
GoogleLatest multimodal embedding model generating vectors from text, image, audio, and video inputs.