Model Catalog
Approved foundation models available to OneMain teams across OpenAI, AWS Bedrock, and Google Gemini.
GPT-5.5
OpenAIOpenAI's newest frontier reasoning model for the most complex coding and professional work, with text and image input and configurable reasoning effort.
GPT-5.5 pro
OpenAIHighest-compute variant of GPT-5.5 that thinks harder to deliver consistently better answers on the hardest professional tasks.
GPT-5.4
OpenAIAffordable frontier reasoning model for complex coding and professional tasks with text and image input.
GPT-5.4 mini
OpenAIOpenAI's strongest mini model for coding, computer use, and subagents at lower cost.
GPT-5.4 nano
OpenAIOpenAI's cheapest GPT-5-class reasoning model for simple, high-volume tasks with text and image input.
GPT-5.4 pro
OpenAIHigher-compute variant of GPT-5.4 for more reliable answers on demanding professional tasks.
GPT Realtime 2
OpenAIReasoning model for realtime voice interactions with stronger instruction following and reliable tool use for voice-agent workflows.
GPT Realtime 1.5
OpenAIDefault realtime voice model for natural two-way audio conversations, voice agents, and customer support.
GPT Realtime Translate
OpenAIStreaming speech-to-speech translation model for live multilingual audio, priced per minute of audio.
GPT Realtime Whisper
OpenAIStreaming speech-to-text model for low-latency realtime transcription, priced per minute of audio.
GPT-4o Transcribe
OpenAISpeech-to-text model powered by GPT-4o with improved word error rate and language recognition over original Whisper.
GPT-4o mini Transcribe
OpenAILighter, lower-cost speech-to-text model for high-volume transcription.
GPT Image 2
OpenAIState-of-the-art image generation and editing model accepting text and high-fidelity image inputs.
Sora 2
OpenAIVideo generation model priced per second of generated video.
Sora 2 Pro
OpenAIHigher-quality video generation model priced per second of generated video.
text-embedding-3-large
OpenAIOpenAI's most capable embedding model for English and non-English text retrieval and similarity tasks.
text-embedding-3-small
OpenAICost-efficient embedding model for text retrieval and similarity at high volume.
Claude Opus 4.8
AnthropicAnthropic's most capable model for complex reasoning, long-horizon agentic coding, and high-autonomy work, available on Bedrock.
Claude Sonnet 4.6
AnthropicAnthropic's balanced model offering the best combination of speed and intelligence with extended and adaptive thinking.
Claude Haiku 4.5
AnthropicAnthropic's fastest model with near-frontier intelligence and extended thinking support.
Claude Opus 4.5
AnthropicOpus 4.5 reasoning model with extended thinking, still available on Bedrock.
Claude Sonnet 4.5
AnthropicBalanced Sonnet 4.5 model with extended thinking, available on Bedrock.
Claude 3.5 Haiku
AnthropicFast, cost-efficient Claude 3.5 generation model still listed as available on Bedrock.
Claude Opus 4.1
AnthropicClaude Opus 4.1, deprecated by Anthropic and scheduled for retirement on August 5, 2026.
Amazon Nova Pro
AmazonAmazon's balanced multimodal model offering strong accuracy, speed, and cost across text, image, and video inputs.
Amazon Nova 2 Lite
AmazonAmazon's cost-efficient next-gen multimodal model for automation, document processing, and customer support across text, images, and video.
Amazon Nova Lite
AmazonAmazon's low-cost multimodal model that processes text, image, and video inputs for tasks like document analysis and visual Q&A.
Amazon Nova Micro
AmazonAmazon's fastest text-only model, optimized for speed and low cost in summarization, translation, and classification.
Amazon Nova Canvas
AmazonAmazon's image generation model that creates studio-quality images from text and image prompts with watermarking and content moderation.
Amazon Nova Reel
AmazonAmazon's video generation model that creates short videos from text and image prompts with camera motion controls.
Amazon Nova Sonic
AmazonAmazon's speech-to-speech model enabling natural, real-time voice conversations with low latency and multilingual support.
Amazon Titan Text Embeddings V2
AmazonAmazon's second-generation text embeddings model with configurable output dimensions and improved retrieval accuracy.
Amazon Nova Multimodal Embeddings
AmazonAmazon's embedding model that converts text, images, and video into vector representations for search and retrieval.
Amazon Nova Premier
AmazonAmazon's multimodal Nova model for complex reasoning and model distillation; marked Legacy with end-of-life September 14, 2026.
Llama 4 Maverick 17B Instruct
MetaMeta's 17-billion active-parameter mixture-of-experts model with 128 experts, optimized for multimodal chat and instruction following.
Llama 4 Scout 17B Instruct
MetaMeta's 17-billion active-parameter mixture-of-experts model with 16 experts and a 10M-token context window for long-document tasks.
Llama 3.3 70B Instruct
MetaMeta's 70-billion-parameter model with improved efficiency, delivering strong reasoning and coding with a 128K context window.
Llama 3.2 90B Instruct
MetaMeta's 90-billion-parameter multimodal model that processes text and images for visual reasoning and document understanding.
Llama 3.2 11B Instruct
MetaMeta's 11-billion-parameter multimodal model that processes both text and images with a 128K context window.
Llama 3.1 405B Instruct
MetaMeta's largest open model with 405 billion parameters and a 128K context window, supporting tool use and multilingual tasks.
Llama 3.1 70B Instruct
MetaMeta's 70-billion-parameter model with a 128K context window and support for tool use and code generation.
Llama 3.1 8B Instruct
MetaMeta's compact 8-billion-parameter model with a 128K context window, suitable for edge deployment and fine-tuning.
Mistral Large 3
MistralMistral AI's 675-billion-parameter model with strong performance on coding, reasoning, and multilingual tasks.
Mistral Large
MistralMistral AI's flagship model with strong reasoning, multilingual support, and a 32K context window for complex enterprise tasks.
Pixtral Large
MistralMistral AI's 124-billion-parameter multimodal model that processes text and images for visual reasoning and document understanding.
Mistral Small
MistralMistral AI's cost-efficient model optimized for low-latency tasks like classification, translation, and customer support.
Mixtral 8x7B Instruct
MistralMistral AI's sparse mixture-of-experts model with 8 experts of 7B parameters each, delivering strong performance at fast inference speeds.
Cohere Command R+
CohereCohere's model for complex RAG workflows, multi-step tool use, and enterprise tasks with a 128K context window.
Cohere Command R
CohereCohere's scalable LLM optimized for retrieval-augmented generation and tool use in enterprise applications with a 128K context window.
Cohere Embed v4
CohereCohere's unified multimodal embedding model that processes text, images, and mixed content in a single model for search and RAG.
OpenAI gpt-oss-120b
OpenAIOpenAI's open-weight 120-billion-parameter model available on Amazon Bedrock for reasoning and coding workloads.
OpenAI gpt-oss-20b
OpenAIOpenAI's open-weight 20-billion-parameter model available on Amazon Bedrock for lightweight reasoning and chat.
Gemini 3.1 Pro
GoogleGoogle's most advanced reasoning model for complex problem-solving, deep reasoning, and frontier coding tasks.
Gemini 3.5 Flash
GoogleMost intelligent flash-class model for sustained frontier performance at high volume.
Gemini 3 Flash
GoogleFrontier-class flash model delivering performance rivaling larger models at lower cost.
Gemini 3.1 Flash-Lite
GoogleMost cost-efficient multimodal model in the Gemini 3 family for high-throughput tasks.
Gemini 3.1 Flash Live
GoogleHigh-quality, low-latency Live API model for real-time bidirectional voice and video dialogue.
Gemini 2.5 Pro
GoogleAdvanced reasoning and coding model with deep thinking capabilities for complex tasks.
Gemini 2.5 Flash
GoogleBest price-performance model for low-latency, high-volume tasks with thinking support.
Gemini 2.5 Flash-Lite
GoogleFastest and most budget-friendly multimodal model for high-frequency, cost-sensitive tasks.
Gemini 2.5 Flash Native Audio
GoogleNative audio dialog model producing natural conversational speech from multimodal input.
Gemini 2.5 Computer Use
GoogleSpecialized model that controls a browser or UI by reasoning over screenshots to take actions.
Gemini 3 Pro Image
GoogleHigh-quality Gemini-native image generation and editing model for detailed visual creation.
Gemini 2.5 Flash Image (Nano Banana)
GoogleGemini-native conversational image generation and editing model known as Nano Banana.
Imagen 4
GoogleGoogle's dedicated text-to-image generation model offering Fast, Standard, and Ultra quality tiers.
Veo 3.1
GoogleState-of-the-art cinematic text-to-video and image-to-video generation model with native audio.
Veo 3
GoogleHigh-quality text-to-video and image-to-video generation model with synchronized audio.
Gemini Embedding 2
GoogleLatest multimodal embedding model generating vectors from text, image, audio, and video inputs.
Gemini Embedding
GoogleText embedding model producing high-dimensional vectors for retrieval and similarity tasks.