Model Catalog
Approved foundation models available to OneMain teams across OpenAI, AWS Bedrock, and Google Gemini.
GPT-5.5
OpenAI
OpenAI's newest frontier reasoning model for the most complex coding and professional work, with text and image input and configurable reasoning effort.
GPT-5.5 pro
OpenAI
Highest-compute variant of GPT-5.5 that thinks harder to deliver consistently better answers on the hardest professional tasks.
GPT-5.4
OpenAI
Affordable frontier reasoning model for complex coding and professional tasks with text and image input.
GPT-5.4 mini
OpenAI
OpenAI's strongest mini model for coding, computer use, and subagents at lower cost.
GPT-5.4 nano
OpenAI
OpenAI's cheapest GPT-5-class reasoning model for simple, high-volume tasks with text and image input.
GPT-5.4 pro
OpenAI
Higher-compute variant of GPT-5.4 for more reliable answers on demanding professional tasks.
GPT Realtime 2
OpenAI
Reasoning model for realtime voice interactions with stronger instruction following and reliable tool use for voice-agent workflows.
GPT Realtime 1.5
OpenAI
Default realtime voice model for natural two-way audio conversations, voice agents, and customer support.
GPT Realtime Translate
OpenAI
Streaming speech-to-speech translation model for live multilingual audio, priced per minute of audio.
GPT Realtime Whisper
OpenAI
Streaming speech-to-text model for low-latency realtime transcription, priced per minute of audio.
GPT-4o Transcribe
OpenAI
Speech-to-text model powered by GPT-4o with improved word error rate and language recognition over original Whisper.
GPT-4o mini Transcribe
OpenAI
Lighter, lower-cost speech-to-text model for high-volume transcription.
GPT Image 2
OpenAI
State-of-the-art image generation and editing model accepting text and high-fidelity image inputs.
Sora 2
OpenAI
Video generation model priced per second of generated video.
Sora 2 Pro
OpenAI
Higher-quality video generation model priced per second of generated video.
text-embedding-3-large
OpenAI
OpenAI's most capable embedding model for English and non-English text retrieval and similarity tasks.
text-embedding-3-small
OpenAI
Cost-efficient embedding model for text retrieval and similarity at high volume.
Claude Opus 4.8
Anthropic
Anthropic's most capable model for complex reasoning, long-horizon agentic coding, and high-autonomy work, available on Bedrock.
Claude Sonnet 4.6
Anthropic
Anthropic's balanced model offering the best combination of speed and intelligence with extended and adaptive thinking.
Claude Haiku 4.5
Anthropic
Anthropic's fastest model with near-frontier intelligence and extended thinking support.
Claude Opus 4.5
Anthropic
Opus 4.5 reasoning model with extended thinking, still available on Bedrock.
Claude Sonnet 4.5
Anthropic
Balanced Sonnet 4.5 model with extended thinking, available on Bedrock.
Claude 3.5 Haiku
Anthropic
Fast, cost-efficient Claude 3.5 generation model still listed as available on Bedrock.
Claude Opus 4.1
Anthropic
Claude Opus 4.1, deprecated by Anthropic and scheduled for retirement on August 5, 2026.
Amazon Nova Pro
Amazon
Amazon's balanced multimodal model offering strong accuracy, speed, and cost across text, image, and video inputs.
Amazon Nova 2 Lite
Amazon
Amazon's cost-efficient next-gen multimodal model for automation, document processing, and customer support across text, images, and video.
Amazon Nova Lite
Amazon
Amazon's low-cost multimodal model that processes text, image, and video inputs for tasks like document analysis and visual Q&A.
Amazon Nova Micro
Amazon
Amazon's fastest text-only model, optimized for speed and low cost in summarization, translation, and classification.
Amazon Titan Text Embeddings V2
Amazon
Amazon's second-generation text embeddings model with configurable output dimensions and improved retrieval accuracy.
Amazon Nova Premier
Amazon
Amazon's multimodal Nova model for complex reasoning and model distillation; marked Legacy with end-of-life September 14, 2026.
Llama 4 Maverick 17B Instruct
Meta
Meta's 17-billion active-parameter mixture-of-experts model with 128 experts, optimized for multimodal chat and instruction following.
Llama 4 Scout 17B Instruct
Meta
Meta's 17-billion active-parameter mixture-of-experts model with 16 experts and a 10M-token context window for long-document tasks.
Llama 3.3 70B Instruct
Meta
Meta's 70-billion-parameter model with improved efficiency, delivering strong reasoning and coding with a 128K context window.
Llama 3.2 90B Instruct
Meta
Meta's 90-billion-parameter multimodal model that processes text and images for visual reasoning and document understanding.
Llama 3.2 11B Instruct
Meta
Meta's 11-billion-parameter multimodal model that processes both text and images with a 128K context window.
Llama 3.1 405B Instruct
Meta
Meta's largest open model with 405 billion parameters and a 128K context window, supporting tool use and multilingual tasks.
Llama 3.1 70B Instruct
Meta
Meta's 70-billion-parameter model with a 128K context window and support for tool use and code generation.
Llama 3.1 8B Instruct
Meta
Meta's compact 8-billion-parameter model with a 128K context window, suitable for edge deployment and fine-tuning.
Mistral Large 3
Mistral
Mistral AI's 675-billion-parameter model with strong performance on coding, reasoning, and multilingual tasks.
Mistral Large
Mistral
Mistral AI's flagship model with strong reasoning, multilingual support, and a 32K context window for complex enterprise tasks.
Pixtral Large
Mistral
Mistral AI's 124-billion-parameter multimodal model that processes text and images for visual reasoning and document understanding.
Mistral Small
Mistral
Mistral AI's cost-efficient model optimized for low-latency tasks like classification, translation, and customer support.
Mixtral 8x7B Instruct
Mistral
Mistral AI's sparse mixture-of-experts model with 8 experts of 7B parameters each, delivering strong performance at fast inference speeds.
Cohere Command R+
Cohere
Cohere's model for complex RAG workflows, multi-step tool use, and enterprise tasks with a 128K context window.
Cohere Command R
Cohere
Cohere's scalable LLM optimized for retrieval-augmented generation and tool use in enterprise applications with a 128K context window.
OpenAI gpt-oss-120b
OpenAI
OpenAI's open-weight 120-billion-parameter model available on Amazon Bedrock for reasoning and coding workloads.
OpenAI gpt-oss-20b
OpenAI
OpenAI's open-weight 20-billion-parameter model available on Amazon Bedrock for lightweight reasoning and chat.
Gemini 3.1 Pro
Google's most advanced reasoning model for complex problem-solving, deep reasoning, and frontier coding tasks.
Gemini 3.5 Flash
Most intelligent flash-class model for sustained frontier performance at high volume.
Gemini 3 Flash
Frontier-class flash model delivering performance rivaling larger models at lower cost.
Gemini 3.1 Flash-Lite
Most cost-efficient multimodal model in the Gemini 3 family for high-throughput tasks.
Gemini 3.1 Flash Live
High-quality, low-latency Live API model for real-time bidirectional voice and video dialogue.
Gemini 2.5 Pro
Advanced reasoning and coding model with deep thinking capabilities for complex tasks.
Gemini 2.5 Flash
Best price-performance model for low-latency, high-volume tasks with thinking support.
Gemini 2.5 Flash-Lite
Fastest and most budget-friendly multimodal model for high-frequency, cost-sensitive tasks.
Gemini 2.5 Flash Native Audio
Native audio dialog model producing natural conversational speech from multimodal input.
Gemini 2.5 Computer Use
Specialized model that controls a browser or UI by reasoning over screenshots to take actions.
Gemini 3 Pro Image
High-quality Gemini-native image generation and editing model for detailed visual creation.
Gemini 2.5 Flash Image (Nano Banana)
Gemini-native conversational image generation and editing model known as Nano Banana.
Imagen 4
Google's dedicated text-to-image generation model offering Fast, Standard, and Ultra quality tiers.
Veo 3.1
State-of-the-art cinematic text-to-video and image-to-video generation model with native audio.
Veo 3
High-quality text-to-video and image-to-video generation model with synchronized audio.
Gemini Embedding 2
Latest multimodal embedding model generating vectors from text, image, audio, and video inputs.
Gemini Embedding
Text embedding model producing high-dimensional vectors for retrieval and similarity tasks.