Model Catalog
Approved foundation models available to OneMain teams across OpenAI, AWS Bedrock, and Google Gemini.
GPT-5.5
OpenAI
OpenAI's newest frontier reasoning model for the most complex coding and professional work, with text and image input and configurable reasoning effort.
GPT-5.5 pro
OpenAI
Highest-compute variant of GPT-5.5 that thinks harder to deliver consistently better answers on the hardest professional tasks.
GPT-5.4
OpenAI
Affordable frontier reasoning model for complex coding and professional tasks with text and image input.
GPT-5.4 mini
OpenAI
OpenAI's strongest mini model for coding, computer use, and subagents at lower cost.
GPT-5.4 nano
OpenAI
OpenAI's cheapest GPT-5-class reasoning model for simple, high-volume tasks with text and image input.
GPT-5.4 pro
OpenAI
Higher-compute variant of GPT-5.4 for more reliable answers on demanding professional tasks.
GPT Realtime 2
OpenAI
Reasoning model for realtime voice interactions with stronger instruction following and reliable tool use for voice-agent workflows.
GPT Realtime 1.5
OpenAI
Default realtime voice model for natural two-way audio conversations, voice agents, and customer support.
GPT Realtime Translate
OpenAI
Streaming speech-to-speech translation model for live multilingual audio, priced per minute of audio.
GPT Realtime Whisper
OpenAI
Streaming speech-to-text model for low-latency realtime transcription, priced per minute of audio.
GPT-4o Transcribe
OpenAI
Speech-to-text model powered by GPT-4o with improved word error rate and language recognition over original Whisper.
GPT-4o mini Transcribe
OpenAI
Lighter, lower-cost speech-to-text model for high-volume transcription.
GPT Image 2
OpenAI
State-of-the-art image generation and editing model accepting text and high-fidelity image inputs.
Sora 2
OpenAI
Video generation model priced per second of generated video.
Sora 2 Pro
OpenAI
Higher-quality video generation model priced per second of generated video.
text-embedding-3-large
OpenAI
OpenAI's most capable embedding model for English and non-English text retrieval and similarity tasks.
text-embedding-3-small
OpenAI
Cost-efficient embedding model for text retrieval and similarity at high volume.
Claude Opus 4.8
Anthropic
Anthropic's most capable model for complex reasoning, long-horizon agentic coding, and high-autonomy work, available on Bedrock.
Claude Sonnet 4.6
Anthropic
Anthropic's balanced model offering the best combination of speed and intelligence with extended and adaptive thinking.
Claude Haiku 4.5
Anthropic
Anthropic's fastest model with near-frontier intelligence and extended thinking support.
Claude Opus 4.5
Anthropic
Opus 4.5 reasoning model with extended thinking, still available on Bedrock.
Claude Sonnet 4.5
Anthropic
Balanced Sonnet 4.5 model with extended thinking, available on Bedrock.
Claude 3.5 Haiku
Anthropic
Fast, cost-efficient Claude 3.5 generation model still listed as available on Bedrock.
Amazon Nova Pro
Amazon
Amazon's balanced multimodal model offering strong accuracy, speed, and cost across text, image, and video inputs.
Amazon Nova 2 Lite
Amazon
Amazon's cost-efficient next-gen multimodal model for automation, document processing, and customer support across text, images, and video.
Amazon Nova Lite
Amazon
Amazon's low-cost multimodal model that processes text, image, and video inputs for tasks like document analysis and visual Q&A.
Amazon Nova Micro
Amazon
Amazon's fastest text-only model, optimized for speed and low cost in summarization, translation, and classification.
Amazon Nova Canvas
Amazon
Amazon's image generation model that creates studio-quality images from text and image prompts with watermarking and content moderation.
Amazon Nova Reel
Amazon
Amazon's video generation model that creates short videos from text and image prompts with camera motion controls.
Amazon Nova Sonic
Amazon
Amazon's speech-to-speech model enabling natural, real-time voice conversations with low latency and multilingual support.
Amazon Titan Text Embeddings V2
Amazon
Amazon's second-generation text embeddings model with configurable output dimensions and improved retrieval accuracy.
Amazon Nova Multimodal Embeddings
Amazon
Amazon's embedding model that converts text, images, and video into vector representations for search and retrieval.
Llama 4 Maverick 17B Instruct
Meta
Meta's 17-billion active-parameter mixture-of-experts model with 128 experts, optimized for multimodal chat and instruction following.
Llama 4 Scout 17B Instruct
Meta
Meta's 17-billion active-parameter mixture-of-experts model with 16 experts and a 10M-token context window for long-document tasks.
Llama 3.3 70B Instruct
Meta
Meta's 70-billion-parameter model with improved efficiency, delivering strong reasoning and coding with a 128K context window.
Llama 3.2 90B Instruct
Meta
Meta's 90-billion-parameter multimodal model that processes text and images for visual reasoning and document understanding.
Llama 3.2 11B Instruct
Meta
Meta's 11-billion-parameter multimodal model that processes both text and images with a 128K context window.
Llama 3.1 405B Instruct
Meta
Meta's largest open model with 405 billion parameters and a 128K context window, supporting tool use and multilingual tasks.
Llama 3.1 70B Instruct
Meta
Meta's 70-billion-parameter model with a 128K context window and support for tool use and code generation.
Llama 3.1 8B Instruct
Meta
Meta's compact 8-billion-parameter model with a 128K context window, suitable for edge deployment and fine-tuning.
Mistral Large 3
Mistral
Mistral AI's 675-billion-parameter model with strong performance on coding, reasoning, and multilingual tasks.
Mistral Large
Mistral
Mistral AI's flagship model with strong reasoning, multilingual support, and a 32K context window for complex enterprise tasks.
Pixtral Large
Mistral
Mistral AI's 124-billion-parameter multimodal model that processes text and images for visual reasoning and document understanding.
Mistral Small
Mistral
Mistral AI's cost-efficient model optimized for low-latency tasks like classification, translation, and customer support.
Mixtral 8x7B Instruct
Mistral
Mistral AI's sparse mixture-of-experts model with 8 experts of 7B parameters each, delivering strong performance at fast inference speeds.
Cohere Command R+
Cohere
Cohere's model for complex RAG workflows, multi-step tool use, and enterprise tasks with a 128K context window.
Cohere Command R
Cohere
Cohere's scalable LLM optimized for retrieval-augmented generation and tool use in enterprise applications with a 128K context window.
Cohere Embed v4
Cohere
Cohere's unified multimodal embedding model that processes text, images, and mixed content in a single model for search and RAG.
OpenAI gpt-oss-120b
OpenAI
OpenAI's open-weight 120-billion-parameter model available on Amazon Bedrock for reasoning and coding workloads.
OpenAI gpt-oss-20b
OpenAI
OpenAI's open-weight 20-billion-parameter model available on Amazon Bedrock for lightweight reasoning and chat.
Gemini 3.5 Flash
Most intelligent flash-class model for sustained frontier performance at high volume.
Gemini 3.1 Flash-Lite
Most cost-efficient multimodal model in the Gemini 3 family for high-throughput tasks.
Gemini 2.5 Pro
Advanced reasoning and coding model with deep thinking capabilities for complex tasks.
Gemini 2.5 Flash
Best price-performance model for low-latency, high-volume tasks with thinking support.
Gemini 2.5 Flash-Lite
Fastest and most budget-friendly multimodal model for high-frequency, cost-sensitive tasks.
Gemini 3 Pro Image
High-quality Gemini-native image generation and editing model for detailed visual creation.
Gemini 2.5 Flash Image (Nano Banana)
Gemini-native conversational image generation and editing model known as Nano Banana.
Imagen 4
Google's dedicated text-to-image generation model offering Fast, Standard, and Ultra quality tiers.
Veo 3
High-quality text-to-video and image-to-video generation model with synchronized audio.
Gemini Embedding 2
Latest multimodal embedding model generating vectors from text, image, audio, and video inputs.
Gemini Embedding
Text embedding model producing high-dimensional vectors for retrieval and similarity tasks.