Model Catalog

Approved foundation models available to OneMain teams across OpenAI, AWS Bedrock, and Google Gemini.

16 models

GPT Realtime 2

OpenAI

Reasoning model for realtime voice interactions with stronger instruction following and reliable tool use for voice-agent workflows.

Approved
GA OpenAI

GPT Realtime 1.5

OpenAI

Default realtime voice model for natural two-way audio conversations, voice agents, and customer support.

Approved
GA OpenAI

GPT Realtime Translate

OpenAI

Streaming speech-to-speech translation model for live multilingual audio, priced per minute of audio.

Approved
GA OpenAI

GPT Realtime Whisper

OpenAI

Streaming speech-to-text model for low-latency realtime transcription, priced per minute of audio.

Approved
GA OpenAI

GPT-4o Transcribe

OpenAI

Speech-to-text model powered by GPT-4o with improved word error rate and language recognition over original Whisper.

Approved
GA OpenAI

GPT-4o mini Transcribe

OpenAI

Lighter, lower-cost speech-to-text model for high-volume transcription.

Approved
GA OpenAI
a

Amazon Nova Sonic

Amazon

Amazon's speech-to-speech model enabling natural, real-time voice conversations with low latency and multilingual support.

Approved
GA Bedrock

Gemini 3.1 Pro

Google

Google's most advanced reasoning model for complex problem-solving, deep reasoning, and frontier coding tasks.

Under review
Preview Gemini

Gemini 3.5 Flash

Google

Most intelligent flash-class model for sustained frontier performance at high volume.

Approved
GA Gemini

Gemini 3 Flash

Google

Frontier-class flash model delivering performance rivaling larger models at lower cost.

Under review
Preview Gemini

Gemini 3.1 Flash Live

Google

High-quality, low-latency Live API model for real-time bidirectional voice and video dialogue.

Under review
Preview Gemini

Gemini 2.5 Pro

Google

Advanced reasoning and coding model with deep thinking capabilities for complex tasks.

Approved
GA Gemini

Gemini 2.5 Flash

Google

Best price-performance model for low-latency, high-volume tasks with thinking support.

Approved
GA Gemini

Gemini 2.5 Flash-Lite

Google

Fastest and most budget-friendly multimodal model for high-frequency, cost-sensitive tasks.

Approved
GA Gemini

Gemini 2.5 Flash Native Audio

Google

Native audio dialog model producing natural conversational speech from multimodal input.

Under review
Preview Gemini

Gemini Embedding 2

Google

Latest multimodal embedding model generating vectors from text, image, audio, and video inputs.

Approved
GA Gemini

Please confirm

Are you sure?