Large Language Models

Access large language models in one governed platform. Choose the model that fits your task and work across models without switching between separate tools.

All (35)
Claude Opus 4.5

Claude Opus 4.5

Anthropic

The next generation of Anthropic's intelligent model, Claude Opus 4.5 is an industry leader across coding, agents, computer use, and enterprise workflows.

Claude Haiku 4.5

Claude Haiku 4.5

Anthropic

Claude Haiku 4.5 delivers near-frontier performance for a wide range of use cases, and stands out as one of the best coding and agent models–with the right speed and cost to power free products and high-volume user experiences.

Claude Opus 5

Claude Opus 5

Anthropic

Claude Opus 5 is Anthropic's most intelligent model, extending the Opus line's leadership across coding, agents, computer use, and enterprise workflows.

Claude Sonnet 4.5

Claude Sonnet 4.5

Anthropic

Claude Sonnet 4.5 is Anthropic's model optimized with significant improvements across all benchmarks.

Claude Sonnet 5

Claude Sonnet 5

Anthropic

Claude Sonnet 5 is Anthropic's most powerful model for powering real-world agents, with industry-leading capabilities around coding and computer use, and the ideal balance of performance and practicality for most internal and external use cases.

Gemini 2.5 Pro

Gemini 2.5 Pro

Google

Gemini 2.5 Pro is Google's advanced reasoning Gemini model, capable of solving complex problems. It can comprehend vast datasets and challenging problems from different information sources.

Nano Banana Pro

Nano Banana Pro

Google

Gemini 3 Pro Image is designed to tackle the most challenging image generation by incorporating state-of-the-art reasoning capabilities. It is the best model for complex and multi-turn image generation and editing.

Nano Banana 2

Nano Banana 2

Google

Gemini 3.1 Flash Image (Nano Banana 2) is optimized for image understanding and generation and offers a balance of price and performance.

Gemini 3.1 Pro

Gemini 3.1 Pro

Google

Gemini 3.1 Pro is Google's most advanced reasoning Gemini model, capable of solving complex problems.

Gemini 3.5 Flash

Gemini 3.5 Flash

Google

Gemini 3.5 Flash delivers strong agentic capabilities at high speed and value, with configurable thinking levels for balancing latency and reasoning depth.

Gemini 3.5 Flash-Lite

Gemini 3.5 Flash-Lite

Google

Gemini 3.5 Flash-Lite trades depth for speed and cost efficiency, making it suitable for high-throughput classification, routing, and extraction tasks.

Gemini 3.6 Flash

Gemini 3.6 Flash

Google

Gemini 3.6 Flash is optimized for multi-step orchestration, full-stack code refactoring, and general reasoning, with improved token efficiency and code generation over Gemini 3.5 Flash.

Gemini 3.7 Flash

Gemini 3.7 Flash

Google

Gemini 3.7 Flash brings Pro-level agentic capability to the Flash tier, with large gains in code generation and terminal execution over Gemini 3.6 Flash.

GPT OSS 120b

GPT OSS 120b

gpt-oss

gpt-oss-120b is OpenAI's most powerful open-weight model, which fits into a single H100 GPU.

GPT OSS 20b

GPT OSS 20b

gpt-oss

gpt-oss-20b is OpenAI's medium-sized open-weight model for low latency, local, or specialized use-cases.

Mistral Medium

Mistral Medium

Mistral

Mistral Medium is a versatile model designed for a wide range of tasks, including programming, mathematical reasoning, understanding long documents, summarization, and dialogue.

Mistral Small

Mistral Small

Mistral

Mistral Small 3.1 (25.03) is the enhanced version of Mistral Small 3, featuring multimodal capabilities and an extended context length of up to 128k.

Mistral Large

Mistral Large

Mistral

Mistral Large 3 (675B parameters) served via AWS Bedrock; strong coding, reasoning, and multilingual capabilities with image input support.

GPT 4.1

GPT 4.1

OpenAI

GPT-4.1 excels at instruction following and tool calling, with broad knowledge across domains. It features a 1M token context window, and low latency without a reasoning step.

GPT 4o

GPT 4o

OpenAI

Fast, intelligent, flexible GPT model. GPT-4o (“o” for “omni”) is OpenAI's versatile, high-intelligence flagship model.

GPT 4o Mini

GPT 4o Mini

OpenAI

GPT-4o mini (“o” for “omni”) is a fast, affordable small model for focused tasks.

GPT 5

GPT 5

OpenAI

GPT-5 is OpenAI's previous model for coding, reasoning, and agentic tasks across domains.

GPT 5.1

GPT 5.1

OpenAI

An earlier GPT-5 model for coding and agentic tasks, with configurable reasoning and non-reasoning effort.

GPT 5.2

GPT 5.2

OpenAI

An earlier GPT-5 generation model for complex professional work.

GPT 5.4

GPT 5.4

OpenAI

OpenAI's cost-efficient model for coding and complex professional work.

GPT 5.5

GPT 5.5

OpenAI

OpenAI's previous frontier model for the most complex coding and professional work.

GPT 5.6 Luna

GPT 5.6 Luna

OpenAI

GPT-5.6 model optimized for cost-sensitive workloads.

GPT 5.6 Sol

GPT 5.6 Sol

OpenAI

OpenAI's frontier model for complex professional work.

GPT 5.6 Terra

GPT 5.6 Terra

OpenAI

GPT-5.6 model that balances intelligence and cost.

Perplexity

Perplexity

Perplexity

A lightweight, cost-effective search model optimized for quick, grounded answers with real-time web search.

Perplexity Pro

Perplexity Pro

Perplexity

An advanced search model designed for complex queries, delivering deeper content understanding with enhanced search result accuracy and 2x more search results than standard Sonar.

TAIDE

TAIDE

TAIDE

Built on Gemma-3-12b-pt, this model has been continually pre-trained on Traditional Chinese datasets and refined through instruction tuning. It is optimized for multi-turn dialogue and everyday office tasks, making it a highly capable assistant for both conversation and professional workflows.

Grok 4.2

Grok 4.2

xAI

Grok 4.2 is xAI's high-performance reasoning model with industry-leading speed and agentic tool calling, combining a very low hallucination rate with strict prompt adherence.

Grok 4.3

Grok 4.3

xAI

Grok 4.3 is xAI's fast, reliable model with strong tool calling and instruction following capabilities.

Grok 4.6

Grok 4.6

xAI

Grok 4.6 is SpaceXAI's frontier model built for coding, agentic tasks, and knowledge work.