Galileo AI: The AI Observability and Evaluation Platform
Galileo
AIModels
See how popular AI Models perform against different benchmarks for AI agents and applications.
Individual AI Model Overview
Nova Lite V1](/content/model-hub/nova-lite-v1-overview/index.html)
Mistral Small 2506](/content/model-hub/mistral-small-2506-overview/index.html)
Magistral Small 2506](/content/model-hub/magistral-small-2506-overview/index.html)
Llama 3.3 70B Instruct](/content/model-hub/llama-3-3-70b-overview/index.html)
Caller](/content/model-hub/caller-overview/index.html)
Qwen3-235B-A22B-Thinking-2507 Overview](/content/model-hub/qwen3-235b-a22b-thinking-2507-overview/index.html)
Nova Pro v1](/content/model-hub/nova-pro-v1-overview/index.html)
Gemini-2.5-flash Overview](/content/model-hub/gemini-2-5-flash-overview/index.html)
Magistral Medium 2506 overview](/content/model-hub/magistral-medium-2506-overview/index.html)
Grok-4-0709 Overview](/content/model-hub/grok-4-0709-overview/index.html)
GPT-4.1-nano overview](/content/model-hub/gpt-4-1-nano-overview/index.html)
Deepseek V3 Overview](/content/model-hub/deepseek-v3-overview/index.html)
Gemini 2.5 Pro overview](/content/model-hub/gemini-2-5-pro-overview/index.html)
GLM 4.5 Air Overview](/content/model-hub/glm-4-5-air-overview/index.html)
Gemini 2.5 flash lite Overview](/content/model-hub/gemini-2-5-flash-lite-overview/index.html)
Qwen3 235B A22B Instruct 2507 Overview](/content/model-hub/qwen3-235b-a22b-instruct-2507-overview/index.html)
Qwen2.5 72B Instruct Overview](/content/model-hub/qwen-2-5-72b-instruct-overview/index.html)
Kimi K2 Instruct Model Overview](/content/model-hub/kimi-k2-instruct-overview/index.html)
GPT-4.1 Mini Overview](/content/model-hub/gpt-4-1-mini-overview/index.html)
Mistral Medium 2508 Overview](/content/model-hub/mistral-medium-2508-overview/index.html)
Claude Sonnet 4 Overview](/content/model-hub/claude-sonnet-4-overview/index.html)