PaLM 2 Bison
Legacy
PaLM 2 generation, retired in 2024 when Vertex moved fully onto Gemini.
PaLM 2 Bison is a frontier AI model from Google. It costs $0.500 per million input tokens and $0.500 per million output tokens (blended $0.500/M), with a 8,000-token context window.
INPUT
$0.500/M
per million input tokens
OUTPUT
$0.500/M
per million output tokens
CONTEXT
8,000
tokens
Benchmarks
No published benchmark scores tracked for this model yet. Frontier and reasoning models from major providers have scores; smaller models and inference-host variants typically inherit the underlying open-weights score.
More from Google
See all 24 →Gemini 3.5 Pro
Frontier · 1,000,000 ctx
in $2.000/Mout $12.000/M
Gemini 3.6 Flash
Balanced · 1,050,000 ctx
in $1.500/Mout $7.500/M
Gemini 3.5 Flash
Balanced · 1,000,000 ctx
in $1.500/Mout $9.000/M
Gemini 3.1 Pro
Frontier · 1,000,000 ctx
in $2.000/Mout $12.000/M
Gemini 3 Pro
Frontier · 1,000,000 ctx
in $2.000/Mout $12.000/M
Gemini 3 Flash
Efficient · 1,000,000 ctx
in $0.500/Mout $3.000/M
Frequently asked questions
How much does PaLM 2 Bison cost?
PaLM 2 Bison costs $0.500 per million input tokens and $0.500 per million output tokens, for a blended reference rate of $0.500 per million tokens.
What is PaLM 2 Bison's context window?
PaLM 2 Bison supports up to 8,000 tokens of context in a single request.
What is PaLM 2 Bison best for?
PaLM 2 generation, retired in 2024 when Vertex moved fully onto Gemini.
Who makes PaLM 2 Bison?
PaLM 2 Bison is developed and served by Google.