Claude Fable 5$22.000/MClaude Opus 5$11.000/MClaude Opus 4.8$11.000/MClaude Opus 4.7$11.000/MClaude Opus 4.6$11.000/MClaude Opus 4.5$33.000/MClaude Sonnet 3.7$6.600/MClaude Opus 3$33.000/MClaude 2.1$12.800/MClaude 2$12.800/MGPT-5.6 Sol$12.500/MGPT-5.6 Terra$5.000/MGPT-5.5$12.500/MGPT-5.2$5.425/MGPT-5.2-Codex$5.425/MGPT-5$3.875/MGPT-4.5$97.500/MGPT-4 Turbo Preview$16.000/MGPT-4$39.000/MGPT-4-32k$78.000/Mo3$19.000/Mo3-mini$2.090/Mo4-mini$2.090/Mo1$28.500/Mo1-mini$5.700/Mo1-preview$28.500/MGemini 3.5 Pro$5.000/MGemini 3.1 Pro$5.000/MGemini 3 Pro$5.000/MGemini 2.5 Pro$3.875/MClaude Fable 5$22.000/MClaude Opus 5$11.000/MClaude Opus 4.8$11.000/MClaude Opus 4.7$11.000/MClaude Opus 4.6$11.000/MClaude Opus 4.5$33.000/MClaude Sonnet 3.7$6.600/MClaude Opus 3$33.000/MClaude 2.1$12.800/MClaude 2$12.800/MGPT-5.6 Sol$12.500/MGPT-5.6 Terra$5.000/MGPT-5.5$12.500/MGPT-5.2$5.425/MGPT-5.2-Codex$5.425/MGPT-5$3.875/MGPT-4.5$97.500/MGPT-4 Turbo Preview$16.000/MGPT-4$39.000/MGPT-4-32k$78.000/Mo3$19.000/Mo3-mini$2.090/Mo4-mini$2.090/Mo1$28.500/Mo1-mini$5.700/Mo1-preview$28.500/MGemini 3.5 Pro$5.000/MGemini 3.1 Pro$5.000/MGemini 3 Pro$5.000/MGemini 2.5 Pro$3.875/M
DeepInfra
DeepInfra
Efficient

Phind-CodeLlama-34B (DI)

Code

Mid-size CodeLlama. Single-GPU friendly code model from 2023.

Phind-CodeLlama-34B (DI) is a efficient AI model from DeepInfra. It costs $0.600 per million input tokens and $0.600 per million output tokens (blended $0.600/M), with a 16,000-token context window.

Profile inherited from upstream CodeLlama 34B — this is a hosted variant of the same open-weights model.

INPUT
$0.600/M
per million input tokens
OUTPUT
$0.600/M
per million output tokens
BLENDED 70/30
$0.600/M
unchanged since 3 May
CONTEXT
16,000
tokens
What it is good at
  • Open weights
  • Code-tuned
  • Single-GPU
Typical use cases
  • Self-hosted code completion (legacy)

Benchmarks

vs. best public score
Scores inherited from CodeLlama 34B — this is a hosted variant of the same open-weights model, so the underlying benchmark scores are identical.
Python function synthesis from docstrings.
Hand-curated from each provider's published reports and public leaderboards. Methodology varies across sources, treat as directional rather than authoritative.

More from DeepInfra

See all 11

Frequently asked questions

How much does Phind-CodeLlama-34B (DI) cost?

Phind-CodeLlama-34B (DI) costs $0.600 per million input tokens and $0.600 per million output tokens, for a blended reference rate of $0.600 per million tokens.

What is Phind-CodeLlama-34B (DI)'s context window?

Phind-CodeLlama-34B (DI) supports up to 16,000 tokens of context in a single request.

What is Phind-CodeLlama-34B (DI) best for?

Phind-CodeLlama-34B (DI) is well suited to Open weights, Code-tuned and Single-GPU.

Who makes Phind-CodeLlama-34B (DI)?

Phind-CodeLlama-34B (DI) is developed and served by DeepInfra. It was released in Aug 2023.

Terms used on this page