Gemma 4 26B A4B
GoogleGemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...
Entrée
$0.09
Sortie
$0.30
Blended
$0.14
Où acheter ce modèle
11 routes d'accès| Route | Entrée | Sortie | Blended | vs |
|---|---|---|---|---|
| ● Darkbloom le moins cher | $0.04 | $0.22 | $0.09 | - |
| ● DekaLLM bf16 | $0.06 | $0.33 | $0.13 | +44% |
| $0.07 | $0.34 | $0.14 | +56% | |
| ● NextBit bf16 | $0.09 | $0.30 | $0.14 | +56% |
| ● Cloudflare | $0.10 | $0.30 | $0.15 | +67% |
| ● Makora | $0.10 | $0.34 | $0.16 | +78% |
| ● Venice bf16 | $0.13 | $0.40 | $0.20 | +122% |
| $0.13 | $0.40 | $0.20 | +122% | |
| $0.13 | $0.40 | $0.20 | +122% | |
| ● SiliconFlow fp8 | $0.14 | $0.40 | $0.21 | +133% |
| $0.15 | $0.60 | $0.26 | +189% |
● API officielle · ● plateforme cloud · ● hébergeur d'inférence · USD par million de tokens. Même modèle, fournisseurs différents - le prix et la quantization varient.
- Ctx
- 262K
- max output
- 236K
- cache read /1M
- $0.05
- Famille
- gemma
- Modalités
- image · text · video
- Fonctions
- json · reasoning · tools
Prix live, mis à jour le 2026-09-15.