Skip to content

Google Gemma 4 26B A4B Instruct

Gemma 4 26B A4B is a Mixture-of-Experts model from Google DeepMind with 26B total parameters and only 4B active per token, offering fast inference at high quality. It handles text, image, and video input, supports 256K context, function calling, and reasoning with configurable thinking modes.

Providers
Capabilities
Venice AI
venice-byok
$0
$0
Unavailable
Usage analytics

Loading usage…

API & code
Uptime & Health
No uptime data yet

These providers haven't been health-probed for this model yet. The router still routes around upstreams that fail live requests — uptime fills in once probe history accrues.

Share cards
Google Gemma 4 26B A4B Instruct share card
Google Gemma 4 26B A4B Instruct
Venice AI upstream share card
Venice AI upstream
Credits
Use your own key

Run Google Gemma 4 26B A4B Instruct on your own key — your requests are billed by the provider. Pool callers pay AnyRouter credits.

No BYOK keys configured for this model yet.

Share a key with the pool to earn credits for every request it serves, covering your plan cost.

Text generation
Context length256,000 tokens
Max output204,800 tokens
ArchitectureTransformer
Categorytext
ReleasedApr 2, 2026
Modalities
→
Capabilities