Gemini

Gemma 4 26B A4B

GoogleToken-based
Alias:Gemma 4-26B-A4B
Create API Key

Compare Gemma 4-26B-A4B API pricing, supported endpoints, capabilities and access options on Modelsell.

文本生成
Starting price
Input / Output · 1M
Context
—
Maximum input window
Modalities
→

Pricing by Supplier

official
官方接口直连
Input$0.15/ 1M
Output$0.6/ 1M
Cache Read$0.015/ 1M

Capabilities / Supported modalities

StreamingFunction callingTools
Input
Output

Provider & data privacy

Provider
GoogleDocs
Tokenizer
SentencePiece (Gemini)
License
Proprietary (commercial)Proprietary
Data retention3 daysNot used for upstream training by default

Performance

Benchmarks

Scores on standardized evaluations. Higher percentages are better — and rank percentile shows

Metrics sourced fromArtificial Analysis 2026-09-30·Gemma 4 26B A4B

This model is not included in the current benchmark snapshot.

Missing models or measurements are not zero scores.

Benchmark charts preserve the source model selection and reasoning settings. Missing models or measurements are not zero scores, and benchmark cost or speed is not this site’s service commitment.

About Gemma 4 26B A4B

在效率与开放部署之间取得平衡

Gemma 4 26B A4B 面向文本生成、代码与推理任务,适合构建知识助手、内容整理工具和开发辅助应用。它采用混合专家架构,每次计算只激活部分参数,使开发者能够在模型容量与推理开销之间寻找合适的平衡。

日常使用可以从边界清楚的文本任务开始:根据提供的资料回答问题、提炼报告要点、解释代码逻辑,或将业务描述转换成初步实现。对于推理问题,要求写清条件、结论和验证方式;对于内容编辑,提供目标受众、语气和必须保留的事实,会更容易得到可用结果。

开放权重使它适合评估自行部署、应用适配和进一步训练的路线。开发者可以围绕自己的任务样本比较推理框架、量化方式与硬件配置。混合专家模型的激活参数较少,并不代表全部权重的存储需求也同样小;实际资源规划仍应结合权重格式、上下文长度和并发测试。

构建知识库问答时,可由检索系统先找出相关段落,再交给模型回答,并要求引用材料编号。这样能将语言能力与你的业务知识连接起来。若接入查询或执行工具,应在应用中定义权限、参数与返回结构,避免将生成建议误当成操作已完成。

开发者常问

开放权重有什么实际价值? 便于自行部署、研究和适配特定任务,部署质量取决于所选实现与配置,需要用真实样本验证。

使用 API 会自动学会公司的资料吗? 请求中提供的资料用于当前任务;持久知识通常由知识库检索或单独训练流程提供。

Use cases and prompting

先用代表任务评估效果

提供清楚的文本输入与输出要求,记录准确性、格式和耗时。

仅依据下面的内部操作说明,回答新同事的问题:如何处理发票抬头更正?
先给操作步骤,再列需要准备的信息和例外情况。
每一步引用资料段落编号;材料没有说明的事项标为待确认,不补充公司政策。
资料:{编号后的说明文档}

代码任务提供语言版本与样例,要求给出可运行实现和边界测试。自行部署时先确认所用权重与对话处理方式,再用同一组任务比较不同配置,避免只依据参数规模判断效果。

API access

API documentation

Code samples

RequestPOST/v1/chat/completions
Example request
Parameters
ParameterTypeDefault / rangeDescription
temperature
number
=10 ~ 2
Sampling temperature; lower is more deterministic
top_p
number
=10 ~ 1
Nucleus sampling probability mass
max_tokens
integer>= 1Maximum number of tokens in the response
frequency_penalty
number
=0-2 ~ 2
Penalises repetition of frequent tokens
presence_penalty
number
=0-2 ~ 2
Encourages introducing new topics
stop
array—Up to 4 strings that stop generation
seed
integer—Deterministic sampling seed (best-effort)
n
integer
=1>= 1
Number of completions to generate
stream
boolean
=false
Stream tokens via Server-Sent Events
response_format
object—Force JSON object or schema-conforming output
tools
array—Tool / function declarations the model may call
tool_choice
string
autononerequired
Tool-choice policy or specific tool name
logprobs
boolean
=false
Return per-token log probabilities
top_logprobs
integer0 ~ 20Number of top log probabilities returned per token
logit_bias
object—Per-token logit bias map
user
string—End-user identifier for abuse monitoring

Replace <YOUR_API_KEY> with the API key from your token settings.

Authentication

All requests must include Authorization: Bearer <TOKEN> header. Anthropic-formatted endpoints accept the x-api-key header instead.

Generate tokens from the Tokens page; you can scope them to specific models, groups, IPs, and rate-limits.

Supported parameters

Generation parameters
ParameterTypeDefault / rangeDescription
temperature
number
=10 ~ 2
Sampling temperature; lower is more deterministic
top_p
number
=10 ~ 1
Nucleus sampling probability mass
max_tokens
integer>= 1Maximum number of tokens in the response
frequency_penalty
number
=0-2 ~ 2
Penalises repetition of frequent tokens
presence_penalty
number
=0-2 ~ 2
Encourages introducing new topics
stop
array—Up to 4 strings that stop generation
seed
integer—Deterministic sampling seed (best-effort)
n
integer
=1>= 1
Number of completions to generate
stream
boolean
=false
Stream tokens via Server-Sent Events
response_format
object—Force JSON object or schema-conforming output
tools
array—Tool / function declarations the model may call
tool_choice
string
autononerequired
Tool-choice policy or specific tool name
logprobs
boolean
=false
Return per-token log probabilities
top_logprobs
integer0 ~ 20Number of top log probabilities returned per token
logit_bias
object—Per-token logit bias map
user
string—End-user identifier for abuse monitoring

Rate limits

SupplierRPMTPMRPD
officialUnlimitedUnlimitedUnlimited

No restriction

Frequently asked questions about Gemma 4-26B-A4B

What is Gemma 4-26B-A4B?

Compare Gemma 4-26B-A4B API pricing, supported endpoints, capabilities and access options on Modelsell.

How do I call Gemma 4-26B-A4B?

Create an API key with access to Gemma 4-26B-A4B, then use the exact model ID and a supported endpoint from the API access section. Request fields depend on the selected endpoint.

How is Gemma 4-26B-A4B priced?

Pricing depends on the selected provider group and the model billing unit. The current input, output, request, or media prices are shown on this page before sign-up.

How should I evaluate Gemma 4-26B-A4B for my project?

Start with the use cases and prompting guidance on this page, then evaluate the model with representative inputs from your project.