build(deps-dev): bump transformers from 4.41.2 to 4.55.0 #2241

dependabot · 2025-08-11T08:35:18Z

Bumps transformers from 4.41.2 to 4.55.0.

Release notes

v4.55.0: New openai GPT OSS model!

Welcome GPT OSS, the new open-source model family from OpenAI!

For more detailed information about this model, we recommend reading the following blogpost: https://huggingface.co/blog/welcome-openai-gpt-oss

GPT OSS is a hugely anticipated open-weights release by OpenAI, designed for powerful reasoning, agentic tasks, and versatile developer use cases. It comprises two models: a big one with 117B parameters (gpt-oss-120b), and a smaller one with 21B parameters (gpt-oss-20b). Both are mixture-of-experts (MoEs) and use a 4-bit quantization scheme (MXFP4), enabling fast inference (thanks to fewer active parameters, see details below) while keeping resource usage low. The large model fits on a single H100 GPU, while the small one runs within 16GB of memory and is perfect for consumer hardware and on-device applications.

Overview of Capabilities and Architecture

21B and 117B total parameters, with 3.6B and 5.1B active parameters, respectively.

4-bit quantization scheme using mxfp4 format. Only applied on the MoE weights. As stated, the 120B fits in a single 80 GB GPU and the 20B fits in a single 16GB GPU.

Reasoning, text-only models; with chain-of-thought and adjustable reasoning effort levels.

Instruction following and tool use support.

Inference implementations using transformers, vLLM, llama.cpp, and ollama.

Responses API is recommended for inference.

License: Apache 2.0, with a small complementary use policy.

Architecture

Token-choice MoE with SwiGLU activations.

When calculating the MoE weights, a softmax is taken over selected experts (softmax-after-topk).

Each attention layer uses RoPE with 128K context.

Alternate attention layers: full-context, and sliding 128-token window.

Attention layers use a learned attention sink per-head, where the denominator of the softmax has an additional additive value.

It uses the same tokenizer as GPT-4o and other OpenAI API models.

Some new tokens have been incorporated to enable compatibility with the Responses API.

The following snippet shows simple inference with the 20B model. It runs on 16 GB GPUs when using mxfp4, or ~48 GB in bfloat16.
from transformers import AutoModelForCausalLM, AutoTokenizer
model_id = "openai/gpt-oss-20b"
tokenizer = AutoTokenizer.from_pretrained(model_id)
model = AutoModelForCausalLM.from_pretrained(
model_id,
device_map="auto",
torch_dtype="auto",
)
messages = [
{"role": "user", "content": "How many rs are in the word 'strawberry'?"},
]
inputs = tokenizer.apply_chat_template(
messages,
</tr></table>

... (truncated)

Commits

06f8004 Release: v4.55.0
c54203a gpt_oss last chat template changes (#39925)
7c38d8f Add GPT OSS model from OpenAI (#39923)
738c1a3 🌐 [i18n-KO] Translated cache_explanation.md to Korean (#39535)
d2ae766 Export SmolvLM (#39614)
c430047 [docs] update object detection guide (#39909)
dedcbd6 run model debugging with forward arg (#39905)
20ce210 Revert "remove dtensors, not explicit (#39840)" (#39912)
2589a52 Fix aria tests (#39879)
6e4a9a5 Fix eval thread fork bomb (#39717)
Additional commits viewable in compare view

Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting @dependabot rebase.

Dependabot commands and options

You can trigger Dependabot actions by commenting on this PR:

@dependabot rebase will rebase this PR
@dependabot recreate will recreate this PR, overwriting any edits that have been made to it
@dependabot merge will merge this PR after your CI passes on it
@dependabot squash and merge will squash and merge this PR after your CI passes on it
@dependabot cancel merge will cancel a previously requested merge and block automerging
@dependabot reopen will reopen this PR if it is closed
@dependabot close will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually
@dependabot show <dependency name> ignore conditions will show all of the ignore conditions of the specified dependency
@dependabot ignore this major version will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself)
@dependabot ignore this minor version will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself)
@dependabot ignore this dependency will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself)

Bumps [transformers](https://github.com/huggingface/transformers) from 4.41.2 to 4.55.0. - [Release notes](https://github.com/huggingface/transformers/releases) - [Commits](huggingface/transformers@v4.41.2...v4.55.0) --- updated-dependencies: - dependency-name: transformers dependency-version: 4.55.0 dependency-type: direct:development update-type: version-update:semver-minor ... Signed-off-by: dependabot[bot] <support@github.com>

dependabot · 2025-08-25T09:51:06Z

Superseded by #2256.

dependabot bot added dependencies Pull requests that update a dependency file python Pull requests that update Python code labels Aug 11, 2025

dependabot bot mentioned this pull request Aug 11, 2025

build(deps-dev): bump transformers from 4.41.2 to 4.54.1 #2237

Closed

dependabot bot requested a review from a team August 11, 2025 08:35

dependabot bot closed this Aug 25, 2025

dependabot bot deleted the dependabot/pip/transformers-4.55.0 branch August 25, 2025 09:51

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Uh oh!

build(deps-dev): bump transformers from 4.41.2 to 4.55.0 #2241

build(deps-dev): bump transformers from 4.41.2 to 4.55.0 #2241

Uh oh!

dependabot bot commented on behalf of github Aug 11, 2025

Uh oh!

dependabot bot commented on behalf of github Aug 25, 2025

Uh oh!

Reviewers

Assignees

Labels

Projects

Milestone

Development

Uh oh!

1 participant

build(deps-dev): bump transformers from 4.41.2 to 4.55.0 #2241

build(deps-dev): bump transformers from 4.41.2 to 4.55.0 #2241

Uh oh!

Conversation

dependabot bot commented on behalf of github Aug 11, 2025

v4.55.0: New openai GPT OSS model!

Welcome GPT OSS, the new open-source model family from OpenAI!

Overview of Capabilities and Architecture

Architecture

Uh oh!

dependabot bot commented on behalf of github Aug 25, 2025

Uh oh!

Reviewers

Assignees

Labels

Projects

Milestone

Development

Uh oh!

1 participant