Bump transformers from 4.53.2 to 4.55.0 #1136

dependabot · 2025-08-11T19:26:31Z

Bumps transformers from 4.53.2 to 4.55.0.

Release notes

v4.55.0: New openai GPT OSS model!

Welcome GPT OSS, the new open-source model family from OpenAI!

For more detailed information about this model, we recommend reading the following blogpost: https://huggingface.co/blog/welcome-openai-gpt-oss

GPT OSS is a hugely anticipated open-weights release by OpenAI, designed for powerful reasoning, agentic tasks, and versatile developer use cases. It comprises two models: a big one with 117B parameters (gpt-oss-120b), and a smaller one with 21B parameters (gpt-oss-20b). Both are mixture-of-experts (MoEs) and use a 4-bit quantization scheme (MXFP4), enabling fast inference (thanks to fewer active parameters, see details below) while keeping resource usage low. The large model fits on a single H100 GPU, while the small one runs within 16GB of memory and is perfect for consumer hardware and on-device applications.

Overview of Capabilities and Architecture

21B and 117B total parameters, with 3.6B and 5.1B active parameters, respectively.

4-bit quantization scheme using mxfp4 format. Only applied on the MoE weights. As stated, the 120B fits in a single 80 GB GPU and the 20B fits in a single 16GB GPU.

Reasoning, text-only models; with chain-of-thought and adjustable reasoning effort levels.

Instruction following and tool use support.

Inference implementations using transformers, vLLM, llama.cpp, and ollama.

Responses API is recommended for inference.

License: Apache 2.0, with a small complementary use policy.

Architecture

Token-choice MoE with SwiGLU activations.

When calculating the MoE weights, a softmax is taken over selected experts (softmax-after-topk).

Each attention layer uses RoPE with 128K context.

Alternate attention layers: full-context, and sliding 128-token window.

Attention layers use a learned attention sink per-head, where the denominator of the softmax has an additional additive value.

It uses the same tokenizer as GPT-4o and other OpenAI API models.

Some new tokens have been incorporated to enable compatibility with the Responses API.

The following snippet shows simple inference with the 20B model. It runs on 16 GB GPUs when using mxfp4, or ~48 GB in bfloat16.
from transformers import AutoModelForCausalLM, AutoTokenizer
model_id = "openai/gpt-oss-20b"
tokenizer = AutoTokenizer.from_pretrained(model_id)
model = AutoModelForCausalLM.from_pretrained(
model_id,
device_map="auto",
torch_dtype="auto",
)
messages = [
{"role": "user", "content": "How many rs are in the word 'strawberry'?"},
]
inputs = tokenizer.apply_chat_template(
messages,
</tr></table>

... (truncated)

Commits

06f8004 Release: v4.55.0
c54203a gpt_oss last chat template changes (#39925)
7c38d8f Add GPT OSS model from OpenAI (#39923)
738c1a3 🌐 [i18n-KO] Translated cache_explanation.md to Korean (#39535)
d2ae766 Export SmolvLM (#39614)
c430047 [docs] update object detection guide (#39909)
dedcbd6 run model debugging with forward arg (#39905)
20ce210 Revert "remove dtensors, not explicit (#39840)" (#39912)
2589a52 Fix aria tests (#39879)
6e4a9a5 Fix eval thread fork bomb (#39717)
Additional commits viewable in compare view

Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting @dependabot rebase.

Dependabot commands and options

You can trigger Dependabot actions by commenting on this PR:

@dependabot rebase will rebase this PR
@dependabot recreate will recreate this PR, overwriting any edits that have been made to it
@dependabot merge will merge this PR after your CI passes on it
@dependabot squash and merge will squash and merge this PR after your CI passes on it
@dependabot cancel merge will cancel a previously requested merge and block automerging
@dependabot reopen will reopen this PR if it is closed
@dependabot close will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually
@dependabot show <dependency name> ignore conditions will show all of the ignore conditions of the specified dependency
@dependabot ignore this major version will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself)
@dependabot ignore this minor version will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself)
@dependabot ignore this dependency will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself)

Bumps [transformers](https://github.com/huggingface/transformers) from 4.53.2 to 4.55.0. - [Release notes](https://github.com/huggingface/transformers/releases) - [Commits](huggingface/transformers@v4.53.2...v4.55.0) --- updated-dependencies: - dependency-name: transformers dependency-version: 4.55.0 dependency-type: direct:production update-type: version-update:semver-minor ... Signed-off-by: dependabot[bot] <support@github.com>

sonarqubecloud · 2025-08-11T19:27:32Z

Quality Gate passed

Issues
0 New issues
0 Accepted issues

Measures
0 Security Hotspots
0.0% Coverage on New Code
0.0% Duplication on New Code

See analysis details on SonarQube Cloud

coveralls · 2025-08-11T19:31:36Z

coverage: 52.919% (+0.04%) from 52.88%
when pulling 41d6c27 on dependabot/pip/transformers-4.55.0
into 3158296 on dev.

dependabot · 2025-08-12T10:57:49Z

OK, I won't notify you again about this release, but will get in touch when a new version is available. If you'd rather skip all updates until the next major or minor version, let me know by commenting @dependabot ignore this major version or @dependabot ignore this minor version. You can also ignore all major, minor, or patch releases for a dependency by adding an ignore condition with the desired update_types to your config file.

If you change your mind, just re-open this PR and I'll resolve any conflicts on it.

dependabot bot added dependencies Pull requests that update a dependency file python Pull requests that update Python code labels Aug 11, 2025

dependabot bot mentioned this pull request Aug 11, 2025

Bump transformers from 4.53.2 to 4.54.1 #1133

Closed

bact closed this Aug 12, 2025

dependabot bot deleted the dependabot/pip/transformers-4.55.0 branch August 12, 2025 10:57

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Uh oh!

Bump transformers from 4.53.2 to 4.55.0 #1136

Bump transformers from 4.53.2 to 4.55.0 #1136

dependabot bot commented on behalf of github Aug 11, 2025

Uh oh!

sonarqubecloud bot commented Aug 11, 2025

Uh oh!

coveralls commented Aug 11, 2025

Uh oh!

dependabot bot commented on behalf of github Aug 12, 2025

Uh oh!

Reviewers

Assignees

Labels

Projects

Milestone

Development

Uh oh!

3 participants

Bump transformers from 4.53.2 to 4.55.0 #1136

Bump transformers from 4.53.2 to 4.55.0 #1136

Conversation

dependabot bot commented on behalf of github Aug 11, 2025

v4.55.0: New openai GPT OSS model!

Welcome GPT OSS, the new open-source model family from OpenAI!

Overview of Capabilities and Architecture

Architecture

Uh oh!

sonarqubecloud bot commented Aug 11, 2025

Quality Gate passed

Uh oh!

coveralls commented Aug 11, 2025

Uh oh!

dependabot bot commented on behalf of github Aug 12, 2025

Uh oh!

Reviewers

Assignees

Labels

Projects

Milestone

Development

Uh oh!

3 participants