Repository navigation
fix(deps): update dependency transformers to v5.18.0 - #48
Open
renovate[bot] wants to merge 1 commit into
Open
renovate[bot] wants to merge 1 commit into
renovate[bot] wants to merge 1 commit into
Conversation
This branch has not been deployed
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
This PR contains the following updates:
==5.17.0→==5.18.0Release Notes
huggingface/transformers (transformers)
v5.18.0: Release 5.18.0Compare Source
New Model additions
Nemotron 3 Diarization
Nemotron 3 Diarization is an open-weight streaming speaker diarization model designed to determine "who spoke when" in real-world audio. It supports both streaming and offline inference, handles up to eight speakers, and orders speaker outputs by each speaker's first arrival in the input audio.
The model uses the Arrival-Order Speaker Cache (AOSC) 1 and FIFO queue introduced for Streaming Sortformer 1, 2. A single checkpoint supports configurable latency profiles, from an 80 ms input buffer to a 30.4 s offline-style buffer, and configurable output frame resolution in multiples of 10 ms. With chunked inference, the maximum audio duration is not limited.
Links: Documentation
NemotronH Omni
NemotronH Omni is a multimodal reasoning model from NVIDIA that pairs the NemotronH hybrid
Mamba-Transformer language model with a RADIO vision encoder and an optional Parakeet-based sound encoder.
Image (and video) patches are projected through a RADIO tower and a pixel-shuffle MLP into the language model's
embedding space at the
<image>/<video>context-token positions; audio clips are projected in the same way at<audio>positions. The result is a single autoregressive model that reasons jointly over text, images, video andsound.
Links: Documentation
HyperCLOVAX Vision V2
HyperCLOVAX Vision V2 is a multimodal vision-language model developed by NAVER. It combines the HyperClovaX language model backbone with a Qwen2.5-VL vision encoder. The model supports text, image, and video inputs and is capable of chain-of-thought reasoning via built-in thinking tokens (
<think>...</think>).Links: Documentation
GTE
GTE was proposed in mGTE: Generalized Long-Context Text Representation and Reranking Models for Multilingual Text Retrieval by Xin Zhang, Yanzhao Zhang, Dingkun Long, Wen Xie, Ziqi Dai, Jialong Tang, Huan Lin, Baosong Yang, Pengjun Xie, Fei Huang, Meishan Zhang, Wenjie Li and Min Zhang.
GTE is a BERT-style bidirectional encoder that replaces absolute position embeddings with RoPE, uses a gated MLP, and applies layer normalization after each residual connection. The same architecture backs Alibaba's
gte-*-v1.5,gte-multilingual-*andgte-en-mlm-*checkpoints as well as Snowflake'ssnowflake-arctic-embed-m-v2.0.Links: Documentation
Breaking changes
Kernels] Bump version (#48714) by @vasquBugfixes and improvements
AutoImageProcessorrequiring torchvision when only Pillow is installed (#48616) by @blipbyteMoE] Fix eager EP (#48653) by @vasquminimaxfailing withoutput_mismatch(tensor values differ (2)) (#48515) by @sergereview[bot]mistralfailing withother(other (2)) (#48429) by @sergereview[bot]MemoryCleanupMixinclass for tests (#48681) by @tarekziadeflex_olmofailing withother(other (1)) (#48668) by @sergereview[bot]RTDetrModel/SEWDForCTCloads (wrongbase_model_prefix) (#48744) by @peftversion requirement (#48716) by @shniuboboedgetamfailing withimport_or_config(other (12)) (#48322) by @sergereview[bot]AutoModel.from_pretrainednot restoringmodules_to_saveweights (#48595) by @shniuboboreseton the dynamic cache layers (#48809) by @jiqing-fengStaticCachefor Mllama and enabletorch.compile(#48141) by @jiqing-fengfeat] Allow untyinghidden_states[-1]fromlast_hidden_statevia the model config (#48087) by @tomaarsendeepseek_vlfailing withother(other (3)) (#48536) by @sergereview[bot]DSA] Only save latents on dsa with indexer as well (#48876) by @vasquprefix_allowed_tokens_fnoverride model-infand raise an exception on unsatisfiable generation constraints. (#48927) by @ksh108405update_candidate_strategy(#48982) by @Cyrilvalleztorch.compile(#48975) by @jiqing-fengconfig.output_router_logitsin the MoE VLM wrappers (#48885) by @qgallouedecattention_maskasNonein OPT's causal mask creation (#49002) by @jiqing-fengTrainer.end(#48875) by @qgallouedeccvtfailing withoutput_mismatch(tensor values differ (2)) (#49076) by @sergereview[bot]hy_v3failing withoutput_mismatch(tensor values differ (2)) (#49068) by @sergereview[bot]pvt_v2failing withother(#49067) by @sergereview[bot]jambafailing withother(other (2)) (#49044) by @sergereview[bot]auto_find_batch_size(#49198) by @qgallouedecKernels] Sync mamba version (#49205) by @vasquSignificant community contributions
The following contributors have made significant changes to the library over the last release:
MemoryCleanupMixinclass for tests (#48681)Configuration
📅 Schedule: (UTC)
🚦 Automerge: Disabled by config. Please merge this manually once you are satisfied.
♻ Rebasing: Whenever PR is behind base branch, or you tick the rebase/retry checkbox.
🔕 Ignore: Close this PR and you won't be reminded about this update again.
This PR was generated by Mend Renovate. View the repository job log.