Transformers.js V4: Native WebGPU EP, repo restructuring, and more! #1382

xenova · 2025-07-31T03:16:21Z

No description provided.

* ONNX Runtime improvements (experimental native webgpu; fix iOS) (#1231) * customize the wasm paths * update implementation * allow using 'webgpu' in nodejs binding * update version of onnxruntime-node * Upgrade onnxruntime-web to same version as onnxruntime-node * Update list of supported devices --------- Co-authored-by: Joshua Lochner <[email protected]> * customize the wasm paths (#1250) * customize the wasm paths * update implementation * [internal] Add is_decoder option to session retrieval for preferred output location * Update tests * Formatting * Bump ort versions * Bump onnxruntime-node version * Bump versions * Bump ORT versions * Bump versions * Only check webgpu fp16 for non-node environments * Fix * Assume node supports webgpu * Update ORT node support comment * Relax test strictness * Update conversion script versions * Downgrade onnxslim * cleanup * Update package-lock.json * Update onnxruntime versions * Update post-build script * Use built-in session release function * Call garbage collection after each tokenizer test * Do not double-throw error * Fix race-condition in build process with file removal * Update versions * Bump jinja version * [version] Update to 3.6.3 * Bump jinja version to support new features * [version] Update to 3.6.3 * Add support for LFM2 models (#1367) * Use prefix in lfm2 output location (#1369) * Update package-lock.json * Run `npm audit fix` * Add special tokens in text-generation pipeline if tokenizer requires (#1370) * Add special tokens in text-generation pipeline if tokenizer requires * Fix logits processors tests * Update bundles.test.js * Update comment * Formatting * Add support for ModernBERT Decoder (#1371) * Use from/to buffer instead of string Actually fixes #1343 * Add support for Voxtral (#1373) * Support longform voxtral processing (#1375) * [version] Update to 3.7.0 * Add support for Arcee (#1377) * Optimize tensor.slice() (#1381) * Optimize tensor.slice() The performance of executing `tensor.slice()` is super poor, especially for the 'logits' tensor with large dimensions. ``` const logits = outputs.logits.slice(null, -1, null);` ``` This is because currently implementation of the `slice` method manually iterates through each element and calculate indices which is a big time consuming if the tensor shape is large. For cases like `slice(null, -1, null)`, where the slicing operation is contiguous along certain dimensions, which can be optimized by bulk copy by using `TypeArray.subarray()` and `TypeArray.set()`. * nit * Add a few more tensor slice unit tests --------- Co-authored-by: Joshua Lochner <[email protected]> --------- Co-authored-by: Yulong Wang <[email protected]> Co-authored-by: Wanming Lin <[email protected]>

xenova and others added 30 commits December 23, 2024 14:10

Move adaptive retrieval demo

1f1c645

Move code completion demo

363aefa

Move florence-2 demo

6c96600

Move depth anything demo

c4b3a76

Move tokenizer playground demo

29bdc5b

Move node audio processing demo

98958ff

Move remove background web demo

ac5b758

Move webgpu whisper demo

b3ac89e

Move depth estimation video demo

48b46d6

Move vanilla js demo

a361b38

Move node demo

1634399

Move whisper word timestamps demo

4f7b7f4

Move musicgen web demo

b0ac200

Move cross encoder demo

4ccb405

Move text-to-speech client demo

bc006e5

Move video object detection demo

7cdea27

Move video background removal demo

b053db2

Move Segment Anything demo

eacdbd1

Move zero-shot classification demo

e890716

Move browser extension template

1ee3d46

Move semantic image search demo

56297b0

Move semantic audio search demo

97a3c48

Move webgpu embedding benchmark demo

0bf2fec

Merge branch 'main' into move-examples

c033537

Merge branch 'main' into move-examples

e20167a

Merge branch 'main' into v4

140f106

Upgrade sharp version

bf12225

Bump versions

5e2c942

Delete gh-pages.yml

9f9f6f2

xenova added 19 commits July 30, 2025 18:14

Move next-server and next-client examples

ee2f0b4

Move react-translator demo

c755653

Remove old demo site example

3913313

Move webgpu-chat demo to phi-3.5-webgpu

08b3391

Remove old server-side semantic image search demo

5ecbb85

Move electron example app

0cfda75

Move webgpu clip demo app

18d675f

Replace webgpu-vlm demo

e8151c5

Folder no longer exists

64a0836

Remove examples table, replacing with link to new examples repo

5979250

Rename files

3204661

Merge branch 'main' into move-examples

32017ae

Merge branch 'move-examples' into v4

a2b00c0

Update .prettierrc

0c76a03

Apply to js, mjs, and cjs

05df971

Format source code

2a32dc0

Update .prettierrc

276f3e7

Formatting

3c7cf47

Fix type issues

3b4f4c4

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Uh oh!

Transformers.js V4: Native WebGPU EP, repo restructuring, and more! #1382

Transformers.js V4: Native WebGPU EP, repo restructuring, and more! #1382

Uh oh!

xenova commented Jul 31, 2025

Uh oh!

Uh oh!

Transformers.js V4: Native WebGPU EP, repo restructuring, and more! #1382

Are you sure you want to change the base?

Transformers.js V4: Native WebGPU EP, repo restructuring, and more! #1382

Uh oh!

Conversation

xenova commented Jul 31, 2025

Uh oh!

Uh oh!