[WebNN EP] Automatically move input CPU tensors to ml-tensor #23073

egalli · 2024-12-11T01:31:30Z

Description

If it would improve performance, this patch moves the CPU to ml-tensor before sending the to the ONNXRuntime WebNN EP.

Motivation and Context

We are currently performing 2 extra copies on input tensors located in the CPU when using the WebNN EP (JS -(copy)-> wasm heap -(copy)-> JS -> WebNN API). This patch removes these extra copies.

### Description If it would improve performance, this patch moves the CPU to ml-tensor before sending the to the ONNXRuntime WebNN EP. ### Motivation and Context We are currently performing 2 extra copies on input tensors located in the CPU when using the WebNN EP (JS -(copy)-> wasm heap -(copy)-> JS -> WebNN API). This patch removes these extra copies.

fdwr

Yulong or Guenther can review this one effectively (as it impacts TypeScript interfaces). I'll start the CI's though.

fdwr · 2024-12-11T03:26:59Z

/azp run ONNX Runtime Web CI Pipeline,Windows GPU CI Pipeline,Linux Android Emulator QNN CI Pipeline

fdwr · 2024-12-11T03:27:02Z

/azp run Linux CPU CI Pipeline,Linux CPU Minimal Build E2E CI Pipeline,Linux GPU CI Pipeline,Linux GPU TensorRT CI Pipeline,Linux OpenVINO CI Pipeline,Linux QNN CI Pipeline,MacOS CI Pipeline,Windows ARM64 QNN CI Pipeline,Windows CPU CI Pipeline

fdwr · 2024-12-11T03:27:04Z

/azp run Windows GPU CUDA CI Pipeline,Windows GPU DML CI Pipeline,Windows GPU Doc Gen CI Pipeline

fdwr · 2024-12-11T03:27:07Z

/azp run Windows GPU TensorRT CI Pipeline,onnxruntime-binary-size-checks-ci-pipeline,orttraining-linux-ci-pipeline,orttraining-linux-gpu-ci-pipeline,orttraining-ortmodule-distributed,Windows x64 QNN CI Pipeline,Big Models

azure-pipelines · 2024-12-11T03:27:13Z

Azure Pipelines successfully started running 2 pipeline(s).

azure-pipelines · 2024-12-11T03:27:19Z

Azure Pipelines successfully started running 3 pipeline(s).

azure-pipelines · 2024-12-11T03:27:25Z

Azure Pipelines successfully started running 4 pipeline(s).

azure-pipelines · 2024-12-11T03:27:36Z

Azure Pipelines successfully started running 9 pipeline(s).

Honry

👍

js/web/lib/wasm/jsep/backend-webnn.ts

Honry

LGTM % a nit.

js/web/lib/wasm/jsep/backend-webnn.ts

Honry · 2024-12-12T00:46:49Z

@fs-eire, @guschmue, pls. take another look, thanks!

egalli · 2024-12-13T08:12:17Z

js/web/lib/wasm/wasm-core-impl.ts

+          if (!createTemporaryTensor || !uploadTensor) {
+            throw new Error('Tensor location "ml-tensor" is not supported without using WebNN.');
+          }
+          const tensorId = await createTemporaryTensor(dataTypeEnum, dims as number[]);


Found an issue while debugging microsoft/webnn-developer-preview#69

We can't safety use await and expect WebNNBackend.activeSessionId to be valid.

We'll need to manually pass the sessionHandle/Id to createTemporaryTensor and isGraphInput

Appears you pushed more commits related to sessionHandle. So is this comment resolveable now?

…sion

guschmue · 2024-12-17T18:00:44Z

/azp run ONNX Runtime Web CI Pipeline,Windows GPU CI Pipeline,Linux Android Emulator QNN CI Pipeline

guschmue · 2024-12-17T18:00:52Z

/azp run Linux CPU CI Pipeline,Linux CPU Minimal Build E2E CI Pipeline,Linux GPU CI Pipeline,Linux GPU TensorRT CI Pipeline,Linux OpenVINO CI Pipeline,Linux QNN CI Pipeline,MacOS CI Pipeline,Windows ARM64 QNN CI Pipeline,Windows CPU CI Pipeline

guschmue · 2024-12-17T18:01:00Z

/azp run Windows GPU TensorRT CI Pipeline,onnxruntime-binary-size-checks-ci-pipeline,orttraining-linux-ci-pipeline,orttraining-linux-gpu-ci-pipeline,orttraining-ortmodule-distributed,Windows x64 QNN CI Pipeline,Big Models

azure-pipelines · 2024-12-17T18:01:01Z

Azure Pipelines successfully started running 2 pipeline(s).

guschmue · 2024-12-17T18:01:07Z

/azp run Windows GPU CUDA CI Pipeline,Windows GPU DML CI Pipeline,Windows GPU Doc Gen CI Pipeline

azure-pipelines · 2024-12-17T18:01:22Z

Azure Pipelines successfully started running 3 pipeline(s).

azure-pipelines · 2024-12-17T18:01:23Z

Azure Pipelines successfully started running 4 pipeline(s).

azure-pipelines · 2024-12-17T18:01:34Z

Azure Pipelines successfully started running 9 pipeline(s).

fdwr reviewed Dec 11, 2024

View reviewed changes

Honry reviewed Dec 11, 2024

View reviewed changes

js/web/lib/wasm/jsep/backend-webnn.ts Outdated Show resolved Hide resolved

js/web/lib/wasm/jsep/backend-webnn.ts Outdated Show resolved Hide resolved

PR feedback

be01b60

Honry approved these changes Dec 12, 2024

View reviewed changes

js/web/lib/wasm/jsep/backend-webnn.ts Outdated Show resolved Hide resolved

More renames from tensor(s) to tensorId(s)

5e3295f

egalli commented Dec 13, 2024

View reviewed changes

egalli added 2 commits December 13, 2024 14:07

Merge remote-tracking branch 'origin/main' into promote_inputs

21edcaf

Pass sessionHandle/Id directly to function instead of using activeSes…

d66258f

…sion

guschmue added the ep:WebNN WebNN execution provider label Dec 16, 2024

fdwr requested a review from fs-eire December 18, 2024 22:13

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

[WebNN EP] Automatically move input CPU tensors to ml-tensor #23073

[WebNN EP] Automatically move input CPU tensors to ml-tensor #23073

egalli commented Dec 11, 2024

fdwr left a comment

fdwr commented Dec 11, 2024

fdwr commented Dec 11, 2024

fdwr commented Dec 11, 2024

fdwr commented Dec 11, 2024

azure-pipelines bot commented Dec 11, 2024

azure-pipelines bot commented Dec 11, 2024

azure-pipelines bot commented Dec 11, 2024

azure-pipelines bot commented Dec 11, 2024

Honry left a comment

Honry left a comment

Honry commented Dec 12, 2024

egalli Dec 13, 2024

fdwr Dec 18, 2024

guschmue commented Dec 17, 2024

guschmue commented Dec 17, 2024

guschmue commented Dec 17, 2024

azure-pipelines bot commented Dec 17, 2024

guschmue commented Dec 17, 2024

azure-pipelines bot commented Dec 17, 2024

azure-pipelines bot commented Dec 17, 2024

azure-pipelines bot commented Dec 17, 2024

[WebNN EP] Automatically move input CPU tensors to ml-tensor #23073

Are you sure you want to change the base?

[WebNN EP] Automatically move input CPU tensors to ml-tensor #23073

Conversation

egalli commented Dec 11, 2024

Description

Motivation and Context

fdwr left a comment

Choose a reason for hiding this comment

fdwr commented Dec 11, 2024

fdwr commented Dec 11, 2024

fdwr commented Dec 11, 2024

fdwr commented Dec 11, 2024

azure-pipelines bot commented Dec 11, 2024

azure-pipelines bot commented Dec 11, 2024

azure-pipelines bot commented Dec 11, 2024

azure-pipelines bot commented Dec 11, 2024

Honry left a comment

Choose a reason for hiding this comment

Honry left a comment

Choose a reason for hiding this comment

Honry commented Dec 12, 2024

egalli Dec 13, 2024

Choose a reason for hiding this comment

fdwr Dec 18, 2024

Choose a reason for hiding this comment

guschmue commented Dec 17, 2024

guschmue commented Dec 17, 2024

guschmue commented Dec 17, 2024

azure-pipelines bot commented Dec 17, 2024

guschmue commented Dec 17, 2024

azure-pipelines bot commented Dec 17, 2024

azure-pipelines bot commented Dec 17, 2024

azure-pipelines bot commented Dec 17, 2024