Skip to content

[js/webgpu] enable f16 for concat - #18528

Merged
Yulong Wang (fs-eire) merged 1 commit into
microsoft:mainfrom
qjia7:concat_f16
Nov 21, 2023
Merged

[js/webgpu] enable f16 for concat#18528
Yulong Wang (fs-eire) merged 1 commit into
microsoft:mainfrom
qjia7:concat_f16

Conversation

@qjia7

Copy link
Copy Markdown
Contributor

Description

With this PR realesrgan-t64-f16 models becomes 32.8 ms from 1052.55 ms. Now the whole model run on jsep.

@qjia7

Copy link
Copy Markdown
Contributor Author

@satyajandhyala Satya Kumar Jandhyala (satyajandhyala) added the ep:WebGPU ort-web webgpu provider label Nov 21, 2023
@fs-eire

Copy link
Copy Markdown
Contributor

/azp run Windows ARM64 QNN CI Pipeline,Windows x64 QNN CI Pipeline,Windows CPU CI Pipeline,Windows GPU CI Pipeline,Windows GPU TensorRT CI Pipeline,ONNX Runtime Web CI Pipeline,Linux CPU CI Pipeline,Linux CPU Minimal Build E2E CI Pipeline,Linux GPU CI Pipeline,Linux GPU TensorRT CI Pipeline

@fs-eire

Copy link
Copy Markdown
Contributor

/azp run Linux OpenVINO CI Pipeline,Linux QNN CI Pipeline,MacOS CI Pipeline,orttraining-amd-gpu-ci-pipeline,orttraining-linux-ci-pipeline,orttraining-linux-gpu-ci-pipeline,orttraining-ortmodule-distributed,onnxruntime-python-checks-ci-pipeline,onnxruntime-binary-size-checks-ci-pipeline,Android CI Pipeline

@fs-eire

Copy link
Copy Markdown
Contributor

/azp run iOS CI Pipeline,ONNX Runtime React Native CI Pipeline

@azure-pipelines

Copy link
Copy Markdown
Azure Pipelines successfully started running 1 pipeline(s).

2 similar comments
@azure-pipelines

Copy link
Copy Markdown
Azure Pipelines successfully started running 1 pipeline(s).

@azure-pipelines

Copy link
Copy Markdown
Azure Pipelines successfully started running 1 pipeline(s).

@fs-eire
Yulong Wang (fs-eire) merged commit ac8598a into microsoft:main Nov 21, 2023
@qjia7
Jiajia Qin (qjia7) deleted the concat_f16 branch November 23, 2023 02:12
kleiti (kleiti) pushed a commit to kleiti/onnxruntime that referenced this pull request Mar 22, 2024
### Description
With this PR `realesrgan-t64-f16` models becomes 32.8 ms from 1052.55
ms. Now the whole model run on jsep.
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

ep:WebGPU ort-web webgpu provider

Projects

None yet

Development

Successfully merging this pull request may close these issues.

3 participants