[ExecuTorch][WebGPU] Add Gemma 4 MTP operator and route support - #21675
[ExecuTorch][WebGPU] Add Gemma 4 MTP operator and route support#21675JCNTH wants to merge 1 commit into
Conversation
🔗 Helpful Links🧪 See artifacts and rendered test results at hud.pytorch.org/pr/pytorch/executorch/21675
Note: Links to docs will display an error until the docs builds have been completed. ❌ 2 New Failures, 1 PendingAs of commit 60a9937 with merge base ceca90f ( NEW FAILURES - The following jobs have failed:
This comment was automatically generated by Dr. CI and updates every 15 minutes. |
This PR needs a
|
|
Hi @JCNTH! Thank you for your pull request. We require contributors to sign our Contributor License Agreement, and yours needs attention. You currently have a record in our system, but the CLA is no longer valid, and will need to be resubmitted. ProcessIn order for us to review and merge your suggested changes, please sign at https://code.facebook.com/cla. If you are contributing on behalf of someone else (eg your employer), the individual CLA may not be sufficient and your employer may need to sign the corporate CLA. Once the CLA is signed, our tooling will perform checks and validations. Afterwards, the pull request will be tagged with If you have received this in error or have any questions, please contact us at cla@meta.com. Thanks! |
Stack from ghstack (oldest at bottom):
Add the backend-only operator surface required by Gemma 4 MTP
k2_round. The diff introduces exact-shape staged TopK, duplicate-safe generic Scatter, a provenance-certified unique-index WG64 Scatter route, long-context 2D dispatch closure, integer subtraction metadata, and a guarded raw-fp32 M=3 Q4 path.Key changes:
[1,1,2048]to[1,1,32]values/indices with WG64 metadata.aten.scatter.srcduplicate-safe and exposes parallel execution only through certifiedet_vk.scatter_src_unique.default.MTP delegation is instance-scoped; plain Gemma exposes neither TopK nor certified Scatter.
Co-authored-with: Claude Code.
Differential Revision: D115234084