ExecuTorch MLX delegate by metascroy · Pull Request #16718 · pytorch/executorch

metascroy · 2026-01-20T22:10:33Z

Summary

This PR adds an MLX backend for ExecuTorch, enabling Metal-accelerated inference on Apple Silicon. It runs Llama, Qwen, Gemma, Whisper, Voxtral, and Parakeet models end-to-end, with 637 passing op tests and multithreaded execution support. For many models, it offers best performance among all ExecuTorch backends on Apple Silicon, offering 2-6x speedups over what was previously possible with ExecuTorch, and up to 30% smaller model sizes compared to XNNPACK due to BF16 support and tied quantized embedding support.

The PR is large due to extensive op coverage, testing, and documentation, but almost all changes are confined to backends/mlx/. The design is described in backends/mlx/README.md.

Suggested review approach:

Review files outside backends/mlx/ carefully — these integrate with ExecuTorch's build system and are the most likely to need changes.
For backends/mlx/, focus on structural design (see README) and test coverage (CI job is .github/workflows/mlx.yml)

Prerequisite PRs

These fixes were developed alongside the MLX backend. Once merged, this PR can be rebased to remove the duplicated changes:

#17257 — Improve lowering time with NamedDataMap
#17679 — Allow transform passes in etLLM
#17678 — Fix dynamic shape bug in remove_noop_pass
#17378 — Fix pocketfft intermittent bus errors on macOS (upstream fix)

Tests

CI is defined in .github/workflows/mlx.yml:

test_ops.py: 637 passing op tests
Multithreading: launches models on 50 threads, verifies correctness
GenAI E2E: parakeet, voxtral, etLLM (stories110m), HF LLMs (llama, qwen, gemma)
backend-tester: 380 passed, 1 failed, 86 skipped operator tests; 34 passed, 3 failed, 5 skipped model tests

The 1 failing operator test is a test-side issue being fixed in #17539. The 3 failing model tests will be addressed in follow-ups — they are not an initial focus compared to the GenAI models above.

pytorch-bot · 2026-01-20T22:10:38Z

🔗 Helpful Links

🧪 See artifacts and rendered test results at hud.pytorch.org/pr/pytorch/executorch/16718

📄 Preview Python docs built from this PR

Note: Links to docs will display an error until the docs builds have been completed.

❌ 2 Cancelled Jobs

As of commit 240b241 with merge base 3319157 ():

CANCELLED JOBS - The following jobs were cancelled. Please retry:

Test ARM Backend / test-arm / test-backend-linux (arm_tosa_fp, models) / linux-job (gh)
##[error]The operation was canceled.
Test ARM Backend / test-arm / test-backend-linux (arm_vgf_fp, models) / linux-job (gh)
##[error]The operation was canceled.

This comment was automatically generated by Dr. CI and updates every 15 minutes.

github-actions · 2026-01-20T22:11:20Z

This PR needs a `release notes:` label

If your change should be included in the release notes (i.e. would users of this library care about this change?), please use a label starting with release notes:. This helps us keep track and include your important work in the next release notes.

To add a label, you can comment to pytorchbot, for example
@pytorchbot label "release notes: none"

For more information, see
https://github.com/pytorch/pytorch/wiki/PyTorch-AutoLabel-Bot#why-categorize-for-release-notes-and-how-does-it-work.

metascroy requested review from cccclai, kirklandsign, larryliu0820 and shoumikhin as code owners January 20, 2026 22:10

meta-cla bot added the CLA Signed This label is managed by the Facebook bot. Authors need to sign the CLA before a PR can be reviewed. label Jan 20, 2026

mergennachin requested review from manuelcandales and mergennachin January 20, 2026 22:37

metascroy force-pushed the mlx-delegate branch from 5e84aa5 to 1867cfc Compare January 27, 2026 01:58

metascroy requested review from JacobSzwejbka and lucylq as code owners January 30, 2026 06:33

metascroy force-pushed the mlx-delegate branch from 6541f23 to f296724 Compare February 9, 2026 04:41

metascroy force-pushed the mlx-delegate branch 6 times, most recently from 5eef84c to 6e41f05 Compare February 24, 2026 07:28

metascroy and others added 11 commits February 24, 2026 16:34

MLX delegate

403bdc9

backends/apple/mlx/patches/mlx_json.patch

1602189

Exclude auto-generated files from git

6372e7f

up

54981f7

more testing

f832757

up

d158e52

up

913f129

up

8de6b93

parakeet

137892c

up

bb3dd35

up

365158d

metascroy and others added 28 commits February 24, 2026 16:34

up

a8db934

up

0720468

up

4c2448f

up

e1c0f6d

up

37a674f

up

0b9a664

up

f0c02a9

up

709bf9f

up

8e2f9c8

up

a3ceab3

up

56c3196

up

f17a4a7

up

ed4d45c

up

76adc1d

up

c7c22c9

up

da81245

up

72560af

up

7591f32

up

3b0473e

up

2541ca5

up

8609f0e

up

3ed8760

up

303a1ce

up

feab5bb

Fix intermittent bus errors in portable pocketfft op

fb43eb4

Create pocketfft_aligned_alloc.patch

ab15e4f

up

ffeeafc

up

240b241

metascroy force-pushed the mlx-delegate branch from e0e015c to 240b241 Compare February 25, 2026 01:06

metascroy changed the title ~~[draft] MLX delegate~~ ExecuTorch MLX delegate Feb 25, 2026

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

ExecuTorch MLX delegate#16718

ExecuTorch MLX delegate#16718
metascroy wants to merge 69 commits intomainfrom
mlx-delegate

metascroy commented Jan 20, 2026 •

edited

Loading

Uh oh!

pytorch-bot bot commented Jan 20, 2026 •

edited

Loading

Uh oh!

github-actions bot commented Jan 20, 2026

Uh oh!

Reviewers

Assignees

Labels

Projects

Milestone

Development

Uh oh!

1 participant

Conversation

metascroy commented Jan 20, 2026 • edited Loading Uh oh! There was an error while loading. Please reload this page.

Uh oh!

Summary

Prerequisite PRs

Tests

Uh oh!

pytorch-bot bot commented Jan 20, 2026 • edited Loading Uh oh! There was an error while loading. Please reload this page.

Uh oh!

🔗 Helpful Links

🧪 See artifacts and rendered test results at hud.pytorch.org/pr/pytorch/executorch/16718

❌ 2 Cancelled Jobs

Uh oh!

github-actions bot commented Jan 20, 2026

This PR needs a release notes: label

Uh oh!

Reviewers

Assignees

Labels

Projects

Milestone

Development

Uh oh!

1 participant

metascroy commented Jan 20, 2026 •

edited

Loading

pytorch-bot bot commented Jan 20, 2026 •

edited

Loading

This PR needs a `release notes:` label