mirror of
https://github.com/priyanshujain/torchmlx.git
synced 2026-10-02 11:07:13 +00:00
build torchmlx
This commit is contained in:
commit
b0ef50eec2
15 files changed
+2008
No files matched your search
@@ -0,0 +1,13 @@
|
||||
# Compatibility
|
||||
|
||||
TorchMLX targets the common transformer operations used by GPT-2, Llama 3, Qwen 3, and GPT-OSS style implementations.
|
||||
|
||||
Supported MLX operations include embeddings, linear layers, normalization building blocks, dropout, activations, causal attention, tensor shape operations, masks, top-k routing, and AdamW training through `Trainer`.
|
||||
|
||||
MLX arrays remain native arrays. Torch-style tensor methods are installed on the native array type for the supported subset.
|
||||
|
||||
Boolean expert routing and `unique` execute eagerly because their output shapes control Python flow.
|
||||
|
||||
Set `TORCHMLX_BACKEND=torch` before import to use native PyTorch for unsupported programs. TorchMLX never changes backend during an operation.
|
||||
|
||||
The referenced OpenArch model files contain source errors independent of TorchMLX, including invalid constructor calls and undefined attributes. Correct those errors before using either backend.
|
||||
Reference in new issue
Block a user