Loading…
Implement Metal/MPS kernels for PyTorch operators.
This skill enables developers to write Metal/MPS kernels specifically for PyTorch operators, facilitating the addition of MPS device support, the implementation of Metal shaders, and the migration of CUDA kernels to Apple Silicon. It covers the necessary steps for updating dispatch in `native_functions.yaml`, implementing Metal kernels, and creating host-side operators, ensuring efficient and maintainable code for Apple devices.