PyTorch 2.13: FlexAttention on Apple Silicon Is 12x Faster
PyTorch 2.13 brings FlexAttention to Apple Silicon with hand-written Metal kernels, delivering up to 12x speedup over SDPA on sparse attention patterns. Here ...
Latest tech industry news, product launches, and company announcements