News
PyTorch 2.13: FlexAttention on Apple Silicon and 4x LLM Memory Savings
PyTorch 2.13 brings FlexAttention to Apple Silicon with 12x sparse attention speedup and a fused ...




