PyTorch 2.13 brings FlexAttention to Apple Silicon, cuts peak memory by up to 4× for large-vocabulary models with nn.LinearCrossEntropyLoss, and updates distributed training, compilation, profiling, and on-device inference.
Our live Q&A examined CUDA version support, CuTeDSL in
Tensors and neural networks in Python with strong hardware acceleration. PyTorch is an open source project at the Linux Foundation. #PyTorchFoundation
Joined September 2016





