Autotuning in PyTorch & Triton
May 4, 2025
torch.compile offers some knobs for controlling the trade-off of execution performance with longer compile times. This is particularly useful for inference, where the same model will be running for a long time.
