Torchaudio Documentation

Torchaudio is a library for audio and signal processing with PyTorch. It provides I/O, signal and data processing functions, datasets, model implementations and application components.

Tutorials

Loading waveform Tensors from files and saving them

Topics: I/O

Learn how to query/load audio files and save waveform tensors to files, using torchaudio.info, torchaudio.load and torchaudio.save functions.

Streaming media decoding with StreamReader

Topics: I/O, StreamReader

Learn how to load audio/video to Tensors using torchaudio.io.StreamReader class.

Device input, synthetic audio/video, and filtering with StreamReader

Topics: I/O, StreamReader

Learn how to load media from hardware devices, generate synthetic audio/video, and apply filters to them with torchaudio.io.StreamReader.

Streaming media encoding with StreamWriter

Topics: I/O, StreamWriter

Learn how to save audio/video with torchaudio.io.StreamWriter.

Playing media with StreamWriter

Topics: I/O, StreamWriter

Learn how to play audio/video with torchaudio.io.StreamWriter.

Hardware accelerated video I/O with NVDEC/NVENC

Topics: I/O, StreamReader, StreamWriter

Learn how to setup and use HW accelerated video I/O.

Citing torchaudio

If you find torchaudio useful, please cite the following paper:

Yang, Y.-Y., Hira, M., Ni, Z., Chourdia, A., Astafurov, A., Chen, C., Yeh, C.-F., Puhrsch, C., Pollack, D., Genzel, D., Greenberg, D., Yang, E. Z., Lian, J., Mahadeokar, J., Hwang, J., Chen, J., Goldsborough, P., Roy, P., Narenthiran, S., Watanabe, S., Chintala, S., Quenneville-Bélair, V, & Shi, Y. (2021). TorchAudio: Building Blocks for Audio and Speech Processing. arXiv preprint arXiv:2110.15018.

In BibTeX format:

@article{yang2021torchaudio,
  title={TorchAudio: Building Blocks for Audio and Speech Processing},
  author={Yao-Yuan Yang and Moto Hira and Zhaoheng Ni and
          Anjali Chourdia and Artyom Astafurov and Caroline Chen and
          Ching-Feng Yeh and Christian Puhrsch and David Pollack and
          Dmitriy Genzel and Donny Greenberg and Edward Z. Yang and
          Jason Lian and Jay Mahadeokar and Jeff Hwang and Ji Chen and
          Peter Goldsborough and Prabhat Roy and Sean Narenthiran and
          Shinji Watanabe and Soumith Chintala and
          Vincent Quenneville-Bélair and Yangyang Shi},
  journal={arXiv preprint arXiv:2110.15018},
  year={2021}
}

Torchaudio Documentation

Tutorials

Loading waveform Tensors from files and saving them

Streaming media decoding with StreamReader

Device input, synthetic audio/video, and filtering with StreamReader

Streaming media encoding with StreamWriter

Playing media with StreamWriter

Hardware accelerated video I/O with NVDEC/NVENC

Citing torchaudio

Docs

Tutorials

Resources