Movatterモバイル変換

PyTorch

From Wikipedia, the free encyclopedia

Open source machine learning library

PyTorch

Original author(s)	Adam Paszke Sam Gross Soumith Chintala Gregory Chanan
Developer(s)	Meta AI
Initial release	September 2016; 8 years ago (2016-09)^[1]

Stable release	2.6.0^[2] / 29 January 2025; 52 days ago (29 January 2025)

Repository	github.com/pytorch/pytorch
Written in	Python C++ CUDA
Operating system	Linux macOS Windows
Platform	IA-32,x86-64,ARM64
Available in	English
Type	Library formachine learning anddeep learning
License	BSD-3^[3]
Website	pytorch.org

Machine learning anddata mining
Part of a series on
Paradigms Supervised learning Unsupervised learning Semi-supervised learning Self-supervised learning Reinforcement learning Meta-learning Online learning Batch learning Curriculum learning Rule-based learning Neuro-symbolic AI Neuromorphic engineering Quantum machine learning
Problems Classification Generative modeling Regression Clustering Dimensionality reduction Density estimation Anomaly detection Data cleaning AutoML Association rules Semantic analysis Structured prediction Feature engineering Feature learning Learning to rank Grammar induction Ontology learning Multimodal learning
Supervised learning (classification • regression) Apprenticeship learning Decision trees Ensembles Bagging Boosting Random forest k-NN Linear regression Naive Bayes Artificial neural networks Logistic regression Perceptron Relevance vector machine (RVM) Support vector machine (SVM)
Clustering BIRCH CURE Hierarchical k-means Fuzzy Expectation–maximization (EM) DBSCAN OPTICS Mean shift
Dimensionality reduction Factor analysis CCA ICA LDA NMF PCA PGD t-SNE SDL
Structured prediction Graphical models Bayes net Conditional random field Hidden Markov
Anomaly detection RANSAC k-NN Local outlier factor Isolation forest
Artificial neural network Autoencoder Deep learning Feedforward neural network Recurrent neural network LSTM GRU ESN reservoir computing Boltzmann machine Restricted GAN Diffusion model SOM Convolutional neural network U-Net LeNet AlexNet DeepDream Neural radiance field Transformer Vision Mamba Spiking neural network Memtransistor Electrochemical RAM (ECRAM)
Reinforcement learning Q-learning SARSA Temporal difference (TD) Multi-agent Self-play
Learning with humans Active learning Crowdsourcing Human-in-the-loop RLHF
Model diagnostics Coefficient of determination Confusion matrix Learning curve ROC curve
Mathematical foundations Kernel machines Bias–variance tradeoff Computational learning theory Empirical risk minimization Occam learning PAC learning Statistical learning VC theory Topological deep learning
Journals and conferences ECML PKDD NeurIPS ICML ICLR IJCAI ML JMLR
Related articles Glossary of artificial intelligence List of datasets for machine-learning research List of datasets in computer vision and image processing Outline of machine learning
v t e

PyTorch is amachine learning library based on theTorch library,^[4]^[5]^[6] used for applications such ascomputer vision andnatural language processing,^[7] originally developed byMeta AI and now part of theLinux Foundation umbrella.^[8]^[9]^[10]^[11] It is one of the most populardeep learning frameworks, alongside others such asTensorFlow,^[12] offeringfree and open-source software released under themodified BSD license. Although thePython interface is more polished and the primary focus of development, PyTorch also has aC++ interface.^[13]

A number of pieces ofdeep learning software are built on top of PyTorch, includingTesla Autopilot,^[14]Uber's Pyro,^[15]Hugging Face's Transformers,^[16]^[17] and Catalyst.^[18]^[19]

PyTorch provides two high-level features:^[20]

Tensor computing (likeNumPy) with strong acceleration viagraphics processing units (GPU)
Deep neural networks built on a tape-basedautomatic differentiation system

History

[edit]

Meta (formerly known as Facebook) operates both PyTorch and Convolutional Architecture for Fast Feature Embedding (Caffe2), but models defined by the two frameworks were mutually incompatible. The Open Neural Network Exchange (ONNX) project was created by Meta andMicrosoft in September 2017 for converting models between frameworks. Caffe2 was merged into PyTorch at the end of March 2018.^[21] In September 2022, Meta announced that PyTorch would be governed by the independent PyTorch Foundation, a newly created subsidiary of theLinux Foundation.^[22]

PyTorch 2.0 was released on 15 March 2023, introducingTorchDynamo, a Python-levelcompiler that makes code run up to 2x faster, along with significant improvements in training and inference performance across majorcloud platforms.^[23]^[24]

PyTorch tensors

[edit]

Main article:Tensor (machine learning)

PyTorch defines a class called Tensor (torch.Tensor) to store and operate on homogeneous multidimensional rectangular arrays of numbers. PyTorch Tensors are similar toNumPy Arrays, but can also be operated on aCUDA-capableNVIDIA GPU. PyTorch has also been developing support for other GPU platforms, for example, AMD'sROCm^[25] and Apple'sMetal Framework.^[26]

PyTorch supports various sub-types of Tensors.^[27]

Note that the term "tensor" here does not carry the same meaning as tensor in mathematics or physics. The meaning of the word in machine learning is only superficially related to its original meaning as a certain kind of object inlinear algebra. Tensors in PyTorch are simply multi-dimensional arrays.

PyTorch neural networks

[edit]

Main article:Neural network (machine learning)

PyTorch defines a module called nn (torch.nn) to describe neural networks and to support training. This module offers a comprehensive collection of building blocks for neural networks, including various layers and activation functions, enabling the construction of complex models. Networks are built by inheriting from thetorch.nn module and defining the sequence of operations in theforward() function.

Example

[edit]

The following program shows the low-level functionality of the library with a simple example.

importtorchdtype=torch.floatdevice=torch.device("cpu")# Execute all calculations on the CPU# device = torch.device("cuda:0")  # Executes all calculations on the GPU# Create a tensor and fill it with random numbersa=torch.randn(2,3,device=device,dtype=dtype)print(a)# Output: tensor([[-1.1884,  0.8498, -1.7129],#                  [-0.8816,  0.1944,  0.5847]])b=torch.randn(2,3,device=device,dtype=dtype)print(b)# Output: tensor([[ 0.7178, -0.8453, -1.3403],#                  [ 1.3262,  1.1512, -1.7070]])print(a*b)# Output: tensor([[-0.8530, -0.7183,  2.58],#                  [-1.1692,  0.2238, -0.9981]])print(a.sum())# Output: tensor(-2.1540)print(a[1,2])# Output of the element in the third column of the second row (zero based)# Output: tensor(0.5847)print(a.max())# Output: tensor(0.8498)

The following code-block defines a neural network with linear layers using thenn module.

fromtorchimportnn# Import the nn sub-module from PyTorchclassNeuralNetwork(nn.Module):# Neural networks are defined as classesdef__init__(self):# Layers and variables are defined in the __init__ methodsuper().__init__()# Must be in every network.self.flatten=nn.Flatten()# Construct a flattening layer.self.linear_relu_stack=nn.Sequential(# Construct a stack of layers.nn.Linear(28*28,512),# Linear Layers have an input and output shapenn.ReLU(),# ReLU is one of many activation functions provided by nnnn.Linear(512,512),nn.ReLU(),nn.Linear(512,10),)defforward(self,x):# This function defines the forward pass.x=self.flatten(x)logits=self.linear_relu_stack(x)returnlogits

References

[edit]

^Chintala, Soumith (1 September 2016)."PyTorch Alpha-1 release".GitHub.
^"PyTorch 2.6.0 Release". 29 January 2025. Retrieved2 February 2025.
^Claburn, Thomas (12 September 2022)."PyTorch gets lit under The Linux Foundation".The Register.
^Yegulalp, Serdar (19 January 2017)."Facebook brings GPU-powered machine learning to Python".InfoWorld. Retrieved11 December 2017.
^Lorica, Ben (3 August 2017)."Why AI and machine learning researchers are beginning to embrace PyTorch". O'Reilly Media. Retrieved11 December 2017.
^Ketkar, Nikhil (2017). "Introduction to PyTorch".Deep Learning with Python. Apress, Berkeley, CA. pp. 195–208.doi:10.1007/978-1-4842-2766-4_12.ISBN 9781484227657.
^Moez Ali (Jun 2023)."NLP with PyTorch: A Comprehensive Guide".datacamp.com. Retrieved2024-04-01.
^Patel, Mo (2017-12-07)."When two trends fuse: PyTorch and recommender systems".O'Reilly Media. Retrieved2017-12-18.
^Mannes, John."Facebook and Microsoft collaborate to simplify conversions from PyTorch to Caffe2".TechCrunch. Retrieved2017-12-18.FAIR is accustomed to working with PyTorch – a deep learning framework optimized for achieving state of the art results in research, regardless of resource constraints. Unfortunately in the real world, most of us are limited by the computational capabilities of our smartphones and computers.
^Arakelyan, Sophia (2017-11-29)."Tech giants are using open source frameworks to dominate the AI community".VentureBeat. Retrieved2017-12-18.
^"PyTorch strengthens its governance by joining the Linux Foundation".pytorch.org. Retrieved2022-09-13.
^"Top 30 Open Source Projects".Open Source Project Velocity by CNCF. Retrieved2023-10-12.
^"The C++ Frontend".PyTorch Master Documentation. Retrieved2019-07-29.
^Karpathy, Andrej (6 November 2019)."PyTorch at Tesla - Andrej Karpathy, Tesla".YouTube.
^"Uber AI Labs Open Sources Pyro, a Deep Probabilistic Programming Language".Uber Engineering Blog. 2017-11-03. Retrieved2017-12-18.
^PYTORCH-TRANSFORMERS: PyTorch implementations of popular NLP Transformers, PyTorch Hub, 2019-12-01, retrieved2019-12-01
^"Ecosystem Tools".pytorch.org. Retrieved2020-06-18.
^GitHub - catalyst-team/catalyst: Accelerated DL & RL, Catalyst-Team, 2019-12-05, retrieved2019-12-05
^"Ecosystem Tools".pytorch.org. Retrieved2020-04-04.
^"PyTorch – About".pytorch.org. Archived fromthe original on 2018-06-15. Retrieved2018-06-11.
^"Caffe2 Merges With PyTorch". 2018-04-02.
^Edwards, Benj (2022-09-12)."Meta spins off PyTorch Foundation to make AI framework vendor neutral".Ars Technica.
^"Dynamo Overview".
^"PyTorch 2.0 brings new fire to open-source machine learning".VentureBeat. 15 March 2023. Retrieved16 March 2023.
^"Installing PyTorch for ROCm".rocm.docs.amd.com. 2024-02-09.
^"Introducing Accelerated PyTorch Training on Mac".pytorch.org. Retrieved2022-06-04.
^"An Introduction to PyTorch – A Simple yet Powerful Deep Learning Library".analyticsvidhya.com. 2018-02-22. Retrieved2018-06-11.

External links

[edit]

Official website

v t e Deep learning software
Comparison
Open source	Apache MXNet Apache SINGA Caffe Deeplearning4j DeepSpeed Dlib Keras Microsoft Cognitive Toolkit ML.NET OpenNN PyTorch TensorFlow Theano Torch ONNX OpenVINO MindSpore
Proprietary	Apple Core ML IBM Watson Neural Designer Wolfram Mathematica MATLAB Deep Learning Toolbox
Category

v t e Differentiable computing
General	Differentiable programming Information geometry Statistical manifold Automatic differentiation Neuromorphic computing Pattern recognition Ricci calculus Computational learning theory Inductive bias
Hardware	IPU TPU VPU Memristor SpiNNaker
Software libraries	TensorFlow PyTorch Keras scikit-learn Theano JAX Flux.jl MindSpore
Portals Computer programming Technology

Retrieved from "https://en.wikipedia.org/w/index.php?title=PyTorch&oldid=1281854264"

Categories:

Hidden categories:

[8]ページ先頭

Movatterモバイル変換

History

PyTorch tensors

PyTorch neural networks

Example

See also

References

External links