• m532@lemmy.ml
      link
      fedilink
      arrow-up
      1
      arrow-down
      1
      ·
      4 hours ago

      The article says “to run AI models” so it probably means both inference and training, not just inference.

          • geneva_convenience@lemmy.ml
            link
            fedilink
            arrow-up
            1
            ·
            50 minutes ago

            Training has a lot of extra functionality like calculating how to update the weights of the model during training to make it more performant on the dataset(backpropagation and gradient descent) and much more

            Meanwhile inference is mostly running the weights of the model as they are. The model isn’t being adjusted in any way. And Nvidia holds a strong grip on training libraries through Cuda