DeepSeek to order 160000 Huawei AI chips over Nvidia
-
This post did not contain any content.
-
This post did not contain any content.
Only for inference like the article says or also training?
-
Only for inference like the article says or also training?
they might not need more training infrastructure at this point
-
Only for inference like the article says or also training?
The article says "to run AI models" so it probably means both inference and training, not just inference.
-
The article says "to run AI models" so it probably means both inference and training, not just inference.
Running refers to inference usually
-
Running refers to inference usually
If they can be used for inference, I would assume they can be used for training
-
Running refers to inference usually
They run during training too. Also, why would they get 160000 chips only for inference?
-
If they can be used for inference, I would assume they can be used for training
Training has a lot of extra functionality like calculating how to update the weights of the model during training to make it more performant on the dataset(backpropagation) and much more
Meanwhile inference is mostly running the weights of the model as they are. The model isn't being adjusted in any way. And Nvidia holds a strong grip on training libraries through Cuda
Ciao! Sembra che tu sia interessato a questa conversazione, ma non hai ancora un account.
Stanco di dover scorrere gli stessi post a ogni visita? Quando registri un account, tornerai sempre esattamente dove eri rimasto e potrai scegliere di essere avvisato delle nuove risposte (tramite email o notifica push). Potrai anche salvare segnalibri e votare i post per mostrare il tuo apprezzamento agli altri membri della comunità.
Con il tuo contributo, questo post potrebbe essere ancora migliore 💗
Registrati Accedi
Citiverse è un progetto che si basa su NodeBB ed è federato! | Categorie federate | Chat | 📱 Installa web app o APK | 🧡 Donazioni | Privacy Policy