Machine Learning

1 readers

1 users here now

Community Rules:

Be nice. No offensive behavior, insults or attacks: we encourage a diverse community in which members feel safe and have a voice.
Make your post clear and comprehensive: posts that lack insight or effort will be removed. (ex: questions which are easily googled)
Beginner or career related questions go elsewhere. This community is focused in discussion of research and new projects that advance the state-of-the-art.
Limit self-promotion. Comments and posts should be first and foremost about topics of interest to ML observers and practitioners. Limited self-promotion is tolerated, but the sub is not here as merely a source for free advertisement. Such posts will be removed at the discretion of the mods.

founded 2 years ago

MODERATORS

communick@academy.garden

[D] Difference between CUDA and Tensor Cores (alien.top)

submitted 2 years ago by 3DHydroPrints@alien.top to c/machinelearning@academy.garden

13 comments fedilink hide all child comments

CUDA cores (or shader cores in general) have long been used to compute graphics. A very often used operation in computer graphics are matrix multiplications, just like in deep learning. Back in the days (AlexNet) NNs were computed using shader cores, but now have completely moved to be computed on Tensor cores. My question are:

Why have these workloads been seperated? (Yes obviously the tensor cores are more specialized and leave out a bunch of unnecessary operations, but how and why not integrate it into the CUDA cores to boost MM operations for computer graphics?)
Why isn't the workload offloaded to the other cores when the mathematical operations are the same
What makes tensor cores so much more efficient and faster?

you are viewing a single comment's thread
view the rest of the comments

[–] smokingPimphat@alien.top 1 points 2 years ago

Building a solution that solves the specific problem you have is always going to yield faster and more efficient results compared to something that is even 90% a solution.

A recent example would be bitcoin ASICs. While not into crypto personally it was amazing to see just how fast bitcoin ASICs got rolled out.

Its now at a point where no one in their right mind would use anything else and if a faster one gets released people clamor to grab as many as they can afford.

Having dedicated hardware to do the specific math for ML is the logical move.