Episode Details

Back to Episodes
More powerful deep learning with transformers (Ep. 84)

More powerful deep learning with transformers (Ep. 84)

Episode 80 Published 6 years, 6 months ago
Description

Some of the most powerful NLP models like BERT and GPT-2 have one thing in common: they all use the transformer architecture.
Such architecture is built on top of another important concept already known to the community: self-attention.
In this episode I explain what these mechanisms are, how they work and why they are so powerful.

Don't forget to subscribe to our Newsletter or join the discussion on our Discord server

 

References
Listen Now

Love PodBriefly?

If you like Podbriefly.com, please consider donating to support the ongoing development.

Support Us