Vision Transformers Change the AI Acceleration Rules
Transformers were first introduced by the team at Google Brain in 2017 in their paper, “Attention is All You Need“. Since their introduction, transformers have inspired a flurry of investment and research which have produced some of the most impactful model architectures and AI products to-date, including ChatGPT which is an acronym for Chat Generative Pre-trained Transformer.
Transformers are also being employed for vision applications (ViTs). This new class of models was empirically proven to be viable alternatives to more traditional Convolutional Neural Networks (CNNs) in the paper “An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale“, published by the team at Google Brain in 2021.
Vision transformers are of a size and scale that are approachable for SoC designers targeting the high-performance, edge AI market. There’s just one problem: vision transformers are not CNNs and many of the assumptions made by the designers of first-generation Neural Processing Unit (NPU) and AI hardware accelerators found in today’s SoCs do not translate well to this new class of models.
What makes Vision Transformers so special?
To read the full article, click here
Related Semiconductor IP
- Safety Enhanced GPNPU Processor IP
- GPNPU Processor IP - 32 to 864TOPs
- GPNPU Processor IP - 16 to 108 TOPs
- GPNPU Processor IP - 4 to 28 TOPs
- GPNPU Processor IP - 1 to 7 TOPs
Related Blogs
- From ChatGPT to Computer Vision Processing: How Deep-Learning Transformers Are Shaping Our World
- Vision Transformers Have Already Overtaken CNNs: Here’s Why and What’s Needed for Best Performance
- CNNs and Transformers: Decoding the Titans of AI
- Rethinking AI Infrastructure: The Rise of PCIe Switches
Latest Blogs
- Implementing I3C Host Controller support in the open source I3C Core
- SystemLens Gives System Engineers Real-Time Visibility Into Hidden Silicon Issues Under Functional Tests
- As Vehicles Become Data Centers: Why Interface IP Integration Matters in SoC Design
- Implementing a multi-stage secure boot chain on an Efinix Titanium Ti375
- Last-Level Caches Matter Even More in the HBM Era