Expedera NPUs Run Large Language Models Natively on Edge Devices
Highlights
- Expedera NPU IP adds native support for LLMs, including stable diffusion
- Origin NPUs deliver the power-performance profile needed to run LLMs on edge and portable devices
LLMs bring a new level of natural language processing and understanding capabilities, making them versatile tools for enhancing communication, automation, and data analysis tasks. They unlock new capabilities in chatbots, content generation, language translation, sentiment analysis, text summarization, question-answering systems, and personalized recommendations. Due to their large model size and the extensive processing required, most LLM-based applications have been confined to the cloud. However, many OEMs want to reduce reliance on costly, overburdened data centers by deploying LLMs at the edge. Additionally, running LMM-based applications on edge devices improves reliability, reduces latency, and provides a better user experience.
"Edge AI designs require a careful balance of performance, power consumption, area, and latency," said
Expedera's patented packet-based NPU architecture eliminates the memory sharing, security, and area penalty issues that conventional layer-based and tiled AI accelerator engines face. The architecture is scalable to meet performance needs from the smallest edge nodes to smartphones to automobiles. Origin NPUs deliver up to 128 TOPS per core with sustained utilization averaging 80%—compared to the 20-40% industry norm—avoiding dark silicon waste.
For more information or to contact an Expedera representative in your region, visit www.expedera.com.
About Expedera
Expedera provides customizable neural engine semiconductor IP that dramatically improves performance, power, and latency while reducing cost and complexity in edge AI inference applications. Successfully deployed in over 10 million consumer devices, Expedera's Neural Processing Unit (NPU) solutions are scalable and produce superior results in applications ranging from edge nodes and smartphones to automotive. The platform includes an easy-to-use TVM-based software stack that allows the importing of trained networks, provides various quantization options, automatic completion, compilation, estimator, and profiling tools, and supports multi-job APIs. Headquartered in
Media Contact:
Paul Karazuba, Vice President of Marketing
408-421-2119
[email protected]
View original content to download multimedia:https://www.prnewswire.com/news-releases/expedera-npus-run-large-language-models-natively-on-edge-devices-302027586.html
SOURCE Expedera, Inc
Serious News for Serious Traders! Try StreetInsider.com Premium Free!
You May Also Be Interested In
- HyGen Promotes Omid Ghobadi to Chief Operating Officer
- Wilson Legal Becomes Love Legacy Estate Planning, Cumming Law Firm Helps Clients Commemorate "Heart Message" in their Estate Plans
- Rigel to Participate in Upcoming September Investor Conferences
Create E-mail Alert Related Categories
PRNewswire, Press ReleasesSign up for StreetInsider Free!
Receive full access to all new and archived articles, unlimited portfolio tracking, e-mail alerts, custom newswires and RSS feeds - and more!



Tweet
Share