AI on ESP32: AI Model on ESP32

The esp32-ai project is an open-source initiative that brings a language model to run entirely on the ESP32-S3 microcontroller, without the need for a GPU, internet, or API. This model boasts 28.9 million parameters and can operate with 512KB SRAM, 8MB PSRAM, and 16MB Flash.

The key to its efficiency lies in keeping the main computation within the fast memory, while the embedding table of approximately 25 million parameters is stored in Flash, with the system only reading the necessary parts when generating tokens. The model currently supports writing short stories for children but is not capable of answering questions or writing code.

However, this project demonstrates the expanding boundaries of AI on ultra-small devices. The model's specifications include 28.9 million parameters, a 4-bit model size of about 14.9MB, and a speed of around 9.5 tokens per second.

It can display 文本 directly on a small screen. For more information, visit the @@N8NLINK0@@.

References

These external sources were used to verify the article and provide deeper context.

Source Images

Conclusion

The esp32-ai project showcases the potential of running complex AI models on minimal hardware, paving the way for innovative applications of AI in constrained environments.

标签

你怎么认为?

发表回复 Cancel reply

Your email address will not be published. Required fields are marked *

相关文章

Graph Engineering Explained

Discover the concept of Graph Engineering and its significance in organizing workflows with 6 key components, including Node, Edge, and Conditional

阅读更多
联系我们

与我们合作进行数字创新

我们随时了解您的目标并为您的业务设计正确的解决方案 - 无论是人工智能自动化、营销系统、品牌推广还是数字化转型。

告诉我们您需要什么。我们将帮助您构建正确的方法。

请致电:+84 587 22 88 66
与我们合作您可以获得什么:
接下来会发生什么?
1

我们会在您方便的时候安排咨询

2

我们分析您的需求并定义正确的框架

3

我们准备符合您目标的战略提案

安排免费咨询
公司/组织
公司邮箱
我们能为您提供什么帮助?