esp32-ai
A 28.9 million parameter language model runs on an $8 ESP32-S3 microcontroller entirely on-device, generating simple stories at about 9.5 tokens per second.
它提供了完整的高性能语音克隆服务栈,具备 A100 上 0.148 RTF 的实测性能、开发者友好的 HTTP API、Web 管理后台以及音频后处理功能,同时将声纹特征提取从每次推理中解耦,大幅降低重复上传开销。
Production-ready CosyVoice serving with Triton, reusable Speaker Registry, Public HTTP API and Web console.
根据分类、Topic 和编程语言匹配的相似项目。
A 28.9 million parameter language model runs on an $8 ESP32-S3 microcontroller entirely on-device, generating simple stories at about 9.5 tokens per second.

A comprehensive guide for building and configuring a high-end local machine to run state-of-the-art LLMs, with detailed hardware choices, BIOS tuning, and Docker-based model serving.
Cindy is an open-source AI agent that runs locally on your machine, integrates multiple AI harnesses and models, and provides memory, skills, and automation to perform real work in your projects and apps.