chiprook
← AI
AIOctober 5, 2026, 16:07

Ito: speech synthesis in 4.89 MB per voice on ESP32-S3

Lokutor released the Ito inference engine, ESP32-S3 firmware and two English voice models: 4.05M parameters and 4.89 MB of weights per voice, producing 24 kHz speech with integer arithmetic. The firmware is verified in Espressif's QEMU emulator, but nothing has run on a physical board yet.

Ito: speech synthesis in 4.89 MB per voice on ESP32-S3
#ESP32-S3#Lokutor#Ito
Read next
AI

Oído: open-source speech recognition for ESP32-S3 without a command list

AI

421M-parameter Laya model runs in a browser tab, download up to 478 MB

Science

Palladium membrane enables ammonia synthesis from water in South Korea

AI

Suno Launches Speech Beta, Generating Voice and Music in One Model