Or run the model on the phone, free.
A cloud voice model charges you for every second it listens, and an assistant on your face listens constantly — so your bill grows with the engagement you were trying to win. A local model is free instead. Extentos runs the whole conversation on the user’s phone: speech in, the model thinking, speech out, with no server to pay for. Same handler code, same tools, same events. Only the brain moves.
- Published for iOS and Android — one dependency line, not a research project
- Automatic serves the best model each phone can genuinely sustain, and the cloud when none fits
- Eleven neural voices that run on the phone too — and it all keeps working with no signal