newvui for iPhone is in beta

Teaching machines
to speak human.

We build the whole voice stack — recognition, reasoning and speech — from the weights up. Small enough to run on a phone. Cheap enough to build a business on.

fig. 1 — a real call
preview · muted
  1. Call
  2. Call
  3. Call
Agents built on the Fluxions API, answering real business lines. Play one — or press Call and ring it yourself, no account needed.

$0.10/ min

for a whole voice agent — recognition, LLM and voice, all-in

300Mparams

in vui-nano, speaking in real time on the iPhone GPU

$0.20/ hour

of transcription, with speakers and non-speech events

§01 — Phone agents

Answers the phone like your best receptionist.

Agents built on our API take bookings, triage call-outs and fill appointment books for real businesses — you heard three of them in fig. 1.

$0.10 a minute, all-in. Numbers are bring-your-own or provided at cost.

fig. 2 — what answers the phone
  1. akro — listens

    Speech to text — who is speaking, and the breaths, laughs and hesitations between words.

  2. LLM — thinks

    Decides what to say, and calls your tools — bookings, lookups, web search.

  3. vui — speaks

    Text to an expressive voice, streamed back while it is still being written.

All three stages are ours, so there is one bill: $0.10 a minute, all-in.

§02 — vui for iPhone

A voice assistant that never leaves your phone.

vui runs speech recognition, the language model and a 300M-parameter voice entirely on the iPhone GPU. No server does the thinking, so nothing you say is sent anywhere.

  • Real voices
  • Talk to it
  • Private by construction
  • Works offline

Free during beta · iPhone with iOS 26+ · ~1.2 GB one-time download

fig. 3 — vui-nano, on an iPhone
35 seconds, sound on — every word is generated on the phone.

§03 — API

The whole stack. One endpoint. One bill.

We train our own recognition, reasoning and voice, so nothing is stacked on top — no separate LLM bill, no TTS bill, no keys to bring. Billed by the second.

Voice agents, all-in
$0.10/ min
Transcription
$0.20/ hour
Text-to-speech
$9/ 1M chars
lst. 1 — render speech
curl -X POST https://api.fluxions.ai/vui/v1/tts \
-H "Authorization: YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"voice": "maeve.h8ff7e07da",
"input": "[sigh] fine, I will say it one more time."
}' \
--output speech.wav
POST /vui/v1/tts. Square brackets are performed, not read: [sigh], [laugh], [gasp], [hesitate].

§04 — Licensing

Run it on your own metal.

The same models, air-gapped in your infrastructure or on your own devices. Annual licence with SLAs and dedicated support.