Skip to content
Ajith Thaduri

Voice & realtime

AI you can talk to, that answers fast enough to feel like a conversation.

Voice agents live or die on timing. I build pipelines around a latency budget — speech in, reasoning and tools in the middle, speech out — with turn-taking and interruption handled properly.

  • What this covers
  • Speech-to-text, reasoning and ElevenLabs speech out
  • Streaming end to end, with a latency budget per stage
  • Turn-taking and barge-in
  • Tool calls mid-conversation without awkward silence

Work in this area

Customer support

Realtime Voice Agent

Speech in, reasoning and tool calls in the middle, ElevenLabs speech out. Most of the engineering is about timing — a reply that arrives a beat late stops feeling like a conversation — so the pipeline is built around latency.

Contact

Working on something
like this?

I'm open to AI engineering, architecture and training work. Tell me what you're building and what the constraints are — that's usually enough to start.

Prefer a short form? Send a project brief
  • Taking on new projects
  • Usually replies within a day