Persona Designer
Tone, personality traits, and greeting style - every conversation sounds like your brand.
Turn your data into a sovereign voice agent - persona, knowledge, tools, and voice, live on your site in one line of code.

Real businesses running Oshara agents on live traffic. Click any one and talk to it yourself.
Name it, describe it, set personality and greeting - no code.
Upload documents, connect URLs and tools so your agent can answer and act.
Choose the LLM, our Lyric or Fantom speech models, and the TTS voice that powers your agent.
Set the widget title, brand colors, avatar, and display settings - previewed live as you edit.
Drop one script tag on your site, or call the agent-session API. The widget appears automatically.
Persona, knowledge, tools, models, appearance, and forms - tuned in one place, with a live preview of your widget the whole way.
Tone, personality traits, and greeting style - every conversation sounds like your brand.
PDFs and URLs, chunked and retrieved automatically mid-conversation.
MCP servers, HTTP endpoints, and client events - your agent can act, not just talk.
Any LLM, our Lyric (90+ languages) or Fantom (1600+ languages) speech models, and a TTS voice library.
Name, colors, avatar, and watermark - previewed live as you edit.
Forms your agent surfaces mid-call to book demos, capture leads, or run intake.
Production voices tuned for English, Nepali, and Hindi - preview each one before you assign it to an agent.
Sovereign, low-latency, and fine-tuned on your data - the infrastructure behind every agent.
All audio and data processed locally - nothing sent to foreign clouds.
Optimized for real-time voice UX, production-ready out of the box.
Fine-tuned on Nepali and regional dialects - nuance, grammar, and vocabulary included.
The voice model adapts tone to emotional context, reducing churn.
OFTA handles data prep, training, and evaluation - bring your own data, we run the pipeline.
Wire up call-transfer rules so a conversation moves - with full context - to the right specialist.
Decide when a conversation escalates, then preview the hand-off.
One agent routes a live conversation out to the right specialist.
Your agent listens and speaks in the same live session - voice and persona set independently, so it never waits its turn.
One channel at a time. Fast, but each side still waits for a turn.
Both channels open. Speech overlaps, as it does with people.
Pick a template or build your own — your agent surfaces the right form mid-call, fields pre-filled from the conversation.
Click a template to preview it - your agent picks automatically mid-call.
Forms that fill themselves - detected mid-call, pre-filled, ready to confirm.
Embed the hosted widget with one snippet, or build a fully custom experience on top of the agent-session API.
We host it - paste one script tag on any page and the widget appears.
<script
src="https://api.oshara.ai/widget.js"
data-agent="your-agent-id">
</script>Mint a session with your API key and connect your own UI to the live agent.
curl -X POST https://api.oshara.ai/api/agents/agent-session/ \
-H "x-api-key: sk_..." \
-H "Content-Type: application/json" \
-d '{ "agent": "your-agent-id" }'Every conversation stays on infrastructure you control.
You pay for the audio your agent actually generates - a flat $0.01 per minute of text-to-speech. No seats, no platform fee, no idle-time billing.
That's $0.60 an hour of generated speech - the same rate whether you ship one agent or a hundred.
100 min × $0.01 per minute
≈ 1.7 hours of generated audio
Final usage is billed per minute generated.
Design the persona, feed it your knowledge, and deploy to your site in an afternoon - no ML team required.