← All case studies

Friendsy: a real-time Voice AI platform for phone-based AI agents

A full-cycle engagement covering architecture, backend, dashboard, and the real-time audio pipeline, delivered as a scalable SaaS platform where users deploy low-latency AI agents on real phone lines.

ClientFriendsy
IndustrySaaS · Voice AI
RoleFull-stack & AI development
StackPipecat · Twilio · WebRTC · Next.js · FastAPI · PostgreSQL · Redis · Docker · AWS/GCP
Friendsy AI analytics dashboard: 1,234 total calls, 94.2% success rate, 0.8s average AI response time, call volume and duration charts
Live analytics: call volume, success rate, escalation rate, and per-call API costs across all deployed numbers.

The challenge

Friendsy needed a production platform where non-technical users can deploy AI agents that answer and place real phone calls, and hold a natural conversation. The hard part of voice AI is latency: every extra half-second between the caller finishing a sentence and the AI replying makes the agent feel robotic. The system had to run speech-to-text, an LLM, and text-to-speech in a real-time streaming loop, stay fast under load, and remain flexible enough to swap providers at every stage.

What we built

We took the product from concept to production. At the core is a Pipecat-based real-time audio pipeline over WebRTC, connected to phone networks through Twilio and other VOIP providers for both inbound and outbound calling. An STT → LLM → TTS routing layer lets each agent mix and match providers (Deepgram or OpenAI Whisper for transcription; GPT-4.1, Claude, or Gemini for reasoning; Cartesia or ElevenLabs for voice) with live cost estimation for any combination.

Around the pipeline sits the SaaS layer: a Next.js dashboard with a step-by-step agent builder, phone number management, call logs and recordings, real-time analytics over WebSockets, and a flexible API for programmatic control. The backend runs on FastAPI with PostgreSQL, Redis, and S3 storage, shipped in Docker containers to AWS/GCP through CI/CD pipelines, engineered to scale from the first customer without a rewrite.

Friendsy agent builder: three-step configuration wizard with AI model, language, and STT/LLM/TTS provider selection
The agent builder: a three-step wizard for configuring model, language, call limits, and AI providers.
Friendsy price calculator: building a voice AI stack from transport, AI model, speech-to-text, text-to-speech and telephony providers
The price calculator: users assemble their voice stack and see per-minute and monthly costs instantly.

The result

A production SaaS platform live on real phone lines: 0.8-second average AI response latency, a 94.2% call success rate, and thousands of calls handled, with an architecture the Friendsy team can confidently extend, from new AI providers to new telephony integrations.

"Working with your team was a strong experience from start to finish. You delivered a production ready SaaS solution with solid architecture, clean code, and a clear focus on scalability and long term maintainability."

Thomas Timm Stensrud, CTO of Friendsy

Building a SaaS product?

We'll help you get to production, and stay there.

Get a free proposal