BeFreed

BeFreed builds AI-powered audio lessons that sound like real conversations

BeFreed is an AI-powered learning platform that turns books, research papers, expert talks, and podcasts into personalized audio lessons. With more than 1M+ learners, the company is building a learning experience designed to fit around everyday life, from commutes to workouts to household chores.

Products used
Fish Audio S2 ProText-to-SpeechCloud API

BeFreed turns books into conversations worth listening to

BeFreed takes dense source material and turns it into short, personalized lessons. Learners can choose the voice, tone, and pace, then listen wherever they are. The experience is built around audio rather than the screen, making learning possible during the parts of the day when reading is not.

The company's standout format takes that idea further. Instead of simply narrating a book or paper, BeFreed turns it into an AI podcast: a back-and-forth between two distinct voices discussing the material. Each episode can run 10, 20, or 40 minutes, giving learners something closer to a podcast they commissioned for themselves.

That format made voice a core part of the product. The two speakers had to sound like distinct people having a conversation, not two voices reading generated text.

A 40-minute conversation is a different TTS problem

A voice that sounds good for a sentence does not necessarily hold up for an entire episode. Over 20 or 40 minutes, small problems become more noticeable: flat delivery, repetitive expression, or two voices that sound like they are speaking past each other.

BeFreed also needed personalization to work across the experience. Learners choose the voices they want to spend an entire episode with, so the platform needed a broad range of distinct voices rather than a single default narrator.

And because BeFreed ships in more than 10 languages, the same experience had to work across markets. Naturalness, expression, voice variety, and language coverage all had to hold up together.

Audio is the heart of how our users learn, so voice quality is non-negotiable. Fish Audio gives us natural, emotionally expressive, multi-voice audio at the scale and cost we need to make every piece of content listenable.

Jisong Liu

Founder, BeFreed

BeFreed needed voices that could perform, not just narrate

During its evaluation, BeFreed focused on how voices performed across the type of long-form conversational content at the center of its product.

Fish Audio stood out for its naturalness and emotional range over longer passages. Two voices could maintain distinct identities while delivering a more expressive back-and-forth, giving BeFreed the foundation for its AI podcast format.

The breadth of voices also mattered. Because personalization is built into the product, BeFreed could give learners more choice over who they listen to rather than designing the experience around a single voice.

Fish Audio gives BeFreed a voice layer built for long-form audio

BeFreed integrated Fish Audio's text-to-speech API directly into its content pipeline. When a learner selects a voice and tone, BeFreed generates the corresponding audio as part of the in-app experience.

The same voice infrastructure supports both monologue lessons and two-voice conversational episodes. That lets BeFreed build different learning formats without creating a separate voice production workflow for each one.

The economics matter as well. BeFreed can generate audio on demand rather than manually producing every lesson, making it practical to turn a large library of source material into personalized listening experiences.

One voice layer across a global learning platform

BeFreed now ships AI-personalized audio lessons in more than 10 languages to a global community of more than 1,000,000 learners.

Fish Audio's broader language coverage gives the team room to extend that experience without rebuilding its voice infrastructure for every market. As BeFreed adds formats and expands its interactive audio experience, the same voice layer can support the product underneath.

The architecture keeps voice inside the product rather than treating it as a separate content-production workflow.

Building a learning experience that talks back

BeFreed is extending personalized audio beyond generated lessons into interactive experiences where learners can talk to the lesson in real time.

That moves the product from listening to interacting. The same principle behind its AI podcasts continues into these experiences: learning can adapt to the person, the moment, and the way they want to engage.

Fish Audio provides the voice layer underneath that evolution, giving BeFreed a foundation for audio experiences that can become more personalized, conversational, and expressive over time.

Read more customer stories

Fish Audio

Ready when you are

Talk to our team about your deployment. We'll come prepared.