Audify

No phone app

5.6No. 19 of 23
in AI Podcast Generators
  • Recognised40% of the score17
  • Phone app26% of the score0
  • Documented20% of the score85
  • Free plan14% of the score30
Free plan
No
Runs on
api, self-hosted, Web

Summary

Audify is ranked #19 of 23 in AI podcast generators on Samsung Mobile US Press. It runs on API, Self-hosted, Web.

Compared on AI podcast generators

Host dialogue
Yesgithub.com
Source imports
PDF, DOC, DOCXgithub.com
Audio export
mp3github.com

Facts

Purpose
Audify converts PDF, DOC, and DOCX documents into editable two-speaker podcast scripts and downloadable MP3 episodes.github.com · 4 Oct 2026
Document processing
It extracts text from PDF, DOC, and DOCX files and uses OCR as a fallback for scanned PDFs with little detected text.github.com · 4 Oct 2026
Script editing
Users can review the generated dialogue and edit speaker turns, wording, and flow before audio generation.github.com · 4 Oct 2026
Voice selection
The interface lets users choose host and guest voices from available text-to-speech options.github.com · 4 Oct 2026
Script generation
The LLM service uses Ollama as its primary script-generation path and OpenAI as a fallback, with conversational, educational, and professional tone options.github.com · 4 Oct 2026
Audio generation
The TTS service generates speech for dialogue turns, combines them into an MP3 podcast, and provides playback and download.github.com · 4 Oct 2026
Integrations
The documented external providers are Ollama for local model inference and OpenAI for fallback script generation and text-to-speech.github.com · 4 Oct 2026
Deployment
Audify is a FastAPI microservices application with a React and Vite frontend, orchestrated for local deployment with Docker Compose.github.com · 4 Oct 2026
API
The gateway exposes endpoints for document upload, script generation, audio generation, job status, and MP3 download.github.com · 4 Oct 2026
Requirements
The setup instructions require Docker Engine 24.x or later and Docker Compose Plugin v2 or later; an OpenAI API key is required for TTS, while Ollama is optional.github.com · 4 Oct 2026
Upload limit
The configured maximum upload size is 10 MB, and the README recommends keeping uploads under 10 MB.github.com · 4 Oct 2026
Processing limit
The TTS service configuration accepts up to 100 dialogue turns and allows up to five concurrent requests.github.com · 4 Oct 2026
Job persistence
Gateway and TTS job stores are in memory, so jobs reset when containers restart unless persistent storage is added.github.com · 4 Oct 2026
Accuracy guidance
The project advises users to review scripts, validate OCR output for scanned documents, and check generated audio and dialogue for factual accuracy and tone.github.com · 4 Oct 2026
License
The repository identifies the project license as MIT.github.com · 4 Oct 2026

Best Audify alternatives

See all 12

Where it ranks on Samsung Mobile US Press

Is Audify yours?

Claim it for free: prove the domain, then correct facts, plans and screenshots. An editor reviews every change.

Sources