Welcome to the official issue tracker, developer documentation, and community hub for FreeAudioToText.
FreeAudioToText is a privacy-first, zero-friction AI speech-to-text platform that converts audio and video recordings into high-precision transcripts and subtitles in seconds. We believe basic transcription is a fundamental utility — which is why our core audio-to-text conversion is 100% free with unlimited usage, zero account registration, and no daily caps.
Unlike traditional paid tools (Otter.ai, Descript, Rev) that enforce strict 3-file daily caps or charge $9-$30/month subscriptions just for basic transcription, FreeAudioToText is built for developers, indie hackers, researchers, and creators:
- 🚀 100% Free & No Signup: Upload immediately without creating an account or verifying an email.
- ⚡️ Blazing-Fast Local Apple Silicon Architecture: Under the hood, we run optimized FunASR (Paraformer) and Faster-Whisper inference engines on local Apple Silicon (M2 Max 32GB) orchestrated via Cloudflare Edge (D1 & R2). A 1-hour audio file is processed in as little as 13 minutes (~783 seconds).
- 🎙️ High-Precision Speaker Diarization: Automatically identifies and separates multiple speakers (
SPK_0,SPK_1) in podcasts, interviews, and meetings. - 🌍 90+ Languages & Dialects: Auto-detects and transcribes English, Mandarin Chinese, Japanese, Spanish, French, German, Arabic, and bilingual mixed content with industry-leading accuracy.
- 🔒 True Ephemeral Privacy: All uploaded audio files and database records are automatically and permanently purged within 24 hours. We never use your private audio for AI model training.
Are you a developer looking to integrate speech-to-text into your Telegram bots, Slack workflows, Notion wikis, or automation scripts without paying massive API bills?
We offer a clean, developer-friendly Speech-to-Text API:
- REST & OpenAI-compatible endpoints: Easily swap into existing Whisper API integrations.
- Webhook & Asynchronous processing: Perfect for long podcast episodes and meeting recordings.
- Affordable Pay-as-you-go credits: No recurring monthly subscriptions.
👉 Explore our API Documentation & Start Building →
We provide specialized, zero-click converters optimized for every media format and use case:
| Audio Converters | Video Converters | Use Case Specialized |
|---|---|---|
| 🎵 MP3 to Text | 🎬 Video to Text | 🎙️ Podcast Transcript |
| 📱 M4A to Text | 📺 YouTube to Text | 🤝 Interview Transcription |
| 🎙️ WAV to Text | 📱 TikTok Transcript | 🧠 AI Analysis Reports |
| 📽️ MP4 to Text | ⚡️ Speech-to-Text API | 🛠️ All Tools Directory |
How do we keep basic transcription 100% free without ads or paywalls? We operate a healthy, self-sustained Freemium & Pay-as-you-go model:
- Core Transcription (100% Free): Transcribe any audio/video to text with speaker labels and timestamps for free.
- AI Structural Analysis Reports ($1.99 one-time): For power users needing an AI-generated Executive Summary, Structural Outline, Action Items, and an interactive shareable public link (
?share=1), we charge a simple one-time fee per report. No recurring subscriptions. - Developer API: Pay-as-you-go credits for programmatic API usage.
Please note that this repository does not contain the core AI inference engine source code. Our proprietary multi-model routing logic and Apple Silicon worker orchestration remain closed-source to protect our infrastructure from abuse.
This repository serves as our official community hub to:
- Track bug reports and technical issues.
- Collect community feedback and feature requests.
- Host public developer documentation and integration guides.
We love hearing from developers and users! If you encounter an issue or have an idea to make FreeAudioToText even better:
- Report a Bug: Open a Bug Report
- Request a Feature: Open a Feature Request
If you discover a security vulnerability, please review our Security Policy instead of opening a public issue.
Built with ❤️ for indie hackers, developers, and creators worldwide.
Try FreeAudioToText Now →
