Skip to content
#

speech-to-text-transcription

Here are 54 public repositories matching this topic...

This repository contains a web application for multi-lingual transcription using OpenAI's Whisper Automatic Speech Recognition (ASR) model. Users can upload audio files in WAV, MP3, or M4A formats and get transcriptions in various languages. The application is designed with accessibility and data privacy in mind.

  • Updated Feb 12, 2026
  • Python

CLI that turns a YouTube URL into a text transcript: yt-dlp for the audio, Whisper for the transcription, running 100% locally with GPU support.

  • Updated Sep 4, 2026
  • Python

Successfully developed an interview preparation guide using Langchain which can effectively guide users in their interview preparation process and job search journeys by providing valuable insights and feedback regarding their performance. It generates a comprehensive list of questions pertaining to a user query as well.

  • Updated Apr 3, 2025
  • Python

Lipi is a lightweight, local-first desktop speech-to-text and voice note application. Built with Tauri v2, Rust, and React 19, it lets you record audio, transcribe speech with OpenAI Whisper or any OpenAI-compatible local/remote endpoint, and manage your notes seamlessly.

  • Updated Sep 28, 2026
  • TypeScript

Add this topic to your repo

To associate your repository with the speech-to-text-transcription topic, visit your repo's landing page and select "manage topics."

Learn more