TypeScript speech-to-text

Open-source TypeScript projects categorized as speech-to-text

Top 13 TypeScript speech-to-text Projects

  • Leon

    🧠 Leon is your open-source personal assistant.

    Project mention: Rabbit R1, Designed by Teenage Engineering | news.ycombinator.com | 2024-01-09

    It's indeed suspicious. You're sending your voice samples, your various services accounts, your location and more private data to some proprietary black box in some public cloud. Sorry, but this is a privacy nightmare. It should be open source and self-hosted like Mycroft (https://mycroft.ai) or Leon (https://getleon.ai) to be trustworthy.

  • audapolis

    an editor for spoken-word audio with automatic transcription

    Project mention: Audapolis: An editor for spoken-word audio with automatic transcription | news.ycombinator.com | 2023-08-23
  • SurveyJS

    Open-Source JSON Form Builder to Create Dynamic Forms Right in Your App. With SurveyJS form UI libraries, you can build and style forms in a fully-integrated drag & drop form builder, render them in your JS app, and store form submission data in any backend, inc. PHP, ASP.NET Core, and Node.js.

  • voice-assistant

    Voice assistant for Visual Studio Code.

  • gdansk-ai

    🦭 Full stack AI voice chatbot (speech-to-text, LLM, text-to-speech) with integrations to Auth0, OpenAI, Google Cloud and Stripe - Web App, Web API and AI API

    Project mention: Ask HN: How do you find contributors to open source projects? | news.ycombinator.com | 2023-10-12

    - Gdańsk AI - Full stack AI voice chatbot (speech-to-text, LLM, text-to-speech) with integrations to Auth0, OpenAI, Google Cloud and Stripe - Web App, Web API and AI API https://github.com/jmaczan/gdansk-ai [Python, TypeScript, Next.js]

  • web-voice-processor

    A library for real-time voice processing in web browsers

  • deepgram-js-sdk

    Official JavaScript SDK for Deepgram's automated speech recognition APIs.

  • simple-obs-stt

    Speech-to-text and keyboard input captions for OBS.

  • InfluxDB

    Power Real-Time Data Analytics at Scale. Get real-time insights from all types of time series data with InfluxDB. Ingest, query, and analyze billions of data points in real-time with unbounded cardinality.

  • svelte-speech-recognition

    Speech recognition library for Svelte

  • transcript.fish

    Unofficial No Such Thing As A Fish episode transcripts.

    Project mention: Ask HN: Tell us about your project that's not done yet but you want feedback on | news.ycombinator.com | 2023-08-16

    I have been working on this podcast transcription project for a couple months and it's been super rewarding.

    I listen to a podcast called No Such Thing As A Fish, where some researchers talk about their favorite facts they learned that week. Then they riff on it and are generally smart and funny. I listened to the series so many times that I decided I wanted to listen to the show on shuffle, not at the episode level, but at the fact level.

    Since I have been playing around with whisper.cpp in python this seemed like a perfect way to combine some technologies I've been wanting to play with.

    I ran whisper over the entire podcast and transcribed all the episodes. I had to do this multiple times because I kept messing up. It eventually took like 7 straight days of my M1 processing to get through ~490 episodes.

    4 million words, and an 800Mb SQLite database later, I got the transcriptions done and have put up a nice site for searching through the data.

    https://transcript.fish

    Now I just need to figure out the rest. Breaking it up into facts. Getting the audio working. Highlighting and linking to words, phrases, etc.

    Some cool info about the process so far:

    1. The SQLite database is chunked up and stored as static files, and the frontend queries the static files directly using HTTP range requests, so it only downloads a couple hundreds kbs when querying.

    2. I've been proper using ChatGPT 3.5 free version to help me write python and SQL. It's been pretty game changing as I feel basically no pain from not knowing what I'm doing.

    The code is here: https://github.com/noman-land/transcript.fish

    Please help if you know how to get whisper speaker diarization working!! I would really appreciate the help.

  • speech-to-element

    A simple way to add speech to text functionality to your website :microphone:

    Project mention: Speech To Element - embed speech to text into your website with ease | /r/github | 2023-08-25

    A GitHub star is always appreciated 🌟 https://github.com/OvidijusParsiunas/speech-to-element

  • deepgram-deno-sdk

    Deno SDK for Deepgram's automated speech recognition APIs

  • open-chatbot-js

    Chatbot with text-generation-webui and ChatGPT 3.5 backend with a sort-of long-term memory.

    Project mention: Chatbot based on GPT 3.5 | /r/Entrepreneur | 2023-04-20

    If you decided you want to do go forward with this you can have a look at my project https://github.com/ChrisDeadman/open-chatbot-js it's MIT license so you can use whatever you want. Good luck :)

  • voice-to-text-notes

    A basic speech to text app.

NOTE: The open source projects on this list are ordered by number of github stars. The number of mentions indicates repo mentiontions in the last 12 Months or since we started tracking (Dec 2020). The latest post mention was on 2024-01-09.

TypeScript speech-to-text related posts

Index

What are some of the best open-source speech-to-text projects in TypeScript? This list will help you:

Project Stars
1 Leon 14,415
2 audapolis 625
3 voice-assistant 281
4 gdansk-ai 173
5 web-voice-processor 161
6 deepgram-js-sdk 102
7 simple-obs-stt 99
8 svelte-speech-recognition 25
9 transcript.fish 15
10 speech-to-element 10
11 deepgram-deno-sdk 5
12 open-chatbot-js 4
13 voice-to-text-notes 3
The modern identity platform for B2B SaaS
The APIs are flexible and easy-to-use, supporting authentication, user identity, and complex enterprise features like SSO and SCIM provisioning.
workos.com