MerchantryTidbits

Public catalog

Browse tools

9 shown. Search ranks against problem language; browsing defaults to stronger public signals.

media-processingclilocal

Insanely Fast Whisper

Transcribing hours of recorded audio with Whisper large-v3 is far too slow on your GPU and you need batched fp16 or Flash Attention 2 throughput

Verified 2026-08-10

media-processingapilocal

Kokoro-FastAPI

You need an OpenAI-compatible text-to-speech endpoint running on your own hardware so audio never leaves the machine and there is no per-character bill

Verified 2026-08-10

llm-inferencelibraryfree

Transformers.js

You want text classification, embeddings, translation, object detection, or speech recognition inside a web app without standing up any inference backend

Verified 2026-08-10

media-processingapilocal

Whisper ASR Box

Multiple apps or services need speech to text and you want one self-hosted REST endpoint instead of bundling Whisper into each of them

Verified 2026-08-10

media-processingclifree

Whisper Standalone Win (Faster-Whisper executables)

You want Faster-Whisper transcription on Windows or Linux without installing Python, pip, or CUDA toolchains

Verified 2026-08-10

media-processingclilocal

whisper.cpp

You need offline on-device speech transcription without installing a Python or PyTorch environment

Verified 2026-08-10

media-processinglibrarylocal

WhisperKit (Argmax OSS SDK)

An iOS or macOS app must transcribe user audio without uploading recordings to any server, for privacy or offline operation

Verified 2026-08-10

media-processinglibrarylocal

WhisperLive

You need live partial transcripts from a microphone, RTSP camera, or HLS broadcast stream instead of waiting for a whole file to finish processing

Verified 2026-08-10

media-processingclilocal

WhisperLiveKit

You need live low-latency captions with speaker labels from meetings or streams without sending audio to a cloud service

Verified 2026-08-10