Public catalog
media-processing
31 shown. Search ranks against problem language; browsing defaults to stronger public signals.
whisper
local speech-to-text transcription offline
Verified 2026-08-10 ยท 44,228 stars
auto-editor
remove silence and repetitive dead space from recordings automatically
Verified 2026-08-10
BetterShot
You need to capture, annotate, and blur-redact screenshots on macOS for docs or bug reports without paying for a CleanShot X subscription
Verified 2026-08-10
Chatterbox
generate natural text to speech locally for narration and voice workflows
Verified 2026-08-10
ComfyUI
build reproducible node-based image and media generation workflows
Verified 2026-08-10
CutItOut
You need to remove backgrounds from product photos at full resolution without a watermark or per-image charge.
Verified 2026-08-10
Dolphin RVZ/ISO Conversion Scripts
Your GameCube and Wii ISO library is eating disk space and you want batch conversion to RVZ, the lossless compressed format Dolphin reads natively
Verified 2026-08-10
elevenlabs-python
use elevenlabs-python for browser automation
Verified 2026-08-10
Faster Whisper TransWithAI ChickenRice
You need to batch-convert Japanese audio or video into Chinese SRT/VTT/LRC subtitles on a local Windows GPU without writing any code
Verified 2026-08-10
faster-whisper
transcribe long audio locally with efficient Whisper inference
Verified 2026-08-10
FFmpeg
convert, trim, combine, encode, filter, or caption audio and video
Verified 2026-08-10
Hunyuan PromptEnhancer
Users type short vague prompts and your text-to-image model drops subjects, styles, or layout details; you want automatic chain-of-thought rewriting into structured prompts before generation
Verified 2026-08-10
Insanely Fast Whisper
Transcribing hours of recorded audio with Whisper large-v3 is far too slow on your GPU and you need batched fp16 or Flash Attention 2 throughput
Verified 2026-08-10
Kokoro-FastAPI
You need an OpenAI-compatible text-to-speech endpoint running on your own hardware so audio never leaves the machine and there is no per-character bill
Verified 2026-08-10
LosslessCut
You need to rough-cut hours of camera, GoPro, or drone footage down to the good parts quickly, reclaiming gigabytes without a slow quality-losing re-encode
Verified 2026-08-10
MetaSort
You exported Google Photos via Takeout and every photo lost its date, camera, and GPS metadata to .json sidecar files
Verified 2026-08-10
Mini QR
You need hundreds of branded QR codes generated from a spreadsheet of URLs without a paid generator
Verified 2026-08-10
OpenAI.fm
You need to compare OpenAI text-to-speech voices and delivery styles in a UI before committing one to your application code
Verified 2026-08-10
Piper
generate speech locally without per-character API charges
Verified 2026-08-10
pyannote.audio
identify who spoke when in meetings, interviews, and podcasts
Verified 2026-08-10
Retrieval-based Voice Conversion WebUI
You have recorded vocals or speech and want them re-voiced in another timbre while keeping the melody, words, and timing
Verified 2026-08-10
TostUI
You want to run a supported image-generation or editing model such as Flux.2 or Qwen Image Edit behind a local web UI without assembling its Python environment manually.
Verified 2026-08-10
video-use
browser-use style agents that watch or control video surfaces
Verified 2026-08-10
Whisper ASR Box
Multiple apps or services need speech to text and you want one self-hosted REST endpoint instead of bundling Whisper into each of them
Verified 2026-08-10
Whisper Diarization
You have a meeting or interview recording and need a transcript that labels which speaker said each sentence, processed locally
Verified 2026-08-10
Whisper Standalone Win (Faster-Whisper executables)
You want Faster-Whisper transcription on Windows or Linux without installing Python, pip, or CUDA toolchains
Verified 2026-08-10
whisper.cpp
You need offline on-device speech transcription without installing a Python or PyTorch environment
Verified 2026-08-10
WhisperKit (Argmax OSS SDK)
An iOS or macOS app must transcribe user audio without uploading recordings to any server, for privacy or offline operation
Verified 2026-08-10
WhisperLive
You need live partial transcripts from a microphone, RTSP camera, or HLS broadcast stream instead of waiting for a whole file to finish processing
Verified 2026-08-10
WhisperLiveKit
You need live low-latency captions with speaker labels from meetings or streams without sending audio to a cloud service
Verified 2026-08-10
WhisperX
produce word-level timestamps and speaker-aware transcripts
Verified 2026-08-10