An MCP server: works with Claude Code, Cursor, and Codex.
Give your agent a podcast or video link, or just a topic, and it gets the words of the part that matters, with the minute each was said, turned into text on your own computer.
omniseek_search(query="speculative decoding", sources=["apple_podcasts"])
# picks an episode from the returned feed, then:
omniseek_transcribe(url="https://www.buzzsprout.com/2529712/episodes/18794851-speculative-speculative-decoding.mp3",
start="1:00", duration="0:45", language="en")
From a real run on 2026-10-10. The first time, it downloads a speech-to-text model (about 900 MB). After that, transcripts are saved and reused.
Tip: ask for the part you need, say minutes 40 to 50, not the whole two-hour episode.
uv tool install "omniseek[asr]" --with torch --with torchaudio claude mcp add omniseek -- omniseek
asr extra and PyTorch). Details and disk sizes: Podcast and video transcription.--torch-backend cpu to the first command, or PyTorch downloads several GB of CUDA libraries you will not use.apple_podcasts: find shows and episode feeds by topicxiaoyuzhou: Chinese podcast episodeschinese_podcasts: Chinese tech and research podcast feedsyoutube: videos, transcripts, and top commentsbilibili: Chinese videospodcast_index: podcast directory (needs a free key)Every source, with whether it needs a key or a login: docs/sources.md. Tool reference: docs/tools.md.