I would end up jumping between a video player, subtitles/transcription, a dictionary, screenshots, audio clips and Anki. So I built SubSmith to bring that workflow together.
You can drop a video or audio file into it, generate a transcript locally and then use the transcript alongside the media to:
* look up words and sentences * replay individual lines * edit the transcript * save useful sentences with their original context/audio * export them as Anki cards
The important part for me is that it works with your own media. It isn't tied to a particular streaming service or library, so I can use the random anime episode, podcast, lecture, etc. that I'm actually interested in studying.
It's an offline-first desktop app, and transcription happens locally rather than sending the media to a transcription API.
I'm sharing it here because I'm now more interested in finding out where this workflow breaks down for other people rather than adding features randomly now that I have solid core/base.
For example:
* Would you actually save sentences from your own media? * Which part of this process feels like too much work? * Does having the audio/context attached make creating an Anki card more useful? * Would you prefer this to work inside your existing video player/browser? * Is installing a desktop app a significant barrier? * And does requiring an account before starting the free trial make you give up?
The current version does require an account to start the trial, and I'm trying to work out whether that's meaningful friction for the people who would actually use this.
It's free to try, and I'd particularly appreciate feedback from people who already learn languages through their own videos, anime, films, podcasts or other media.
I'm the developer, so I'll be around in the comments to answer questions and discuss how it works.
k__•1h ago
While I'm not so sure about learning a language with content from fictional sources, that should at least make it more relevant for the learners.
And you can use it on podcasts too, where people talk normally.
The biggest hurdle is probably that users need to get hold of the media so they can use it as input.
IbrahimF96•54m ago
Yeah that is my suspicion too, I am considering adding an option to transcribe system audio too so it could hook into whatever video or audio is playing on the computer. That would be quite a big change so would like feedback first! hahah