NihonSub:基于 Whisper-Large-v3 与 Groq/DeepSeek 的开源日语动漫实时字幕翻译引擎
I built an open-source real-time Japanese anime subtitle & translation engine powered by Whisper-Large-v3 + Groq / DeepSeek
作者发布开源工具 NihonSub,可将原始日语视频转成上下文双语字幕,并配有同步影院播放器。流水线用 ffmpeg 静音检测做 VAD 切分,Whisper Large-v3 在 Groq LPU 上转写。
| Hey r/LocalLLaMA, Like many anime fans, I've always been frustrated by traditional MT engines (like Google Translate or base DeepL) when dealing with raw Japanese anime: - They completely butcher Japanese honorifics, sentence-ending particles (-tteba, -zo, -desu wa), and character slang. - They struggle with subject dropping (pro-drop grammar), translating pronouns inconsistently line-by-line. - Cloud transcription APIs often choke on background music (OST), loud sound effects, and character screaming. To solve this, I built NihonSub — an open-source tool and synchronized cinema player that turns raw Japanese video files into contextual bilingual subtitles. 🛠️ Architecture & Pipeline:
💡 Why not just rely on standard NMT?LLMs are far superior at resolving who is speaking to whom based on context and tone rather than naive literal dictionary lookup. With zero-cost free-tier APIs (Groq + OpenRouter free models), the entire pipeline runs without subscription costs. Check out the demo video above! - GitHub Repository: https://github.com/Abhishantpadam/NihonSub - License: MIT I'd love your thoughts on the pipeline, optimization ideas for local edge models (like running Whisper.cpp or local Ollama instances), or any feedback! submitted by /u/Grand_Marionberry115[link] [留言] |
来源:r/LocalLLaMA · reddit.com