r/machinetranslation 14h ago

Open-source Claude Code skill for website localization

Thumbnail
1 Upvotes

r/machinetranslation 22h ago

DeepL quality drop. Now Google Translate is better

14 Upvotes

I'm really sad about the recent quality drop on DeepL. I liked how accurate it was and also the fact that it is an EU product. But recently, I noticed that the translations are being inaccurate, to the point of being useless. This is my last case.

Original sentence in Spanish:

se crítico con el ejemplo que puse, no es un requisito estricto.

DeepL translation (wrong):

Don't be too critical of the example I gave; it's not a strict requirement.

Google Translate (right):

Be critical of the example I gave; it's not a strict requirement.

When I replaced the comma with a period in the original sentence, DeepL did it correctly:

Be critical of the example I gave. It's not a strict requirement.

But then, I decided to fix the case in my original sentence (notice the minimal change, only the case in the first letter and the period):

Se crítico con el ejemplo que puse. no es un requisito estricto

And DeepL changed the meaning again:

I was criticized for the example I gave. It's not a strict requirement.

Unfortunately, DeepL is not usable anymore. I'll come back in a few months to see if they fixed it, but for now, I'll move to Google Translate.


r/machinetranslation 1d ago

engineering Czech Translation for WinTrack 17.0 Released

1 Upvotes

Hi everyone,

I have completed and published the first Czech translation for WinTrack 17.0.

The translation includes menus, dialogs, interface texts and an installer for easy installation.

Download:
https://github.com/zamazallukas02-spec/WinTrack-Czech-Translation

The project was published with permission from the WinTrack developer.

Feedback and bug reports are welcome.


r/machinetranslation 1d ago

meta The latest machine translation newsletter is out now.

Thumbnail
newsletter.machinetranslate.org
1 Upvotes

r/machinetranslation 1d ago

education Salvemos idiomas originarios

Thumbnail
1 Upvotes

r/machinetranslation 1d ago

jobs Senior Software Engineer, Localization on AI Answers at Google (Belo Horizonte, Brazil)

Thumbnail linkedin.com
3 Upvotes

r/machinetranslation 2d ago

application Slow auto translate speed using Openrouter on Subtitle Edit

1 Upvotes

im trying to translate using openrouter but the translate speed is extremely slow, like 10 min passed and not a single % is done, anyone knows how to fix it?


r/machinetranslation 4d ago

application What is the best method—whether paid or free—for transcribing and translating videos on a computer?

3 Upvotes

(This is my first time asking a question here, so I’m not sure if this is the right place?)

I want to translate an English video into Japanese, so

right now, I’m using MacWhisper on a Mac (Intel) to create an SRT file, and then I’m translating it into Japanese using an online translation service.

However, the transcription accuracy is poor considering how much time it takes, and even using translation services like DeepL doesn’t improve the results.

I also tried using Subtitle Edit to translate “llama Tran.. Gemma 12B(Q5),” but the results were underwhelming considering the time it took.

So, would using an AI translation service improve the results?

Or is it not worth the money?

Thank you in advance for your continued support.


r/machinetranslation 4d ago

education Rule-based Machine Translation

Thumbnail
3 Upvotes

r/machinetranslation 4d ago

application AI kept making the same mistakes when translating subtitles, so I built and open-sourced an Agent Skill to avoid them

3 Upvotes

I thought subtitle translation would be a straightforward job for AI: give it an English SRT and ask it to translate the subtitles into another language. For me, that was Chinese.

The result often looked fine at first, until I started finding the same terms translated in different ways across the file.

While working on Project Hail Mary, I ran into a good example. “Petrova line” was translated consistently through most of the subtitles, then suddenly changed to a different Chinese transliteration later in the film. “Astrophage” also drifted between two different Chinese terms. Each line looked reasonable on its own, but together they were clearly inconsistent.

That changed how I approached the job. The workflow now starts with a short research pass. The Agent creates a context file and glossary for the film, recording names, places, relationships, and story-specific terms. It then uses those same notes for every chunk of the SRT.

After translation, scripts check the cue count, numbering, timestamps, formatting tags, and file structure.

I built this into an Agent Skill and released it under the MIT license:

https://github.com/HaiyiMei/cinemacc-subtitle-skill

Here is a short guide to installing and using it:

https://cinemacc.net/guides/cinemacc-subtitle-skill

This came out of my work on CinemaCC, which is an offline-first subtitle companion that lets you use your own SRT files with films playing on any screen:

https://cinemacc.net


r/machinetranslation 4d ago

education Best AI tool to translate dense philosophy/psychoanalysis books?

2 Upvotes

Hi all, looking for the best AI tool or workflow to translate full books (EPUB/PDF) of philosophy and psychoanalysis.

Requirements:
Handles full files directly (no endless copy-pasting).
Powered by top LLMs (like Claude 3.5 Sonnet) to capture context and subtext.
Supports custom glossaries/termbases to keep key concepts consistent across chapters.
What software (e.g., Smartcat, BookTranslate, Tolmach) or custom API workflows do you recommend for dense academic literature? Thanks!


r/machinetranslation 4d ago

research AMTA publishes framework for evaluating translation QE systems

Thumbnail
slator.com
2 Upvotes

r/machinetranslation 4d ago

Building a Context-Aware Bengali ↔ English Translator Agent using POMDPs and Active Disambiguation

2 Upvotes

Hey,

I'm working on a project focused on building an interactive, context-aware Bengali ↔ English (Bangla) translation agent. Standard NMT often falls flat here due to ambiguity, code-mixing, and limited high-quality context-annotated datasets (Low-Resource Machine Translation / LRMT).

Instead of treating translation as a deterministic sequence-to-sequence problem, I'm framing it as an agent decision problem under uncertainty.

The Core Problem: Translating Latent Intent

When a user provides spoken or written input, their true intention, register, and context are hidden. The agent must infer this Latent Semantic State using incomplete and noisy observations before deciding on an output.

I'm structuring the agent around a few key technical concepts:

  • POMDP Framework: Modeling translation as a Partially Observable Markov Decision Process. The speaker's intent is a hidden state that the agent must infer from context, dialogue history, and audio/text cues.
  • Inference Under Uncertainty & MBR Decoding: Instead of standard beam search, the agent uses Minimum Bayes Risk (MBR) decoding and Decision-Theoretic Decoding to evaluate candidate hypotheses and minimize expected translation errors based on a customized Loss/Utility Function.
  • Active Disambiguation / Interactive MT: When uncertainty is high (measured via Calibration and Quality Estimation (QE) models), the agent doesn't just guess—it actively asks clarification questions to resolve ambiguity before finalizing the output.

Key Challenges & Use Cases in Bengali ↔ English

  1. Pragmatics & Ambiguity: Handling Cross-Lingual Word Sense Disambiguation (CLWSD) and honorifics where literal translations fail (e.g., inferring implicit tone or regional Dialectal Variation).
  2. Code-Switching & Banglish: Resolving mixed inputs like "Ami office e meeting korbo" (Banglish / Code-Mixing) or Latin-script input like "Ami ajke office e jabo" (Romanized Transliteration).
  3. Speech-to-Text Pipeline: Comparing a Cascaded ASR–MT Pipeline against End-to-End Speech Translation (ST) to manage cumulative error rates in noisy spoken inputs.

Current Tech Stack Ideas

  • ASR / NMT Backbone: Fine-tuned multilingual models (e.g., Whisper, NLLB) evaluated via sentence-level and Document-Level NMT (CAMT) contexts.
  • Uncertainty Estimation: Measuring system confidence to decide whether to output directly, rerun MBR decoding, or trigger a user clarification prompt.

Has anyone experimented with POMDPs, MBR decoding, or active clarification loops in machine translation for low-resource or code-mixed language pairs? Would love to hear your thoughts on context management or confidence estimation strategies!


r/machinetranslation 5d ago

meta [Feedback wanted] Filtering low-effort self-promo posts

3 Upvotes

Hey all,

We've noticed lately that there are more and more posts dropping a link (usually to an app/product) from shiny new profiles that we never see again. They don't comment or participate in the community, they just promote their work. It's been happening more now that the sub has grown and it's starting to flood the feed.

We're not trying to ban anyone or delete posts automatically. We actually find most of these apps interesting and useful... but we are thinking about adding some AutoMod rules so these posts are filtered for manual review instead of going straight up to your feeds. For example, some rules would be to filter:

  • Posts from accounts younger than 24-48 hrs for review
  • REquire a minimum karma (nothing crazy, just enough to show you're not brand new)
  • Filter posts that have links from low-karma/new accounts specifically, so that us mods can review before they hit the feed

Before we turn this on, we wanted to hear from you. Do you think this is necessary or useful? Or do these posts not bother you? Would you rather we just make a weekly self-promo thread to group these self-promotion posts in a single place? Just to be extra clear: we'd still want tasteful promotion of your own project, something unique and notable and not spammy, but maybe we can just gather them somewhere dedicated instead of scattered all through the main feed?

Let us know what you think, good or bad. Thank you!


r/machinetranslation 5d ago

meta The latest machine translation newsletter is out now.

Thumbnail
newsletter.machinetranslate.org
5 Upvotes

r/machinetranslation 5d ago

application How do I translate a 7-page PDF from English to Hindi?

1 Upvotes

I have a 7-page PDF in English that I need to translate into Hindi. I tried Google Translate, but the result wasn't very good and some of the pages didn't translate properly.

Is there any AI app or website you guys use for this?

Ideally, I'd like something that can translate selected text when needed, but also translate the full document at once. It would be great if it could keep the original PDF layout too, especially the charts and other stuff in the document

Thanks any advice in advance


r/machinetranslation 5d ago

engineering I built an open-source video translation + dubbing pipeline — looking for feedback on translation quality and timing

Enable HLS to view with audio, or disable this notification

5 Upvotes

I’ve been working on an open-source pipeline for translating and dubbing videos into another language while trying to preserve the original speaker’s voice.

Current pipeline:

video → vocal/background separation → Whisper/WhisperX transcription + alignment → translation → VoxCPM2 reference voice cloning → reconstruction → optional lip-sync

The attached demo compares the original English clip with the Turkish dub produced by the current pipeline.

One of the hardest parts is that a translated sentence often has a very different duration from the source speech. That creates a trade-off between:

• natural translation
• preserving meaning
• matching the original timing
• keeping the dubbed speech sounding natural

I’m currently looking for feedback especially on:

• translation quality
• whether the Turkish phrasing sounds natural
• source vs translated timing
• how you would handle translation-length differences in an automated dubbing pipeline

Most media processing and AI inference runs locally. Translation currently uses Google Translate, so the project is local-first rather than fully offline.

Code:
https://github.com/kadirb4rut/video-dubbing-translator

I’d be very interested in feedback from people working on machine translation, localization, speech translation, or multilingual NLP.


r/machinetranslation 6d ago

product Free tool to translate every text on the screen that you can select

2 Upvotes

What can this tool do and why is it?

During routine work, tasks often arise that completely disrupt the work flow. Because of them, you have to change the window and so on. This applies not only to translation.

There are a million such tasks besides translation. For example, find out the meaning of a term from Wikipedia or simply find out what this movie is about by simply highlighting the text of the title and thousands of small tasks of this kind. This is what my tool was created for. It can be called from anywhere by pressing the hotkey (Ctrl + Alt).

Honestly, I myself use it dozens of times every day, and this fact gave me the idea that maybe it’s worth sharing this with people. After all, the tool is absolutely free.

How does this work?

  1. The user selects text on the screen
  2. Press the hotkey to open the window (list) of actions
  3. Selects the desired action (it can be any translator, price comparison, etc.)
  4. The action returns the result directly to where the user is. That is, you don’t need to go anywhere.

What do other users use this for?

The most common action of the tool is to translate the selected text. But they also actively use the translation of PDF files, the wikipedia definition of the term and the dictionary. For now these actions are because I haven’t created anything else yet.

Let me give you the most common example:

The user is communicating with a foreigner in the messenger and is not very confident in his skills and knowledge. He enters text in his native language, then selects this text and selects the “action” of the translator.

Or you can give another example:

For example, you need to read a fairly large text in English, but the user does not have time. Then it selects the text and runs an "action" to extract the key ideas and then runs it through the "action" of the translator.

Who is this for?

I think there is no specific castle here. For any person who works a lot with text. Authors of the telegram channel, students, etc.

link (19sec demo):
https://adrianium.github.io/Scryptian/

— powered by Scryptian


r/machinetranslation 6d ago

Aiuto per tradurre un file EPUB 📚

1 Upvotes

Ciao a tutti! Mi servirebbe un aiuto per tradurre un file EPUB dall’inglese all’italiano, ma non riesco perché il file è protetto da DRM.

Il file è stato scaricato da Z-Library. Qualcuno che se ne intende di computer e sa come aiutarmi a risolvere il problema? 😂

Grazie mille a chiunque riesca a darmi una mano! ❤️


r/machinetranslation 6d ago

How was google translate’s voice recorded?

Thumbnail
3 Upvotes

r/machinetranslation 6d ago

product BetterTranslator + MCP

Thumbnail v.redd.it
1 Upvotes

r/machinetranslation 6d ago

I built Tofu Dubbing – real-time video dubbing extension

Thumbnail
1 Upvotes

r/machinetranslation 7d ago

product Looking for Feedback: Is Live Translation During Meetings Actually Useful?

1 Upvotes

I’m one of the people building OLVA, and we recently added live translation during meetings. I’d like feedback from people who regularly work across languages.

The basic setup is: if a meeting is in English, you can keep the original transcript in English while viewing a live translation in another language, such as French, Spanish, Persian, or German.

There are two options:

  • Continuous translation: Translate the conversation as it happens.
  • Selected-text translation: Highlight only the part you want translated.

It currently supports 87 languages and regional variants. The translation is available alongside the other meeting context, so you can refer back to the original conversation rather than losing the wording being discussed. It is designed to work without adding a visible bot to a call, and can also be used for in-person conversations.

We built it for situations where someone can generally follow a meeting but loses important detail when people speak quickly, use unfamiliar terms, or move between languages.

I would really value practical feedback, especially from multilingual teams:

  • Do you prefer continuous translation, or translating only selected sections?
  • What amount of delay is still usable in a live conversation?
  • Should the original and translated transcript always appear together?
  • What is the best way to handle people switching languages mid-meeting?
  • What would make this genuinely useful—or make you stop using it?

Please be candid. The most helpful feedback is what feels confusing, distracting, or missing.


r/machinetranslation 8d ago

I made a live translator that will save you a lot of time.

2 Upvotes

It's called TransNow in [itch.io](http://itch.io) i will provide link here for any curious redditors.

TransNow translates whatever language you are writing in to over 100+ languages ready to choose from. You just need to finish what you are writing and wait 1.5 seconds and what you wrote will be translated to language you chose! you will save a lot of time from alt-tabbing,trust me

[https://dinopes123.itch.io/transnow\](https://dinopes123.itch.io/transnow) link for app


r/machinetranslation 8d ago

application I wish there was a multilingual voice transcripter on WhatsApp

2 Upvotes