#machinetranslation — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #machinetranslation, aggregated by home.social.
-
Kagi Blog: Kagi Translate is back. “Today, Kagi Translate is moving to a freemium model and launching a new standalone Individual subscription for $8/month. Anyone with a Kagi account can continue using Translate for free with a limited monthly allowance. For people who use it more regularly, the new Individual plan includes all Kagi Translate features. Translate will also continue to be […]
https://rbfirehose.com/2026/09/14/kagi-blog-kagi-translate-is-back/ -
Kagi Blog: Kagi Translate is back. “Today, Kagi Translate is moving to a freemium model and launching a new standalone Individual subscription for $8/month. Anyone with a Kagi account can continue using Translate for free with a limited monthly allowance. For people who use it more regularly, the new Individual plan includes all Kagi Translate features. Translate will also continue to be […]
https://rbfirehose.com/2026/09/14/kagi-blog-kagi-translate-is-back/ -
Kagi Blog: Kagi Translate is back. “Today, Kagi Translate is moving to a freemium model and launching a new standalone Individual subscription for $8/month. Anyone with a Kagi account can continue using Translate for free with a limited monthly allowance. For people who use it more regularly, the new Individual plan includes all Kagi Translate features. Translate will also continue to be […]
https://rbfirehose.com/2026/09/14/kagi-blog-kagi-translate-is-back/ -
Kagi Blog: Kagi Translate is back. “Today, Kagi Translate is moving to a freemium model and launching a new standalone Individual subscription for $8/month. Anyone with a Kagi account can continue using Translate for free with a limited monthly allowance. For people who use it more regularly, the new Individual plan includes all Kagi Translate features. Translate will also continue to be […]
https://rbfirehose.com/2026/09/14/kagi-blog-kagi-translate-is-back/ -
Spudart: Deciphering the Matrix Digital Rain. “…if it’s not secretly a recipe for spicy tuna rolls, what is on the screen? I decided to find out the dumbest, most 2026 way possible. I ran it through Google Translate. It’s kind of fun to point today’s ‘instantly understand any language’ technology at a 27-year-old movie prop and just see what happens.”
https://rbfirehose.com/2026/07/25/spudart-deciphering-the-matrix-digital-rain/ -
Spudart: Deciphering the Matrix Digital Rain. “…if it’s not secretly a recipe for spicy tuna rolls, what is on the screen? I decided to find out the dumbest, most 2026 way possible. I ran it through Google Translate. It’s kind of fun to point today’s ‘instantly understand any language’ technology at a 27-year-old movie prop and just see what happens.”
https://rbfirehose.com/2026/07/25/spudart-deciphering-the-matrix-digital-rain/ -
Spudart: Deciphering the Matrix Digital Rain. “…if it’s not secretly a recipe for spicy tuna rolls, what is on the screen? I decided to find out the dumbest, most 2026 way possible. I ran it through Google Translate. It’s kind of fun to point today’s ‘instantly understand any language’ technology at a 27-year-old movie prop and just see what happens.”
https://rbfirehose.com/2026/07/25/spudart-deciphering-the-matrix-digital-rain/ -
Spudart: Deciphering the Matrix Digital Rain. “…if it’s not secretly a recipe for spicy tuna rolls, what is on the screen? I decided to find out the dumbest, most 2026 way possible. I ran it through Google Translate. It’s kind of fun to point today’s ‘instantly understand any language’ technology at a 27-year-old movie prop and just see what happens.”
https://rbfirehose.com/2026/07/25/spudart-deciphering-the-matrix-digital-rain/ -
Unseen Japan: Japanese Social Media Auto-Translation: a Blessing, or a Curse?. “Plenty of people outside Japan browsed Japanese posts, but unless they could speak Japanese themselves or they deliberately hit the ‘Translate Post’ button, conversations mostly stayed within their own language communities. Ah, those sweet, halcyon days. Now, social media sites like X are auto-translating languages, […]
https://rbfirehose.com/2026/07/12/japanese-social-media-auto-translation-a-blessing-or-a-curse-unseen-japan/ -
Unseen Japan: Japanese Social Media Auto-Translation: a Blessing, or a Curse?. “Plenty of people outside Japan browsed Japanese posts, but unless they could speak Japanese themselves or they deliberately hit the ‘Translate Post’ button, conversations mostly stayed within their own language communities. Ah, those sweet, halcyon days. Now, social media sites like X are auto-translating languages, […]
https://rbfirehose.com/2026/07/12/japanese-social-media-auto-translation-a-blessing-or-a-curse-unseen-japan/ -
Unseen Japan: Japanese Social Media Auto-Translation: a Blessing, or a Curse?. “Plenty of people outside Japan browsed Japanese posts, but unless they could speak Japanese themselves or they deliberately hit the ‘Translate Post’ button, conversations mostly stayed within their own language communities. Ah, those sweet, halcyon days. Now, social media sites like X are auto-translating languages, […]
https://rbfirehose.com/2026/07/12/japanese-social-media-auto-translation-a-blessing-or-a-curse-unseen-japan/ -
Unseen Japan: Japanese Social Media Auto-Translation: a Blessing, or a Curse?. “Plenty of people outside Japan browsed Japanese posts, but unless they could speak Japanese themselves or they deliberately hit the ‘Translate Post’ button, conversations mostly stayed within their own language communities. Ah, those sweet, halcyon days. Now, social media sites like X are auto-translating languages, […]
https://rbfirehose.com/2026/07/12/japanese-social-media-auto-translation-a-blessing-or-a-curse-unseen-japan/ -
LibreTranslate is a free, open-source machine translation server that you can use online or self-host for complete control over your data.
It supports multiple languages, offers a simple API, works without relying on Google Translate, and is a great privacy-friendly choice for websites, apps, and personal use.
More details: https://digitalescapetools.com/tools/tool.html?id=libretranslate
#OpenSource #Privacy #SelfHosted #Translation #Linux #FOSS #MachineTranslation
-
LibreTranslate is a free, open-source machine translation server that you can use online or self-host for complete control over your data.
It supports multiple languages, offers a simple API, works without relying on Google Translate, and is a great privacy-friendly choice for websites, apps, and personal use.
More details: https://digitalescapetools.com/tools/tool.html?id=libretranslate
#OpenSource #Privacy #SelfHosted #Translation #Linux #FOSS #MachineTranslation
-
LibreTranslate is a free, open-source machine translation server that you can use online or self-host for complete control over your data.
It supports multiple languages, offers a simple API, works without relying on Google Translate, and is a great privacy-friendly choice for websites, apps, and personal use.
More details: https://digitalescapetools.com/tools/tool.html?id=libretranslate
#OpenSource #Privacy #SelfHosted #Translation #Linux #FOSS #MachineTranslation
-
LibreTranslate is a free, open-source machine translation server that you can use online or self-host for complete control over your data.
It supports multiple languages, offers a simple API, works without relying on Google Translate, and is a great privacy-friendly choice for websites, apps, and personal use.
More details: https://digitalescapetools.com/tools/tool.html?id=libretranslate
#OpenSource #Privacy #SelfHosted #Translation #Linux #FOSS #MachineTranslation
-
https://www.alojapan.com/1499852/washington-post-journalist-visits-unl-plastic-research-labs-news/ Washington Post journalist visits UNL plastic research labs | News #ComputationalLinguistics #FellowsOfTheAmericanAssociationForTheAdvancementOfScience #MachineTranslation #Nebraska #news #Osaka #OsakaNews #ShannonOsaka #UniversityOfNebraska #大阪 #大阪府 Muhammad Saiful Islam (middle) during a Nanoparticle Tracking Analysis demonstration for Shannon Osaka (L) and Dr. Kazi Albab Hussain (R) on Monday, June 8, 2026, at the Biomedical and Obesity Re
-
https://www.alojapan.com/1499852/washington-post-journalist-visits-unl-plastic-research-labs-news/ Washington Post journalist visits UNL plastic research labs | News #ComputationalLinguistics #FellowsOfTheAmericanAssociationForTheAdvancementOfScience #MachineTranslation #Nebraska #news #Osaka #OsakaNews #ShannonOsaka #UniversityOfNebraska #大阪 #大阪府 Muhammad Saiful Islam (middle) during a Nanoparticle Tracking Analysis demonstration for Shannon Osaka (L) and Dr. Kazi Albab Hussain (R) on Monday, June 8, 2026, at the Biomedical and Obesity Re
-
I read a light novel series entirely translated by Google Translate!
And then I wrote up a guide on how I did it.
-
I read a light novel series entirely translated by Google Translate!
And then I wrote up a guide on how I did it.
-
I read a light novel series entirely translated by Google Translate!
And then I wrote up a guide on how I did it.
-
I read a light novel series entirely translated by Google Translate!
And then I wrote up a guide on how I did it.
-
via #AIFoundry : Azure Translator: Improving Translation Quality with Adaptive Datasets and Few‑Shot Learning
https://ift.tt/JGMPZ8D
#AzureTranslator #AdaptiveDatasets #FewShotLearning #MachineTranslation #NLP #AI #Foundry #TechBlog #TranslationQuality #DomainContext #Termino… -
via #AIFoundry : Azure Translator: Improving Translation Quality with Adaptive Datasets and Few‑Shot Learning
https://ift.tt/JGMPZ8D
#AzureTranslator #AdaptiveDatasets #FewShotLearning #MachineTranslation #NLP #AI #Foundry #TechBlog #TranslationQuality #DomainContext #Termino… -
via #AIFoundry : Azure Translator: Improving Translation Quality with Adaptive Datasets and Few‑Shot Learning
https://ift.tt/JGMPZ8D
#AzureTranslator #AdaptiveDatasets #FewShotLearning #MachineTranslation #NLP #AI #Foundry #TechBlog #TranslationQuality #DomainContext #Termino… -
via #AIFoundry : Azure Translator: Improving Translation Quality with Adaptive Datasets and Few‑Shot Learning
https://ift.tt/JGMPZ8D
#AzureTranslator #AdaptiveDatasets #FewShotLearning #MachineTranslation #NLP #AI #Foundry #TechBlog #TranslationQuality #DomainContext #Termino… -
Kagi Blog: An update on Kagi Translate. “Some of you might have noticed that Translate no longer works when you’re signed out, or that it shows as offline. This is because we have temporarily turned off free access while we work through the cost of running Translate. If you have an active subscription, Translate still works.”
https://rbfirehose.com/2026/06/06/kagi-blog-an-update-on-kagi-translate/ -
Kagi Blog: An update on Kagi Translate. “Some of you might have noticed that Translate no longer works when you’re signed out, or that it shows as offline. This is because we have temporarily turned off free access while we work through the cost of running Translate. If you have an active subscription, Translate still works.”
https://rbfirehose.com/2026/06/06/kagi-blog-an-update-on-kagi-translate/ -
Kagi Blog: An update on Kagi Translate. “Some of you might have noticed that Translate no longer works when you’re signed out, or that it shows as offline. This is because we have temporarily turned off free access while we work through the cost of running Translate. If you have an active subscription, Translate still works.”
https://rbfirehose.com/2026/06/06/kagi-blog-an-update-on-kagi-translate/ -
Kagi Blog: An update on Kagi Translate. “Some of you might have noticed that Translate no longer works when you’re signed out, or that it shows as offline. This is because we have temporarily turned off free access while we work through the cost of running Translate. If you have an active subscription, Translate still works.”
https://rbfirehose.com/2026/06/06/kagi-blog-an-update-on-kagi-translate/ -
Stocktonia: Stockton police roll out live AI language translation software for body cameras. “On Tuesday, the Stockton Police Department announced the integration of a new, real-time translation feature into the more than 300 body cameras issued to in-the-field officers. The new tool, said Stockton Police Chief Stanley McFadden during a morning news conference, would eliminate delays in […]
https://rbfirehose.com/2026/05/20/stocktonia-stockton-police-roll-out-live-ai-language-translation-software-for-body-cameras/ -
Stocktonia: Stockton police roll out live AI language translation software for body cameras. “On Tuesday, the Stockton Police Department announced the integration of a new, real-time translation feature into the more than 300 body cameras issued to in-the-field officers. The new tool, said Stockton Police Chief Stanley McFadden during a morning news conference, would eliminate delays in […]
https://rbfirehose.com/2026/05/20/stocktonia-stockton-police-roll-out-live-ai-language-translation-software-for-body-cameras/ -
Stocktonia: Stockton police roll out live AI language translation software for body cameras. “On Tuesday, the Stockton Police Department announced the integration of a new, real-time translation feature into the more than 300 body cameras issued to in-the-field officers. The new tool, said Stockton Police Chief Stanley McFadden during a morning news conference, would eliminate delays in […]
https://rbfirehose.com/2026/05/20/stocktonia-stockton-police-roll-out-live-ai-language-translation-software-for-body-cameras/ -
Stocktonia: Stockton police roll out live AI language translation software for body cameras. “On Tuesday, the Stockton Police Department announced the integration of a new, real-time translation feature into the more than 300 body cameras issued to in-the-field officers. The new tool, said Stockton Police Chief Stanley McFadden during a morning news conference, would eliminate delays in […]
https://rbfirehose.com/2026/05/20/stocktonia-stockton-police-roll-out-live-ai-language-translation-software-for-body-cameras/ -
This week is going to be a bit busy and a bit tough for me.
Not in absolute terms but in terms of what I can deal with at the moment.
For everyone out there with a similar week ahead of them - "hou je taai!".
Aside - I just discovered that many online (AI powered?) translators mis-translate that Dutch phrase!
Sigh.
Time for some more coffee.
-
This week is going to be a bit busy and a bit tough for me.
Not in absolute terms but in terms of what I can deal with at the moment.
For everyone out there with a similar week ahead of them - "hou je taai!".
Aside - I just discovered that many online (AI powered?) translators mis-translate that Dutch phrase!
Sigh.
Time for some more coffee.
-
This week is going to be a bit busy and a bit tough for me.
Not in absolute terms but in terms of what I can deal with at the moment.
For everyone out there with a similar week ahead of them - "hou je taai!".
Aside - I just discovered that many online (AI powered?) translators mis-translate that Dutch phrase!
Sigh.
Time for some more coffee.
-
This week is going to be a bit busy and a bit tough for me.
Not in absolute terms but in terms of what I can deal with at the moment.
For everyone out there with a similar week ahead of them - "hou je taai!".
Aside - I just discovered that many online (AI powered?) translators mis-translate that Dutch phrase!
Sigh.
Time for some more coffee.
-
Why I refuse to use Machine Translation
In the last few years, there has been a lot of talk about how artificial intelligence (actually: commercial chatbots and LLMs) will be transforming our way of working – how it will make some jobs more efficient, and others obsolete. There are also concerns that such systems do not live up to the hype – though this has not stopped CEO and their consultants from pushing them into the workplace, in the hopes of drastically reducing their work force and labor costs even though they cannot substitute for their workers’ process knowledge.
I translate old German folk tales into English, and translation work is already heavily automated these days due to the sheer amount of material that needs to be translated. Thus, it is unsurprising that many people have asked me whether I use machine translation for my work – usually with the assumption that this would save me time.
In this essay, I am going to tell you why I won’t use AI systems for my translation work. I could talk about the ethical concerns – how the work of others is used to train LLM systems without compensation while charging for their output, or how they consume massive amounts of electricity and other resources while our planet and its ecosystems are already on the precipice, or how they are used to build up the mother of all investment bubbles.
I could also add some personal grievances. For instance, in my day job as a bid manager, I also have to price server systems for our customers, and when I recently noticed that a simple 16 GB DDR5 RAM module had a purchase price of €1,600, I realized that something is going very wrong indeed. Furthermore, anonymous bot networks are constantly scraping my websites for LLM training data, forcing me to upgrade my website hosting plan twice last fall to keep outages at a tolerable level.
But since others have elaborated on the ethical concerns in much more detail than I ever could, I won’t be talking about these further. Instead, I will be discussing the practical reasons why machine translation does not fit into my working processes when translating German folk tales.
Reading the Fraktur Typeset
The first challenge for machine translation is parsing the source material. For copyright reasons, I exclusively use public domain works – German folk tale collections which were largely published in the 19th century. And the vast majority of these works were not printed with the modern Antiqua letters, but the old German Fraktur typeset. Here is a reasonably “clean” example of a story I have translated (the source page is here):
Usually, texts that are converted into a new language by machine translation are already in a machine-readable format – but these old digital scans are not. Thus, before I could use machine translations for these texts, I would need to convert them into a machine-readable format. While OCR (“Optical Character Recognition”) tools exist that can handle Fraktur typesets, the output would require additional effort for proofreading, especially since the input data is highly variable in its quality.
Thus, in contrast to the original premise, machine translation would actually increase my workload even before I got to the actual translation step.
Translating Old Words and Phrases
LLM systems are largely trained on the most commonly available modern texts (such as Reddit posts). 19th century German folk tales are not “modern texts”. They are rife with old words and phrases that were only used in some small geographical area and are no longer in modern use. Would a standard machine translation system (i.e., one trained on Reddit) come up with a decent translation for “Bindelbaum” – to pick just one example that stuck in my mind? Especially considering that the old texts that could provide some context were not in a machine-readable format, and thus of limited use for training the LLMs?
Perhaps they could, and perhaps they couldn’t. However, “maybe this is an accurate translation” is not good enough for my purposes, and indeed, it is not sufficient for any professional translator. If I provide a translation for certain old words and prices, I need to be as sure as possible that this translation is accurate – and if I am uncertain, I need to explain that to my readers as well.
Thus, I would have to double-check every machine-translated text I work with with my own research – which, again, would not save me any time. And if I am doing all the research anyway, I might as well skip the machine translation and do it all by myself in the first place.
Providing Context
But truth to be told, the actual translation is the easiest part of my work. German folk tales were told in a specific time and a specific cultural context. The original audience for these tales (mostly 19th century German peasants) were deeply familiar with this context.
A modern audience will usually not be familiar with this context. Many aspects of these folk tales are hard to grasp even for modern Germans – so what chance does an international audience have?
This is why one of my most important tasks as a translator is to explain this context. This is why my books have many hundreds of footnotes, and explanatory commentary following each tale. While I am not primarily writing my books as scientific treatises, I have spent enough years in academia that I have views on providing inaccurate information. Sure, mistakes can and will happen. But allowing errors to proliferate in my manuscripts because I was outsourcing the most critical aspects of my research to LLM systems would be a gross violation of ethical standards (not that this seems to stop a lot of LLM users…).
So I will do my research the proper way. And with each paragraph I translate, I contemplate its hidden meanings and context, and how to convey it to my readers. But if I don’t do the first step of the work myself – that is, translating and thinking about every single sentence – then I have already lost my first opportunity to truly understand the story.
Preserving Unique Voices
German folk tales were told by tens of thousands of people, each of whom had their own unique way of telling their stories. And later on, they were collected by hundreds of folklore researchers, each of whom had their own unique editorial approach. That adds up to a lot of unique voices.
However, LLMs are well-known to generate texts that trend towards the average. They have been trained on vast archives of human-written texts, and their task is to create texts that are “most likely” to fit the prompt – the common denominator, if you will. Worse, it will be the most common denominator of Reddit users and the like. The only LLM system that might even come even close to capturing the unique voices of the original texts would be one that has been trained exclusively on their translations – including my translations.
While I want people to be entertained by my translations, these tales are also part of my country’s cultural heritage. Not even trying to capture the unique voices of these long-ago storytellers and instead replacing them with the generic output of LLMs feels hugely disrespectful.
They deserve better, and my audience deserves better as well.
#LLM #MachineTranslation #Translation -
Why I refuse to use Machine Translation
In the last few years, there has been a lot of talk about how artificial intelligence (actually: commercial chatbots and LLMs) will be transforming our way of working – how it will make some jobs more efficient, and others obsolete. There are also concerns that such systems do not live up to the hype – though this has not stopped CEO and their consultants from pushing them into the workplace, in the hopes of drastically reducing their work force and labor costs even though they cannot substitute for their workers’ process knowledge.
I translate old German folk tales into English, and translation work is already heavily automated these days due to the sheer amount of material that needs to be translated. Thus, it is unsurprising that many people have asked me whether I use machine translation for my work – usually with the assumption that this would save me time.
In this essay, I am going to tell you why I won’t use AI systems for my translation work. I could talk about the ethical concerns – how the work of others is used to train LLM systems without compensation while charging for their output, or how they consume massive amounts of electricity and other resources while our planet and its ecosystems are already on the precipice, or how they are used to build up the mother of all investment bubbles.
I could also add some personal grievances. For instance, in my day job as a bid manager, I also have to price server systems for our customers, and when I recently noticed that a simple 16 GB DDR5 RAM module had a purchase price of €1,600, I realized that something is going very wrong indeed. Furthermore, anonymous bot networks are constantly scraping my websites for LLM training data, forcing me to upgrade my website hosting plan twice last fall to keep outages at a tolerable level.
But since others have elaborated on the ethical concerns in much more detail than I ever could, I won’t be talking about these further. Instead, I will be discussing the practical reasons why machine translation does not fit into my working processes when translating German folk tales.
Reading the Fraktur Typeset
The first challenge for machine translation is parsing the source material. For copyright reasons, I exclusively use public domain works – German folk tale collections which were largely published in the 19th century. And the vast majority of these works were not printed with the modern Antiqua letters, but the old German Fraktur typeset. Here is a reasonably “clean” example of a story I have translated (the source page is here):
Usually, texts that are converted into a new language by machine translation are already in a machine-readable format – but these old digital scans are not. Thus, before I could use machine translations for these texts, I would need to convert them into a machine-readable format. While OCR (“Optical Character Recognition”) tools exist that can handle Fraktur typesets, the output would require additional effort for proofreading, especially since the input data is highly variable in its quality.
Thus, in contrast to the original premise, machine translation would actually increase my workload even before I got to the actual translation step.
Translating Old Words and Phrases
LLM systems are largely trained on the most commonly available modern texts (such as Reddit posts). 19th century German folk tales are not “modern texts”. They are rife with old words and phrases that were only used in some small geographical area and are no longer in modern use. Would a standard machine translation system (i.e., one trained on Reddit) come up with a decent translation for “Bindelbaum” – to pick just one example that stuck in my mind? Especially considering that the old texts that could provide some context were not in a machine-readable format, and thus of limited use for training the LLMs?
Perhaps they could, and perhaps they couldn’t. However, “maybe this is an accurate translation” is not good enough for my purposes, and indeed, it is not sufficient for any professional translator. If I provide a translation for certain old words and prices, I need to be as sure as possible that this translation is accurate – and if I am uncertain, I need to explain that to my readers as well.
Thus, I would have to double-check every machine-translated text I work with with my own research – which, again, would not save me any time. And if I am doing all the research anyway, I might as well skip the machine translation and do it all by myself in the first place.
Providing Context
But truth to be told, the actual translation is the easiest part of my work. German folk tales were told in a specific time and a specific cultural context. The original audience for these tales (mostly 19th century German peasants) were deeply familiar with this context.
A modern audience will usually not be familiar with this context. Many aspects of these folk tales are hard to grasp even for modern Germans – so what chance does an international audience have?
This is why one of my most important tasks as a translator is to explain this context. This is why my books have many hundreds of footnotes, and explanatory commentary following each tale. While I am not primarily writing my books as scientific treatises, I have spent enough years in academia that I have views on providing inaccurate information. Sure, mistakes can and will happen. But allowing errors to proliferate in my manuscripts because I was outsourcing the most critical aspects of my research to LLM systems would be a gross violation of ethical standards (not that this seems to stop a lot of LLM users…).
So I will do my research the proper way. And with each paragraph I translate, I contemplate its hidden meanings and context, and how to convey it to my readers. But if I don’t do the first step of the work myself – that is, translating and thinking about every single sentence – then I have already lost my first opportunity to truly understand the story.
Preserving Unique Voices
German folk tales were told by tens of thousands of people, each of whom had their own unique way of telling their stories. And later on, they were collected by hundreds of folklore researchers, each of whom had their own unique editorial approach. That adds up to a lot of unique voices.
However, LLMs are well-known to generate texts that trend towards the average. They have been trained on vast archives of human-written texts, and their task is to create texts that are “most likely” to fit the prompt – the common denominator, if you will. Worse, it will be the most common denominator of Reddit users and the like. The only LLM system that might even come even close to capturing the unique voices of the original texts would be one that has been trained exclusively on their translations – including my translations.
While I want people to be entertained by my translations, these tales are also part of my country’s cultural heritage. Not even trying to capture the unique voices of these long-ago storytellers and instead replacing them with the generic output of LLMs feels hugely disrespectful.
They deserve better, and my audience deserves better as well.
#LLM #MachineTranslation #Translation -
Why I refuse to use Machine Translation
In the last few years, there has been a lot of talk about how artificial intelligence (actually: commercial chatbots and LLMs) will be transforming our way of working – how it will make some jobs more efficient, and others obsolete. There are also concerns that such systems do not live up to the hype – though this has not stopped CEO and their consultants from pushing them into the workplace, in the hopes of drastically reducing their work force and labor costs even though they cannot substitute for their workers’ process knowledge.
I translate old German folk tales into English, and translation work is already heavily automated these days due to the sheer amount of material that needs to be translated. Thus, it is unsurprising that many people have asked me whether I use machine translation for my work – usually with the assumption that this would save me time.
In this essay, I am going to tell you why I won’t use AI systems for my translation work. I could talk about the ethical concerns – how the work of others is used to train LLM systems without compensation while charging for their output, or how they consume massive amounts of electricity and other resources while our planet and its ecosystems are already on the precipice, or how they are used to build up the mother of all investment bubbles.
I could also add some personal grievances. For instance, in my day job as a bid manager, I also have to price server systems for our customers, and when I recently noticed that a simple 16 GB DDR5 RAM module had a purchase price of €1,600, I realized that something is going very wrong indeed. Furthermore, anonymous bot networks are constantly scraping my websites for LLM training data, forcing me to upgrade my website hosting plan twice last fall to keep outages at a tolerable level.
But since others have elaborated on the ethical concerns in much more detail than I ever could, I won’t be talking about these further. Instead, I will be discussing the practical reasons why machine translation does not fit into my working processes when translating German folk tales.
Reading the Fraktur Typeset
The first challenge for machine translation is parsing the source material. For copyright reasons, I exclusively use public domain works – German folk tale collections which were largely published in the 19th century. And the vast majority of these works were not printed with the modern Antiqua letters, but the old German Fraktur typeset. Here is a reasonably “clean” example of a story I have translated (the source page is here):
Usually, texts that are converted into a new language by machine translation are already in a machine-readable format – but these old digital scans are not. Thus, before I could use machine translations for these texts, I would need to convert them into a machine-readable format. While OCR (“Optical Character Recognition”) tools exist that can handle Fraktur typesets, the output would require additional effort for proofreading, especially since the input data is highly variable in its quality.
Thus, in contrast to the original premise, machine translation would actually increase my workload even before I got to the actual translation step.
Translating Old Words and Phrases
LLM systems are largely trained on the most commonly available modern texts (such as Reddit posts). 19th century German folk tales are not “modern texts”. They are rife with old words and phrases that were only used in some small geographical area and are no longer in modern use. Would a standard machine translation system (i.e., one trained on Reddit) come up with a decent translation for “Bindelbaum” – to pick just one example that stuck in my mind? Especially considering that the old texts that could provide some context were not in a machine-readable format, and thus of limited use for training the LLMs?
Perhaps they could, and perhaps they couldn’t. However, “maybe this is an accurate translation” is not good enough for my purposes, and indeed, it is not sufficient for any professional translator. If I provide a translation for certain old words and prices, I need to be as sure as possible that this translation is accurate – and if I am uncertain, I need to explain that to my readers as well.
Thus, I would have to double-check every machine-translated text I work with with my own research – which, again, would not save me any time. And if I am doing all the research anyway, I might as well skip the machine translation and do it all by myself in the first place.
Providing Context
But truth to be told, the actual translation is the easiest part of my work. German folk tales were told in a specific time and a specific cultural context. The original audience for these tales (mostly 19th century German peasants) were deeply familiar with this context.
A modern audience will usually not be familiar with this context. Many aspects of these folk tales are hard to grasp even for modern Germans – so what chance does an international audience have?
This is why one of my most important tasks as a translator is to explain this context. This is why my books have many hundreds of footnotes, and explanatory commentary following each tale. While I am not primarily writing my books as scientific treatises, I have spent enough years in academia that I have views on providing inaccurate information. Sure, mistakes can and will happen. But allowing errors to proliferate in my manuscripts because I was outsourcing the most critical aspects of my research to LLM systems would be a gross violation of ethical standards (not that this seems to stop a lot of LLM users…).
So I will do my research the proper way. And with each paragraph I translate, I contemplate its hidden meanings and context, and how to convey it to my readers. But if I don’t do the first step of the work myself – that is, translating and thinking about every single sentence – then I have already lost my first opportunity to truly understand the story.
Preserving Unique Voices
German folk tales were told by tens of thousands of people, each of whom had their own unique way of telling their stories. And later on, they were collected by hundreds of folklore researchers, each of whom had their own unique editorial approach. That adds up to a lot of unique voices.
However, LLMs are well-known to generate texts that trend towards the average. They have been trained on vast archives of human-written texts, and their task is to create texts that are “most likely” to fit the prompt – the common denominator, if you will. Worse, it will be the most common denominator of Reddit users and the like. The only LLM system that might even come even close to capturing the unique voices of the original texts would be one that has been trained exclusively on their translations – including my translations.
While I want people to be entertained by my translations, these tales are also part of my country’s cultural heritage. Not even trying to capture the unique voices of these long-ago storytellers and instead replacing them with the generic output of LLMs feels hugely disrespectful.
They deserve better, and my audience deserves better as well.
#LLM #MachineTranslation #Translation -
Why I refuse to use Machine Translation
In the last few years, there has been a lot of talk about how artificial intelligence (actually: commercial chatbots and LLMs) will be transforming our way of working – how it will make some jobs more efficient, and others obsolete. There are also concerns that such systems do not live up to the hype – though this has not stopped CEO and their consultants from pushing them into the workplace, in the hopes of drastically reducing their work force and labor costs even though they cannot substitute for their workers’ process knowledge.
I translate old German folk tales into English, and translation work is already heavily automated these days due to the sheer amount of material that needs to be translated. Thus, it is unsurprising that many people have asked me whether I use machine translation for my work – usually with the assumption that this would save me time.
In this essay, I am going to tell you why I won’t use AI systems for my translation work. I could talk about the ethical concerns – how the work of others is used to train LLM systems without compensation while charging for their output, or how they consume massive amounts of electricity and other resources while our planet and its ecosystems are already on the precipice, or how they are used to build up the mother of all investment bubbles.
I could also add some personal grievances. For instance, in my day job as a bid manager, I also have to price server systems for our customers, and when I recently noticed that a simple 16 GB DDR5 RAM module had a purchase price of €1,600, I realized that something is going very wrong indeed. Furthermore, anonymous bot networks are constantly scraping my websites for LLM training data, forcing me to upgrade my website hosting plan twice last fall to keep outages at a tolerable level.
But since others have elaborated on the ethical concerns in much more detail than I ever could, I won’t be talking about these further. Instead, I will be discussing the practical reasons why machine translation does not fit into my working processes when translating German folk tales.
Reading the Fraktur Typeset
The first challenge for machine translation is parsing the source material. For copyright reasons, I exclusively use public domain works – German folk tale collections which were largely published in the 19th century. And the vast majority of these works were not printed with the modern Antiqua letters, but the old German Fraktur typeset. Here is a reasonably “clean” example of a story I have translated (the source page is here):
Usually, texts that are converted into a new language by machine translation are already in a machine-readable format – but these old digital scans are not. Thus, before I could use machine translations for these texts, I would need to convert them into a machine-readable format. While OCR (“Optical Character Recognition”) tools exist that can handle Fraktur typesets, the output would require additional effort for proofreading, especially since the input data is highly variable in its quality.
Thus, in contrast to the original premise, machine translation would actually increase my workload even before I got to the actual translation step.
Translating Old Words and Phrases
LLM systems are largely trained on the most commonly available modern texts (such as Reddit posts). 19th century German folk tales are not “modern texts”. They are rife with old words and phrases that were only used in some small geographical area and are no longer in modern use. Would a standard machine translation system (i.e., one trained on Reddit) come up with a decent translation for “Bindelbaum” – to pick just one example that stuck in my mind? Especially considering that the old texts that could provide some context were not in a machine-readable format, and thus of limited use for training the LLMs?
Perhaps they could, and perhaps they couldn’t. However, “maybe this is an accurate translation” is not good enough for my purposes, and indeed, it is not sufficient for any professional translator. If I provide a translation for certain old words and prices, I need to be as sure as possible that this translation is accurate – and if I am uncertain, I need to explain that to my readers as well.
Thus, I would have to double-check every machine-translated text I work with with my own research – which, again, would not save me any time. And if I am doing all the research anyway, I might as well skip the machine translation and do it all by myself in the first place.
Providing Context
But truth to be told, the actual translation is the easiest part of my work. German folk tales were told in a specific time and a specific cultural context. The original audience for these tales (mostly 19th century German peasants) were deeply familiar with this context.
A modern audience will usually not be familiar with this context. Many aspects of these folk tales are hard to grasp even for modern Germans – so what chance does an international audience have?
This is why one of my most important tasks as a translator is to explain this context. This is why my books have many hundreds of footnotes, and explanatory commentary following each tale. While I am not primarily writing my books as scientific treatises, I have spent enough years in academia that I have views on providing inaccurate information. Sure, mistakes can and will happen. But allowing errors to proliferate in my manuscripts because I was outsourcing the most critical aspects of my research to LLM systems would be a gross violation of ethical standards (not that this seems to stop a lot of LLM users…).
So I will do my research the proper way. And with each paragraph I translate, I contemplate its hidden meanings and context, and how to convey it to my readers. But if I don’t do the first step of the work myself – that is, translating and thinking about every single sentence – then I have already lost my first opportunity to truly understand the story.
Preserving Unique Voices
German folk tales were told by tens of thousands of people, each of whom had their own unique way of telling their stories. And later on, they were collected by hundreds of folklore researchers, each of whom had their own unique editorial approach. That adds up to a lot of unique voices.
However, LLMs are well-known to generate texts that trend towards the average. They have been trained on vast archives of human-written texts, and their task is to create texts that are “most likely” to fit the prompt – the common denominator, if you will. Worse, it will be the most common denominator of Reddit users and the like. The only LLM system that might even come even close to capturing the unique voices of the original texts would be one that has been trained exclusively on their translations – including my translations.
While I want people to be entertained by my translations, these tales are also part of my country’s cultural heritage. Not even trying to capture the unique voices of these long-ago storytellers and instead replacing them with the generic output of LLMs feels hugely disrespectful.
They deserve better, and my audience deserves better as well.
#LLM #MachineTranslation #Translation -
PCMag: Google Brings Real-Time Headphone Translation to iOS. “The Gemini-powered feature made its Android debut in December and lets you hear translations in real time through your headphones. To give it a try, launch the Translate app, select the desired languages, connect your headphones, and tap the Live Translate button at the bottom left of the home page.”
https://rbfirehose.com/2026/03/27/pcmag-google-brings-real-time-headphone-translation-to-ios/ -
PCMag: Google Brings Real-Time Headphone Translation to iOS. “The Gemini-powered feature made its Android debut in December and lets you hear translations in real time through your headphones. To give it a try, launch the Translate app, select the desired languages, connect your headphones, and tap the Live Translate button at the bottom left of the home page.”
https://rbfirehose.com/2026/03/27/pcmag-google-brings-real-time-headphone-translation-to-ios/ -
PCMag: Google Brings Real-Time Headphone Translation to iOS. “The Gemini-powered feature made its Android debut in December and lets you hear translations in real time through your headphones. To give it a try, launch the Translate app, select the desired languages, connect your headphones, and tap the Live Translate button at the bottom left of the home page.”
https://rbfirehose.com/2026/03/27/pcmag-google-brings-real-time-headphone-translation-to-ios/ -
PCMag: Google Brings Real-Time Headphone Translation to iOS. “The Gemini-powered feature made its Android debut in December and lets you hear translations in real time through your headphones. To give it a try, launch the Translate app, select the desired languages, connect your headphones, and tap the Live Translate button at the bottom left of the home page.”
https://rbfirehose.com/2026/03/27/pcmag-google-brings-real-time-headphone-translation-to-ios/ -
Techdirt: Greater Than Zero: The Anti-AI Pushback On Gaming Preservation Efforts Makes No Sense. “I’m not some AI evangelist. I fully recognize that there are error and other problems with AI… and I imagine there always will be, to some extent. AI is not always, or perhaps even mostly, the right tool to use. Nor will it always have benefits that outweigh problems it creates for we human […]
https://rbfirehose.com/2026/03/23/greater-than-zero-the-anti-ai-pushback-on-gaming-preservation-efforts-makes-no-sense-techdirt/ -
Techdirt: Greater Than Zero: The Anti-AI Pushback On Gaming Preservation Efforts Makes No Sense. “I’m not some AI evangelist. I fully recognize that there are error and other problems with AI… and I imagine there always will be, to some extent. AI is not always, or perhaps even mostly, the right tool to use. Nor will it always have benefits that outweigh problems it creates for we human […]
https://rbfirehose.com/2026/03/23/greater-than-zero-the-anti-ai-pushback-on-gaming-preservation-efforts-makes-no-sense-techdirt/ -
Techdirt: Greater Than Zero: The Anti-AI Pushback On Gaming Preservation Efforts Makes No Sense. “I’m not some AI evangelist. I fully recognize that there are error and other problems with AI… and I imagine there always will be, to some extent. AI is not always, or perhaps even mostly, the right tool to use. Nor will it always have benefits that outweigh problems it creates for we human […]
https://rbfirehose.com/2026/03/23/greater-than-zero-the-anti-ai-pushback-on-gaming-preservation-efforts-makes-no-sense-techdirt/ -
Techdirt: Greater Than Zero: The Anti-AI Pushback On Gaming Preservation Efforts Makes No Sense. “I’m not some AI evangelist. I fully recognize that there are error and other problems with AI… and I imagine there always will be, to some extent. AI is not always, or perhaps even mostly, the right tool to use. Nor will it always have benefits that outweigh problems it creates for we human […]
https://rbfirehose.com/2026/03/23/greater-than-zero-the-anti-ai-pushback-on-gaming-preservation-efforts-makes-no-sense-techdirt/ -
Ars Technica: Kagi Translate’s AI answers the question “What would horny Margaret Thatcher say?”. “This week, many people across the Internet have been bemused to find that the AI-powered Kagi Translate can perform these and countless other unlikely ‘translation’ tasks. And while the collective discovery highlights the playful, creative side of large language models, it also exposes the […]
https://rbfirehose.com/2026/03/20/ars-technica-kagi-translates-ai-answers-the-question-what-would-horny-margaret-thatcher-say/ -
Ars Technica: Kagi Translate’s AI answers the question “What would horny Margaret Thatcher say?”. “This week, many people across the Internet have been bemused to find that the AI-powered Kagi Translate can perform these and countless other unlikely ‘translation’ tasks. And while the collective discovery highlights the playful, creative side of large language models, it also exposes the […]
https://rbfirehose.com/2026/03/20/ars-technica-kagi-translates-ai-answers-the-question-what-would-horny-margaret-thatcher-say/ -
Ars Technica: Kagi Translate’s AI answers the question “What would horny Margaret Thatcher say?”. “This week, many people across the Internet have been bemused to find that the AI-powered Kagi Translate can perform these and countless other unlikely ‘translation’ tasks. And while the collective discovery highlights the playful, creative side of large language models, it also exposes the […]
https://rbfirehose.com/2026/03/20/ars-technica-kagi-translates-ai-answers-the-question-what-would-horny-margaret-thatcher-say/ -
Ich hasse maschinell übersetzte Supportwebseiten großer Unternehmen! Wie sollen deutschsprachige Kund_innen von Osprey bei der Wahrnehmung von deren “all mighty guarantee” denn herausbekommen, dass mit “Ausstellungsort” eigentlich “Problemstelle am Produkt” und mit "Name des Pakets” eigentlich der “Produktname des Rucksacks” gemeint ist?
-
Ich hasse maschinell übersetzte Supportwebseiten großer Unternehmen! Wie sollen deutschsprachige Kund_innen von Osprey bei der Wahrnehmung von deren “all mighty guarantee” denn herausbekommen, dass mit “Ausstellungsort” eigentlich “Problemstelle am Produkt” und mit "Name des Pakets” eigentlich der “Produktname des Rucksacks” gemeint ist?
-
Ich hasse maschinell übersetzte Supportwebseiten großer Unternehmen! Wie sollen deutschsprachige Kund_innen von Osprey bei der Wahrnehmung von deren “all mighty guarantee” denn herausbekommen, dass mit “Ausstellungsort” eigentlich “Problemstelle am Produkt” und mit "Name des Pakets” eigentlich der “Produktname des Rucksacks” gemeint ist?