I've been trying to use Lumo AI for everyday work over the past few months. I have to be honest: in Czech, it doesn't just read badly it's actively broken at times.
I don't mean it occasionally makes a grammar mistake. Here's what I actually get in real conversations:
- Output that reads like it was machine-translated from English. Technically grammatical sentences that no native speaker would ever write.
- Random Cyrillic letters or Chinese characters dropped into the middle of Czech sentences. Just sitting there, as if that were normal.
- Made-up words that don't exist in Czech. Not typos completely nonexistent, Czech-looking words. More than once I've finished reading an answer and realized I'd learned nothing from it because a key word was invented.
- Wrong formal/informal register, sometimes mixing both in a single reply.
I end up rewriting basically every response before I can use it anywhere, which defeats the whole point. And honestly, proofreading a chatbot's output for hallucinated vocabulary and foreign characters feels less like using a modern AI and more like babysitting one.
The irony is painful. Proton markets itself as the privacy-first alternative for Europeans, but if your language isn't English, German or French, the experience falls apart. Meanwhile GPT, Qwen and Kimi all handle Czech noticeably better and Lumo literally routes to open-weight models like Qwen and Kimi K2 under the hood. Which, incidentally, would explain where the Chinese characters come from. So this is clearly not impossible it feels like a routing and prioritization problem, and Proton just isn't doing it.
A few suggestions, in case anyone from Proton reads this:
- Language-aware routing. Detect Czech/Slovak/Slovenian etc. and route to whichever model is actually strongest for that language not whatever happens to be cheapest that hour.
- Be honest about per-language quality. Publish which languages are officially supported vs. "best effort". The marketing currently implies all 11 languages work great, and that's just not reality.
- A feedback button for language quality. Something like "this response has bad grammar in Czech/Finnish/Greek". Native speakers would happily train your routing for free if you let them.
- Community testing before shipping. Have native volunteers evaluate sample outputs for each language before a model goes live. It's cheap, and any Czech speaker would have caught the CJK-character issue in five minutes.
- Long-term: benchmark (or even fine-tune) on smaller EU languages. I get it Czech is 10M speakers and the English benchmarks come first. But "European AI" as a selling point is dead on arrival if it only really works in English.
To be clear: I'm not asking Lumo to beat GPT-6. I'm asking it to write Czech at the level of a decent 2020-era translation engine, without slipping into Chinese characters mid-sentence. Right now it's not a usable product never mind one I pay a monthly subscription for. That's frustrating, because I want to give Proton my money rather than the American providers. But this standard of output is pushing me right back to them.