HomeBlog › Fact-checking

How do I check a claim in a language I do not speak?

Language barriers are the blind spot of online verification. Learn how to trace and verify non-English claims using AI tools and primary source evidence.

Fact-checking · · 7 min read · 1,569 words

W
wyper Fact-Check Team Builders of the wyper Fact-Check extension. Every claim in this article links to its source.
AI transparency

This article was researched and drafted with AI, then checked and released by a person before it went live. The illustration is AI-generated. Every factual claim links to its source, so you can verify it yourself.

A split monitor screen displaying a social media feed in one alphabet on the left and a highlighted official document in a different alphabet on the right. (AI-generated illustration)

AI-generated illustration. A split monitor screen displaying a social media feed in one alphabet on the left and a highlighted official document in a different alphabet on the right.

The short answer

To check a claim in a language you do not speak, you must trace the text back to its original cultural and linguistic root rather than relying on translated summaries. Automated tools now use cross-lingual reasoning to fetch non-English primary sources, read them in their native context, and deliver the evidence chain in your own language.

When you set out to fact-check foreign language claim origins, you will quickly notice that the language barrier is the hardest part of the process. The internet operates globally, but our verification habits remain local. A post claiming a specific event happened in another country often circulates in English, yet the primary evidence required to verify it exists entirely in a different language.

Relying on English-language search results to verify a claim from a non-English region creates a severe blind spot. If you only search in your native language, you are entirely dependent on secondary reporting, translated summaries, or international news agencies deciding the story is worth covering. To get to the truth, you have to reach the root source.

The blind spot of verification tools

For years, the automated analysis of digital claims has focused almost exclusively on English and a few other high-resource languages. This leaves massive gaps in our ability to combat organized manipulation. Manipulative information campaigns specifically target population groups with different cultural backgrounds and often make use of their native languages, as researchers at the FZI Forschungszentrum Informatik noted. The FZI launched the KuKI project to develop systems capable of identifying disinformation across German, Russian, and Turkish, explicitly because existing systems overlook false narratives when they are not in English.

Large language models have shown promise in automating verification, but their effectiveness across global contexts remains highly uneven. A 2026 study published in Scientific Reports evaluated nine established models using 5,000 claims previously assessed by professional fact-checking organizations across 47 languages. The researchers found a concerning pattern resembling the Dunning-Kruger effect: models frequently overestimated their capabilities in multilingual contexts, particularly smaller models.

This inequality is a structural problem in how models are trained and deployed. Research from Hong Kong Baptist University analyzed cross-language inequality in automated fact-checking across nine languages and six language families. They found substantial cross-language disparities, proving that a tool performing well in English might fail entirely when asked to verify a claim originating in Arabic or Hindi.

Moving beyond simple translation

The intuitive approach for most internet users is to translate a foreign post into English using a browser extension, and then paste that translation into a search engine. This approach routinely fails. Translating a claim strips away local context, regional terminology, and cultural nuances that are essential for a successful search query.

Instead of translating the claim to search in English, effective verification requires cross-lingual reasoning. The system must understand the claim in the user's language, formulate a highly specific search query in the source language, retrieve the local documents, read them in their native context, and finally synthesize the evidence back into the user's language.

Computer scientists are actively building frameworks to solve this. A study in the Journal of Intelligent Information Systems proposed a Multilingual and Multimodal Retrieval-Augmented Generation framework to verify news claims in low-resource settings. Similarly, research published on Zenodo demonstrated a system tailored for Indic languages using a fine-tuned transformer model to process English, Hindi, Tamil, and Bengali simultaneously. Another framework, detailed in the AI and Tech in Behavioral and Social Sciences journal, integrates multilingual representations and claim-evidence alignment features to verify facts in environments with limited labeled data.

Reaching the root source

To test whether cross-lingual reasoning actually works in practice, we measured our own systems. In a test of 24 claims where the root source was named in advance, our engine successfully reached the primary sources 63 percent of the time. Legacy media accounted for 14 percent of the retrieved evidence, and fact-checkers made up just 1 percent.

Most importantly, all six non-English roots in the test set were successfully reached. The system retrieved original documentation from Russian authorities, rulings from the Mexican Supreme Court, and local Thai sources. The tool bypassed the English-language media filter completely and pulled the primary evidence directly from the countries of origin.

Sources are never filtered by origin in our system. There is no mainstream bonus, and official sources are not automatically preferred over local independent reporting. The system retrieves the evidence, presents the chain of linked sources, and the reader decides what to trust.

How the architecture handles foreign text

When you submit a claim, the standard check runs a Dual-AI cross-check using Gemini and Grok. Grok searches for itself during this process, meaning the two models do not share one evidence set. They independently formulate queries, retrieve documents in the required languages, and compare their findings.

For a more rigorous deepening, the Full Spectrum check introduces a Triple-AI reading: three models, one evidence set. The third voice does not search the web. It receives exactly the same foreign language sources the main model received and reads them independently. If there is a disagreement in this phase, it is a different reading of the exact same material, not a different search result.

Because video claims often cross borders faster than text, handling audiovisual material is critical. When reachable, the system uses subtitles to read the video content. When subtitles are not available, an AI model actually watches the video within a specific time window. The free tier covers up to 10 minutes on YouTube, while Pro covers 30 minutes. For X videos, the limit is up to 2 minutes free and 7 minutes Pro, with longer videos refused outright rather than partially checked.

Regardless of the language the original source was written or spoken in, the output is delivered in your preferred language. The system supports 12 product languages, returning a truth score of 1-10 alongside an evidence chain of real, linked sources. We never call the truth score proof of trustworthiness, and the tool does not judge the intent of the original poster. Both of those judgments stay firmly with you, the reader.

Comparing your options

If you want to verify a foreign claim yourself, you can use the free wyper web app in a separate tab, or you can run the browser extension directly over the post you are reading.

Manual checkswyper web appwyper extension
Free to performFree tier availableFree tier available
Requires no installationRequires no installationRequires browser installation
User must translate queriesAuto-translates and searchesAuto-translates and searches
User copies text between tabsUser pastes text into the appReads text directly from active page
No data leaves the browserData sent for verificationData sent for verification

Who needs which: Manual checking is for users who want total control over their search history and prefer not to send data to any external tool. The web app is designed for people who want automated cross-lingual search capabilities without installing anything on their device. The browser extension is for users who frequently encounter foreign language claims in their social media feeds and want to check them instantly without switching tabs.

Can AI models understand low-resource languages?
Yes, but their performance varies significantly depending on the specific language and the model used. Major models handle high-resource languages like Spanish or German fluently, but often struggle with regional dialects or languages with limited digital footprints. Cross-lingual verification tools mitigate this by routing queries through models specifically trained on diverse linguistic datasets.
Does translating a post change its factual meaning?
Translating a post often strips away critical cultural context and regional idioms, which can alter the perceived factual meaning. A literal translation might miss sarcasm, local political references, or specific legal terminology. Effective fact-checking requires reading the claim and its supporting evidence in the original language context before summarizing the findings.
How do you find original non-English sources?
You find original non-English sources by identifying the geographic or cultural origin of the claim and running search queries in that specific local language. Automated tools achieve this by using cross-lingual reasoning to translate your English prompt into the target language, querying local search engines, and retrieving the native documents directly.
Does the wyper truth score apply to the translation or the original?
The truth score applies to the factual core of the original claim based on the primary evidence retrieved. The system evaluates the original non-English documents in their native context to determine the accuracy of the event or statement. The resulting 1-10 score and evidence chain are then presented to you in your chosen product language.
Can I check foreign claims without installing an extension?
You can check foreign claims completely in your browser without installing any software by using a dedicated web application. You simply copy the foreign text or the link to the post, paste it into the application, and the system handles the cross-lingual search and verification remotely.

The short version

Fact-checking a claim from another country requires reaching primary sources in their native language, as translating the claim and searching in English often completely misses the evidence. Modern verification tools bridge this gap by using cross-lingual reasoning to formulate local search queries, read foreign documents in context, and deliver the evidence chain in your language. You can automate this process using either a standalone web application or a browser extension to fetch the root sources directly.

Useful? Pass it on: 𝕏 Post it Telegram

Check posts where you read them

The wyper Fact-Check extension runs these checks on the post itself: truth score, evidence chain, and the date gap that catches recycled footage.

The web app installs to your home screen in one tap. No store, no account.
Already using it? ★★★★★ Rate it on the Chrome Web Store. Reviews decide what others get shown first.

Keep reading

Fact-checking · Two vs three AI models for fact-checking? Why two AI models usually agree on a claim, and how adding a third model with different origins creates genuine contradiction in fact-checking. Social Cleanup · Does Deleting a Tweet Remove It From Google? Deleting a post on X removes it from your profile, but Google caches often keep it visible. Learn how to clear deleted posts from search results entirely.