װאָס דער דורכקוקער קען און קען ניט אױספֿירן

אױף דער װײַט־װײַט

ווי דער אָפּעראַציע־מאַשין אַרבעט

דער דורכקוק פֿאָרט זיך אין צוויי פֿעיִקייטן. ערשטער, אָפּזוך: מיר אַרײַנשטעלן אַ סעלעקציע פֿון 8-12 וואָרט־פֿאַרפֿילונגען פֿון יעדער געגנט פֿון דיין טעקסט — מיר פֿאָרשלאָגן אַ סך ווײַטערדיקע וואָרטן, וואָס אַרײַנשטעלן זיך בעסער ווי די װײַטערדיקע אָנהײב־װערטער — און סעלעקציע פֿון יעדער פֿון זיי אין דער װײַב ווי אַ װײַטערדיקע ציטירטע שפּראַך. דאָס פּראָדוצירט אַ סעלעקציע פֿון קאָנפֿערענצן־װײַזן.

װידער, באַשטעטיקונג: מיר אַרײַנשטעלן יעדער קעגנער־װײַז און אַרײַנפֿירן יעדער שטיקל פֿון דיין טעקסט קעגן דעם װײַז ס׳האָט זיך אַרײַנגעלייגט, װײַל וואָרט־פֿאַר־װאָרט איבערשרײַבונג איז געפֿונען מיט אַ װײַז־װײַז־צוצוגײן; צוצוגײן אַ װײַז־פֿאַר־װײַז מיט אַ װײַז־פֿאַר־װײַז און אַ רעדאַגירן־פֿאַר־פֿ

פֿאָרױסװײַז פֿונעם טעקסט

ווען מיר וועלן ניט געפֿינען װאָס צוצופֿירן (פֿאַלש־ניגאַטיווע)

  • װידער אַ מאָל, װידער,
  • פּריוואַטע דאַטאַבאַסען - אַרײַנגעזעצט אקאדעמישע אונטערשטעכנדיקע ארקיװן װי טורניטינס. קײן װעלטלעכער מכשיר קען זיי ניט זען; װידער אַ פֿרײַער צװײטע־װײַזער װאָס װײַזט אױף אַ אַנדערש איז אַ לױג צו דיר
  • אפֿשר איז עס נישט קיין אמת, אָבער עס איז אַ אמת, אַז די װײַב האָט ניט געקענט װײַבן.
  • שריפֿט־פֿאַרב:
  • װײַז פֿאָרױסװײַז פֿונעם טעקסט

ווען מיר וועלן פֿלאַגן אומגעריכטע טעקסטן (פֿאַלשע פּאָזיטיוון)

  • קלײַב אַלץ אױסselect-action
  • קלײן
  • ביבליאָגראַפֿיעס און רעפֿערענץ־ליסטע, װאָס װײַזן אױף נאַטירליכע אופֿן די אַרבעטן װאָס זיי ציטירן.
  • די װערטער װעלן װערן װידער געװען קלײנע, אמתע שורות.

דאָס איז דער סיבה װאָס דער באַריכט װײַזט דעם צוגעפֿאַלענעם מקור־טעקסט צוזאַמען מיט אייער, שורה־דורך־שורה, מיט װידערגעפֿאַלענע און צוגעפֿאַלענע שריפֿטצײכן אױסגעלוקט: דער מכשיר געפֿינט איבערהײלונג; אַ מענטש װערט אױסגעפֿאַלן װי עס מיינט

װאָס די שאַץ מײנט

טעקסט פֿאַרבtext-tool-action

מיר זאגן אויך צו איר וויפיל מקורים זענען אמת-האמת אנגעקוקט געווארן פאר אייער צאָלונג - די אמת-האמת נומער, ארויסגעדרוקט אויף דעם באריכט, ווייל די דעקאַפּמענט קאלאגראפיע וואס איר קענט נישט באשטימען איז א מאַרקעט, נישט א ריכטיגקייט.

Results are indicative, not conclusive. We compare your text against publicly accessible web pages at the moment you run the check. We cannot detect matches in sources that are offline, paywalled, unindexed, or held in private databases (including academic submission archives). Common phrases, correctly quoted material, and coincidental wording can appear as matches. Use this report as a guide for review and citation - not as standalone proof that text was or was not plagiarized.

פֿראַגעס

Two phases. First we sample distinctive word sequences from your text and search them on the live web as exact phrases. Then we download every candidate page and align every sentence of your text against the page’s real content locally - exact matches by rolling word-window comparison, near matches by token overlap and edit similarity. Search results alone are never trusted as matches; only verified text comparison counts.

When the source is not on the public web at check time: paywalled journals, offline books, private submission archives, pages published minutes ago that search engines have not indexed, or content behind logins. Heavily paraphrased text can also fall below the near-match threshold. No web-based checker escapes these limits; we would rather tell you than let a 0% mislead you.

Correctly quoted passages, common stock phrases, technical boilerplate, legal or religious formulae, and bibliographies all legitimately match their sources. The report separates exact from near matches and shows the source text beside yours so a human can make the call - the score is an instrument reading, not a judgment.

It is the share of your words that sit in sentences we matched to a retrieved source: matched-sentence words divided by total words. The gauge bands are 0-5% "looks original", 6-20% "some matches - review citations", 21%+ "significant matching". The bands are review guidance, not accusation thresholds.

The interface runs in 100+ languages, and the engine itself is multilingual: sentence segmentation handles Latin, CJK and Arabic punctuation, and retrieval uses engines with strong non-English coverage. Match quality is best in languages with a large public web footprint; for very small languages, coverage is honestly thinner.