מה יכול בודק זה - ולא יכול - לזהות

לכל בודק ספרותי יש נקודות מתות, רובם קוברים אותם, הדף הזה שלנו, במלואו, כי דו"ח שלא ניתן לחקור הוא דו"ח שאי אפשר לסמוך עליו.

איך המנוע עובד

ראשית, החזרה: אנו בוחנים רצפי מילים ייחודיים של 8-12 מכל אזור בטקסט - מעדיפים מילים נדירות, אשר מחזירות טובות בהרבה מפותחני פתיחה נפוצים - ומחפשות כל אחד מהם ברשת החיה כביטוי מצוטט במדויק. דבר זה יוצר סט של עמודי מועמד.

שנית, אנו מורידים כל דף מועמד ומסדרים כל משפט של הטקסט שלך כנגד התוכן הממשי של הדף, באופן מקומי. עקיפת מילים תחת מילה זוהתה עם השוואה של חלון מילים מתגלגל; parasessade עם חפיפה ואסימון וסימון של סף הסימון. מנוע חיפוש לא מתייחס לעולם כראיה בעצמו - רק כספירת השוואה מאומתת טקסט, וכל התאמה נושאת דמות ביטחון בדו"ח.

לפני כל זה, הטקסט שלך הוא נורמלי: תווים דמויי Unicode (טריק סטנדרטי לדמקה) קרסו בחזרה לשווי ערך רגיל שלהם, כך אותיות קיריליות מוחלפים לא להסתיר גפרורים.

כאשר נפספס גפרורים (שלילים כוזבים)

  • כתבי עת אקדמיים, ארכיון חדשות מאחורי הכניסה.
  • מסדי נתונים פרטיים - כולל ארכיוני הגשה אקדמיים כמו של טורנטין. אף כלי ציבורי אינו יכול לראות אותם; כל דמקה חופשית המרמזת אחרת משקרת לך.
  • מקורות לא מקוונים: ספרים ומסמכים מודפסים שמעולם לא פורסמו ברשת.
  • תוכן טרי מאוד: דפים שפורסמו לפני דקות או שעות שמנועי החיפוש עדיין לא אינדקס.
  • פאראזה כבד: שכתוב שמשתנה רוב המילים נופל מתחת לסף ההתאמה. זיהוי רעיונות (מנוסח) הוא מעבר לכל התאמת טקסט.

כאשר אנו דגל טקסט חף מפשע (חיוביות כוזבות)

  • בדוק את הציטוט ולא את נקודת השיא.
  • ביטויים נפוצים ולוחית דו-צדדית: ביטויי מניות, נוסחה משפטית, תיאורים מתודולוגיים שחוזרים על עצמם בכל תחום.
  • Bibliographies and repression lists, which performally match the works they recript.
  • ניסוח מקרי במשפטים קצרים ועובדותיים.

זו הסיבה שהדו"ח מראה את הטקסט המקורי התואם לצדך, משפט אחר משפט, עם התאמות מדויקות וקרובות מסומנות בנפרד: הכלי מוצא חפיפה; אדם שופט את משמעותו.

מה משמעות הציון?

אחוזי הצפייה השווים הם חלק מהמילים שלך שיושבות במשפטים מיושר למקור משוחזר. להקות המדמד - 0- 95% נראות מקוריות, 6-20% גפרורים מסוימים, 21% + משמעותיים - הן מכוילות הדרכה עבור פרוזה טיפוסית, לא ספסות האשמה. דו"ח 4% עדיין יכול להכיל סעיף אחד שהועתק במלואו שווה תיקון; דו"ח של חומר מצוטט כראוי יכול להיות עבודה ישרה לחלוטין.

אנו גם אומרים לכם כמה מקורות נבדקו עבור הצ'ק שלכם - המספר האמיתי, מודפס על הדו"ח, כי תביעות כיסוי שאתם לא יכולים לאמת הן שיווק, לא דיוק.

Results are indicative, not conclusive. We compare your text against publicly accessible web pages at the moment you run the check. We cannot detect matches in sources that are offline, paywalled, unindexed, or held in private databases (including academic submission archives). Common phrases, correctly quoted material, and coincidental wording can appear as matches. Use this report as a guide for review and citation - not as standalone proof that text was or was not plagiarized.

שאלות לעתים קרובות

Two phases. First we sample distinctive word sequences from your text and search them on the live web as exact phrases. Then we download every candidate page and align every sentence of your text against the page’s real content locally - exact matches by rolling word-window comparison, near matches by token overlap and edit similarity. Search results alone are never trusted as matches; only verified text comparison counts.

When the source is not on the public web at check time: paywalled journals, offline books, private submission archives, pages published minutes ago that search engines have not indexed, or content behind logins. Heavily paraphrased text can also fall below the near-match threshold. No web-based checker escapes these limits; we would rather tell you than let a 0% mislead you.

Correctly quoted passages, common stock phrases, technical boilerplate, legal or religious formulae, and bibliographies all legitimately match their sources. The report separates exact from near matches and shows the source text beside yours so a human can make the call - the score is an instrument reading, not a judgment.

It is the share of your words that sit in sentences we matched to a retrieved source: matched-sentence words divided by total words. The gauge bands are 0-5% "looks original", 6-20% "some matches - review citations", 21%+ "significant matching". The bands are review guidance, not accusation thresholds.

The interface runs in 100+ languages, and the engine itself is multilingual: sentence segmentation handles Latin, CJK and Arabic punctuation, and retrieval uses engines with strong non-English coverage. Match quality is best in languages with a large public web footprint; for very small languages, coverage is honestly thinner.