이 검사기가 감지할 수 있는 것과 감지할 수 없는 것

모든 plagiarist 검사기는 맹점이 있습니다. 대부분은 그들을 묻습니다. 이 페이지는 우리의, 전체, 왜냐하면 당신이 질문할 수없는 보고서는 신뢰할 수없는 보고서입니다.

엔진의 작동 방식

검사는 두 단계로 진행됩니다. 첫째, 검색: 우리는 텍스트의 모든 지역에서 8-12 단어의 독특한 시퀀스를 샘플링합니다 - 흔한 오프닝보다 훨씬 더 잘 검색되는 희귀한 단어를 선호합니다 - 그리고 라이브 웹에서 각 단어를 정확한 인용 구문으로 검색합니다. 이는 후보 페이지의 집합을 생성합니다.

두 번째, 검증: 각 후보 페이지를 다운로드하고 페이지의 실제 내용에 대한 텍스트의 모든 문장을 로컬로 정렬합니다. 단어별 오버랩은 롤링 단어 창 비교로 감지되며, 토큰 오버랩 및 편집 유사성 임계값으로 닫힌 변형을 감지합니다. 검색 엔진 순위는 절대로 페이지 자체를 증거로 취급하지 않습니다. 검증된 텍스트 비교만 계산되며 각 일치 사항은 보고서에서 신뢰도 숫자를 갖습니다.

이것들을 하기 전에, 텍스트는 정상화됩니다: 유니코드 닮은 글자(체커를 속이기 위한 표준 트릭)는 그들의 평범한 대응으로 다시 축소되므로, 교환된 키릴 문자는 일치하는 것을 숨기지 않습니다.

우리가 일치를 놓칠 때 (거짓 음성)

  • 유료 및 구독 소스: 학술 저널, 로그인 뒤에 뉴스 아카이브.
  • 개인 데이터베이스 - Turnitin과 같은 학술 제출 아카이브를 포함. 어떤 공공 도구도 그들을 볼 수 없습니다; 그렇지 않은 무료 체크어를 암시하는 것은 거짓말입니다.
  • 오프라인 출처: 웹에 올리지 않은 인쇄된 책과 논문.
  • 매우 신선한 콘텐츠: 검색엔진이 아직 색인을 작성하지 않은 분 또는 시간 전에 게시된 페이지입니다.
  • 글자 크기를 줄이는 법 : 글자 크기를 줄이는 방법은 글자 크기를 줄이는 방법과 같습니다.

우리가 무고한 텍스트 (거짓 긍정)를 플래그 할 때

  • 올바르게 인용된 자료 - 인용은 그것의 출처와 일치해야 합니다. 인용을 검사하십시오, 강조하지 않습니다.
  • 일반적인 구문과 보일러 플레이트: 주식 표현, 법률 공식, 방법론적 설명 전체 분야에 반복.
  • 참고문헌과 참조 목록, 자연스럽게 그들이 인용 작품과 일치.
  • 짧은 문장에 우연한 표현, 사실 문장.

이것이 보고서가 정확한 일치와 가까운 일치가 별도로 표시된 문장별로 귀하의 텍스트 옆에 일치하는 원본 텍스트를 표시하는 이유입니다. 도구는 중복을 찾아내고 사람은 그것이 무엇을 의미하는지 판단합니다.

점수가 의미하는 바

일치하는 비율은 검색된 출처에 정렬된 문장에 앉아있는 단어의 비율입니다. 게이지 밴드 - 0-5% 원래 보인다, 6-20% 일부 일치, 21% + 중요한 일치 - 비난 임계치가 아닌 일반적인 산문에 대한 보정 검토 지침입니다. 4% 보고서는 여전히 수정할 가치가있는 완전히 복사 한 단락을 포함 할 수 있습니다; 적절하게 인용 자료의 25% 보고서는 완벽하게 정직한 작품이 될 수 있습니다.

우리는 또한 얼마나 많은 소스가 실제로 검사를 위해 검사되었는지 알려줍니다 - 실제 숫자, 보고서에 인쇄, 범위 주장을 확인할 수 없기 때문에 마케팅, 정확성이 아닙니다.

Results are indicative, not conclusive. We compare your text against publicly accessible web pages at the moment you run the check. We cannot detect matches in sources that are offline, paywalled, unindexed, or held in private databases (including academic submission archives). Common phrases, correctly quoted material, and coincidental wording can appear as matches. Use this report as a guide for review and citation - not as standalone proof that text was or was not plagiarized.

자주 묻는 질문

Two phases. First we sample distinctive word sequences from your text and search them on the live web as exact phrases. Then we download every candidate page and align every sentence of your text against the page’s real content locally - exact matches by rolling word-window comparison, near matches by token overlap and edit similarity. Search results alone are never trusted as matches; only verified text comparison counts.

When the source is not on the public web at check time: paywalled journals, offline books, private submission archives, pages published minutes ago that search engines have not indexed, or content behind logins. Heavily paraphrased text can also fall below the near-match threshold. No web-based checker escapes these limits; we would rather tell you than let a 0% mislead you.

Correctly quoted passages, common stock phrases, technical boilerplate, legal or religious formulae, and bibliographies all legitimately match their sources. The report separates exact from near matches and shows the source text beside yours so a human can make the call - the score is an instrument reading, not a judgment.

It is the share of your words that sit in sentences we matched to a retrieved source: matched-sentence words divided by total words. The gauge bands are 0-5% "looks original", 6-20% "some matches - review citations", 21%+ "significant matching". The bands are review guidance, not accusation thresholds.

The interface runs in 100+ languages, and the engine itself is multilingual: sentence segmentation handles Latin, CJK and Arabic punctuation, and retrieval uses engines with strong non-English coverage. Match quality is best in languages with a large public web footprint; for very small languages, coverage is honestly thinner.