표절 통계
데이터 출처

수치의 근거가 되는 자료입니다.

모든 출처는 무료이며 오픈 라이선스로 제공됩니다. 점 표시는 수치 산출 방식(원시 기록 기반 집계 또는 선별된 자료 종합)을 나타냅니다. 전체 산출 과정은방법론 페이지.

Tier 1 — high-authority datasets (the rate data)

Retraction events + categorized reasons (the plagiarism numerator).

Free, Crossref-hosted

Monthly snapshots

Publications, ROR-resolved affiliations, disciplines (denominators + the DOI join).

Free, CC0

Real-time

DOI metadata linking retractions to works.

Free API

Continuous

Canonical institution identifiers (the disambiguation backbone).

Free, CC0

Tier 2 — curated sources (the case data + signal)
Weekly

Notable plagiarism cases (Category:Plagiarism controversies, scientific misconduct).

Free, CC-BY-SA

Daily

Global news-volume index + recent coverage (current-events signal).

Free, DOC 2.0 API

법원 기록(CourtListener)과 기업 공시(SEC EDGAR)는 평가하였으나 제외되었습니다. 전체 텍스트 일치 결과가 신뢰성 있게 귀속되기에는 잡음이 많기 때문입니다. EU 판례와 국가별 연구부정 보고서(영국 QAA, 독일 DFG)는 수작업 출처로 향후 추가 예정입니다.
데이터 출처 및 신뢰도 — 표절 통계