<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Detecton on Michael’s Domain</title><link>https://jeltsch.org/en/tags/detecton/</link><description>Recent content in Detecton on Michael’s Domain</description><generator>Hugo</generator><language>en-us</language><copyright>Copyright © 2002 - 2026 Michael Jeltsch.</copyright><lastBuildDate>Fri, 24 Jul 2026 00:18:18 +0300</lastBuildDate><atom:link href="https://jeltsch.org/en/tags/detecton/index.xml" rel="self" type="application/rss+xml"/><item><title>Ouriginal fails more often than not to detect plagiarism</title><link>https://jeltsch.org/en/plagiarism/</link><pubDate>Thu, 23 May 2024 00:00:00 +0000</pubDate><guid>https://jeltsch.org/en/plagiarism/</guid><description>&lt;p&gt;
 &lt;a href="https://ouriginal.com" target="_blank" rel="noopener noreferrer nofollow"&gt;Ouriginal&amp;nbsp;






 
 
 
 &lt;svg class="svg-inline--fa fas fa-up-right-from-square fa-2xs" fill="currentColor" aria-hidden="true" role="img" viewBox="0 0 512 512" overflow="visible"&gt;&lt;use href="#fas-up-right-from-square"&gt;&lt;/use&gt;&lt;/svg&gt;&lt;/a&gt;
, the plagiarism detection service formerly known as Urkund, has been in use at the 
 &lt;a href="https://www.helsinki.fi" target="_blank" rel="noopener noreferrer nofollow"&gt;University of Helsinki&amp;nbsp;






 
 
 
 &lt;svg class="svg-inline--fa fas fa-up-right-from-square fa-2xs" fill="currentColor" aria-hidden="true" role="img" viewBox="0 0 512 512" overflow="visible"&gt;&lt;use href="#fas-up-right-from-square"&gt;&lt;/use&gt;&lt;/svg&gt;&lt;/a&gt;
 for a very long time. Because of Ouriginal&amp;rsquo;s inability to process some of our students&amp;rsquo; Master&amp;rsquo;s thesis, I spent a little bit of time investigating Ouriginal and its problems. Here&amp;rsquo;s what I found out. &lt;strong&gt;From Swedisch to European to US-American&lt;/strong&gt;It is unclear to me which software we are really using since Ouriginal is a merger of the Swedish Urkund and the German PlagScan, and in 2021 Ouriginal was acquired by the American 
 &lt;a href="https://en.wikipedia.org/wiki/Turnitin" target="_blank" rel="noopener noreferrer nofollow"&gt;Turnitin&amp;nbsp;






 
 
 
 &lt;svg class="svg-inline--fa fas fa-up-right-from-square fa-2xs" fill="currentColor" aria-hidden="true" role="img" viewBox="0 0 512 512" overflow="visible"&gt;&lt;use href="#fas-up-right-from-square"&gt;&lt;/use&gt;&lt;/svg&gt;&lt;/a&gt;
. I don&amp;rsquo;t need any of such software, but our university mandates that theses must be scanned by the software. Since I read and evaluate an increasing number of theses every year, I also increasingly use Ouriginal. &lt;strong&gt;Ouriginal supports many document formats, but support is buggy&lt;/strong&gt; Looking at the history of submissions, I realized that until about mid-May 2024, all submitted PDFs had been dutifully analyzed. But last week, something changed since about half of the submitted documents were not processed. The usual tricks (submitting in another format such as .docx, .odt, .ps, .rtf) did work for some, but not all of the theses. One thesis was especially stubborn, and neither our &amp;ldquo;experts&amp;rdquo; nor the Original helpdesk staff had any clue why their software could not process it (which makes me again wonder what software their backend is actually using). Ouriginal is investigating already for a month without getting back to me with answers… Our university receives funding from the Ministry of Education based on graduating students, and we lose thousands of Euros for each student who fails to graduate on time. Fee-liable students also have a strong incentive to graduate on time in order to avoid additional costs. Therefore, every delay in the graduation schedule is unacceptable. &lt;strong&gt;Ouriginal&amp;rsquo;s performance disappoints with a shocking detection rate&lt;/strong&gt; Since there are dozens of ways how to turn a Word document into a PDF, I tried to isolate the problem by submitting test documents to the Ouriginal service. I used manuscripts that we had published over the last few years. I was shocked when I realized that Ouriginal failed to detect plagiarism in more than half of all cases. Paywalls are not to blame since almost all our publications are freely available. Neither the novelty of the publication nor the quality of the journal made a perceivable difference. However, it seems evident that Ouriginal has more problems with non-English documents. It is very difficult to explain why so many freely available scientific texts from reputable journals (
 &lt;a href="https://www.nature.com/srep/" target="_blank" rel="noopener noreferrer nofollow"&gt;Scientific Reports&amp;nbsp;






 
 
 
 &lt;svg class="svg-inline--fa fas fa-up-right-from-square fa-2xs" fill="currentColor" aria-hidden="true" role="img" viewBox="0 0 512 512" overflow="visible"&gt;&lt;use href="#fas-up-right-from-square"&gt;&lt;/use&gt;&lt;/svg&gt;&lt;/a&gt;
, 
 &lt;a href="https://link.springer.com/journal/10456" target="_blank" rel="noopener noreferrer nofollow"&gt;Angiogenesis&amp;nbsp;






 
 
 
 &lt;svg class="svg-inline--fa fas fa-up-right-from-square fa-2xs" fill="currentColor" aria-hidden="true" role="img" viewBox="0 0 512 512" overflow="visible"&gt;&lt;use href="#fas-up-right-from-square"&gt;&lt;/use&gt;&lt;/svg&gt;&lt;/a&gt;
, 
 &lt;a href="https://www.sciencedirect.com/journal/annals-of-anatomy-anatomischer-anzeiger" target="_blank" rel="noopener noreferrer nofollow"&gt;Annals of Anatomy&amp;nbsp;






 
 
 
 &lt;svg class="svg-inline--fa fas fa-up-right-from-square fa-2xs" fill="currentColor" aria-hidden="true" role="img" viewBox="0 0 512 512" overflow="visible"&gt;&lt;use href="#fas-up-right-from-square"&gt;&lt;/use&gt;&lt;/svg&gt;&lt;/a&gt;
) are not indexed by Ouriginal. &lt;strong&gt;Failure to detect plagiarism from Helsinki University&amp;rsquo;s own thesis repository&lt;/strong&gt; Original does not even index our own university&amp;rsquo;s freely available PhD theses (
 &lt;a href="https://ethesis.helsinki.fi" target="_blank" rel="noopener noreferrer nofollow"&gt;https://ethesis.helsinki.fi&amp;nbsp;






 
 
 
 &lt;svg class="svg-inline--fa fas fa-up-right-from-square fa-2xs" fill="currentColor" aria-hidden="true" role="img" viewBox="0 0 512 512" overflow="visible"&gt;&lt;use href="#fas-up-right-from-square"&gt;&lt;/use&gt;&lt;/svg&gt;&lt;/a&gt;
 ). None of the PhD thesis written in my own lab was recognized. I also checked 
 &lt;a href="https://jeltsch.org/en/phd_thesis/"&gt;my own PhD thesis&lt;/a&gt;
, and it was the only one that was recognized (Ouriginal found it on 
 &lt;a href="https://docslib.org/doc/7864611/vegfr-3-ligands-and-lymphangiogenesis" target="_blank" rel="noopener noreferrer nofollow"&gt;docslib.org&amp;nbsp;






 
 
 
 &lt;svg class="svg-inline--fa fas fa-up-right-from-square fa-2xs" fill="currentColor" aria-hidden="true" role="img" viewBox="0 0 512 512" overflow="visible"&gt;&lt;use href="#fas-up-right-from-square"&gt;&lt;/use&gt;&lt;/svg&gt;&lt;/a&gt;
, not our own university&amp;rsquo;s official thesis repository). Does it make any sense to pay for such an imperfect service? I don&amp;rsquo;t think so. The omission of freely available texts from important journals seems inexcusable since a simple Google search does a better job. I did Google searches for sentences from the unrecognized texts, and in most cases, Google identified the original source without problems. A Python script to split texts into sentences for individual Google searches and to aggregate the results can be written by any programmer in an afternoon. Why would you maintain your own database if Google does a better job than you can? &lt;em&gt;&lt;em&gt;Plagiarism to the following published manuscripts/theses/proceedings&lt;/em&gt; was NOT detected:&lt;/em&gt;*&lt;/p&gt;</description></item></channel></rss>