Digital resources in the Social Sciences and Humanities OpenEdition Our platforms OpenEdition Books OpenEdition Journals Hypotheses Calenda Libraries OpenEdition Freemium Follow us

The path toward ethical science is paved with diamonds

Over the last two decades awareness of a Reproducibility Crisis penetrated all scientific disciplines. Studies recently showed that 90% in STEM and 95% in psychology were aware of this Crisis1,2. At its core this is a crisis of trust. The findings scientists publish and tout as facts, are often not replicable by others. Moreover, a great many are not even computationally reproducible using the original data3,4.  The Open Science Movement developed partly in response to this Crisis5,6. Especially in the last two decades, researchers organized grassroots movements to make science more transparent, reliable and ethical.

One of the many tenets of the Open Science (OS) Movement is open access. Published research results must be accessible to everyone. This is not a new idea. In the 1940s Robert K. Merton developed normative goals necessary to ensure integrity in science7. One was that all scientific findings should be public property. He argued that effective and efficient progress of science depends on open public access. In data science this Movement led to the concept of Open Science by Design – a strategy where scientists plan in advance how they will make all metadata, data and algorithms available and easily re-usable by anyone8.

The OS Movement has been successful. Open access journals and articles increased rapidly. They outpaced the global increase in publications in other formats in the last two decades9. One problem is that these open access articles are overwhelmingly funded by the authors of the studies via Author Processing Charges (APCs). It costs over 10 thousand US$ to publish a study open access in the journal Nature, and most other journals charge at least two thousand. Except for elite institutes and projects with generous third-party fundings, these fees are not usually covered or coverable by universities. A naïve outsider might ask why scientists would pay astronomical fees to make their own hard work publicly available, when they can simply share it online in a free repository like the Open Science Framework, Github or any number of preprint servers based on the ArXiv model?

The answer is competition. The strongest norm governing the practice of science today is publish-or-perish. In every field and science at large, scientists are judged based on their publication record. Ask any academic what the top journals in their area are, and you will get relatively consistent answers by discipline, sub-field and science in general10. Look at the faculty of any ‘top’ university, and the faculty in any discipline will have one or more publications in these ‘top’ journals. Without publications in these journals, a scientist has no chance at a career in science. This is a collective cultural fact. One deeply institutionalized in the organizations, rules, norms, expectations and behaviors of scientists11,12.  

Thus, the founding of new open access journals with lower APCs has little impact on the scientific enterprise because they cannot compete with institutionalized legacy statuses of existing journals. The greatest success story is the non-profit publisher PLOS, which rose swiftly in the rankings, but could not crack into the very top tier. Although far cheaper than Nature, it is still not ‘cheap’, with APCs in their family of journals ranging from 2.5 to 3.2 thousand $US[1].

The most successful open access models are within existing paywalled journals where authors have the option to publish “gold” open access, rather than publish for free. The payment of somewhere between two and 10 thousand US$, gets authors the right to have their single article published open access inside of a closed access journal. This means that despite a massive shift toward open access, the fundamental structures of scientific publishing have not changed.

Ethical Implications

Competition in science supports motivation and innovation, but the publish-or-perish norm is a toxic externality. Scientists willingly prioritize subjective journal ranking, impact factor and increasing their citation counts to get ahead within the competitive scientific enterprise. They do this despite widespread skepticism and evidence that rankings and citations do not correlate strongly with the quality and reliability of published studies14–16. They do this because they must, or at least perceive that they must, in order to follow their scientific career aspirations. The pressure to publish thus motivates rent-seeking behaviors designed to increase publication chances, i.e., career chances. Conducting higher quality science is one of these behaviors. But there are many methods to increase publication chances that are science orthogonal, what are known today as questionable research practices (QRPs)17,18.

Some QRPs are unconscious and learned from supervisors in the process of converting research into a publishable paper. For example, researchers routinely run many statistical models but report only those that show the strongest support of their claims. Researchers who believe that their claim is true in the first place will gravitate toward models that support it, convincing themselves intrinsically that these models are the best tests of their claim.

Decisions based on confirming intrinsic beliefs or window dressing for peer reviewers reduce the replicability of science. They narrow down the multiverse of potential findings into a highly selected set of results. This selectivity is independent of the process that generated the data in the first place, in other words, it is science orthogonal. A classic example is a study published suggesting that hurricanes with feminine names cause more damage than those with masculine names. It turns out that using the available data, the original researchers selected a model that produced regression coefficients that were extremely far away from the central tendency among all other plausible models’ regression coefficients19. We do not know whether this was a conscious decision but their reporting hides the truth and simultaneously increases publication chances.

There are of course conscious and highly unethical behaviors leading to an entirely false representation of reality. Science is filled with scandals of hacking and data-faking. Rent-seeking alone can explain unethical learned unconscious and conscious behaviors of scientists. They seek status and money and job security, or in some cases seek results that support a particular worldview or policy outcome20,21. But rent-seeking is unambiguously the main reason22,23.

If we as a scientific community and science-interested public, want to eliminate the perverse incentive structures that bound and inform rent-seeking behaviors of scientists we need radical change. Status should not be assigned based on journal metrics that often have little to do with the quality of the research being conducted or published and more to do with legacy and embedded norms. The most radical proposal is to eliminate journals altogether. But this is an extremely unlikely outcome no matter how powerful the OS Movement becomes.

Journals became standard in science hundreds of years ago as a means for communicating scientific discoveries across time and space24. They enabled scientists to acquire knowledge without travelling to faraway universities. Publishing was not cheap, and with the dawn of digital media, it became even more expensive as publishers raced to provide their journal both in-print and online. The costs associated with publishing gave publishing firms a great deal of power over time.

The result is that companies like Springer Nature and Elsevier have enough power to dictate to scientists how they perform their research and communicate their results25. They shape academic careers, institutional priorities, governments’ science policies, and perceptions of journals and the publishing enterprise among scientists26. Their primary legal interest is their shareholders. Profit is their priority, and only second is to provide a service to science as their product. A simple mathematical proof confirms that profit is their priority: If they cannot make a profit they will no longer provide the scientific services but if they can make a product without scientific services they have no reason to stop.

Although big publishing firms have shown many draconian practices in their ‘service’ to science27–30, scientists still need a means to communicate their results with each other and the public. This can be done without for-profit publishing but probably not without a journal publication format, or something very similar. For example, we now have a plethora of ‘green’ open access preprint servers where scholars can deposit working papers, or prior versions of their published articles. These are a viable means to communicate science without the perverse incentives generated by big publishing or institutionalized journal rankings. This all still involves the journal article as the standard unit of science production.

Diamond Open Access

If we are ‘stuck’ with journals in science as our primary communication medium, then they should be as free from perversely incentivized bias as much as possible. The first step is thus to remove for-profit publishing from the equation. It is fine to use the services of for-profit publishers. They have shown the capacity to provide print on demand, marketing and scientific journalism. But when they control and direct the scientific enterprise when have an ethical conflict of interest – namely profit versus robust science.

The costs of publishing a journal are the lowest in history. There are publication kits that help associations, institutions and stand-alone journals to take publication into their own hands31. With minimal costs and self-governance, journals have no need to push institutions to purchase journal subscriptions and no need to charge authors astronomical publication fees. Thus, they themselves are not perversely incentivized to perpetually increase their status to make their product profitable.

The optimal existing solution is diamond open access, whereby a journal charges no APCs and is freely readable and downloadable online32. It is the most ethical and equitable by design, and is perceived as the ideal model by most academics33. The OS Movement is overwhelmingly in favor of diamond open access34, as are governance bodies – at least those free from the influence of big publishing like UNESCO35,36. Despite 13 thousand journals indexed in the Directory of Open Access Journals (DOAJ) with no fees as of February 16th, 2026, these journals are not on the radar of most indexing services, and do not belong to the mainstream of journals published by major scientific societies37. This means that successful diamond open access journals are very rare.

Because of such a saturated scientific ‘market’ for publication outlets, starting new diamond open access journals has had little impact on producing a more ethical science. They do not gain reputation. The reason PLOS was so successful, despite failing to break into the highest echelon of science, was because it has a huge cash flow and can use it for branding and promotion. From an ethical science perspective it is valuable, like gold, but it is not as valuable as diamond.

Flipping or Starting Over?

To have the highest ranking and most well-known journals diamond open access, societies need to cancel their contracts with for-profit publishers38. One problem with this is that many academic societies are themselves run like for-profit businesses. They seek to generate as much revenue as possible, and diamond open access would threaten this model. They have embedded relationships with publishers and agree to renew contracts together, mostly independent of the scientists they serve. I witnessed this first hand with the American Sociological Association and Sage39,40. Contracting a for-profit publisher provides a non-profit academic organization a scapegoat for amassing capital via subscription fees.

There are incredible exceptions; however, and they offer model success stories. Computational Linguistics is one of the earliest journals considered to be among the top in its field, to flip to diamond open access. The first step was a move by the Association for Computational Linguistics in 2002 to create the CL Anthology which made all articles published by association journals open access online after an embargo period. Then in 2009 the journal Computational Linguistics flipped to diamond open access. The journal Demography of the Population Association of America is another example.

The embeddedness of big publishing in science, and the capital that associations can raise through their journals when run by big publishers, are major barriers to flipping. Another major barrier is that some big publishers coerce scientific societies into signing away the rights to the titles of their journals in their publishing contracts. This tactic is most intensively deployed by Elsevier. They own the rights to the titles of nearly all the journals they publish. Therefore, when these societies want to change publishers, they cannot. They are trapped. They would have to start a new journal with a different title – what happened for example with the Journal of Infometrics41. There is otherwise no way around this problem because of copyright law.

Therefore, collective efforts to build the popularity of new diamond open access journals would greatly increase the movement toward a more ethical and effective scientific enterprise. If these are journals from societies, which are essentially the same journal but with a new (not copyrighted by Elsevier) name, it requires authors to support this journal and immediately abandon the other. Supporting new diamond open access journals, whether completely new or newly named, requires established scholars to put their status-seeking egos aside in the name of scientific progress and ethics. It is precisely those who have built major scientific reputations who need to engage in this change, because they can give them most clout to new journals by touting them and publishing in them.

Scholars and societies alone cannot carry the burden. Hiring committees need to reward these behaviors. Rather than seeing a publication in a new ‘unranked’ or ‘low ranked’ diamond open access journal as a sign that the article is ‘not high quality enough for top journals’, committees should judge publications only on their content. Then, if two publications are seen as equally scientifically rigorous and high quality, the one in a diamond open access journal should get a greater weight. Scientist who consistently publish high quality research in diamond open access journals should be favorites of hiring committees, all else equal.

A hiring committee would be unlikely to reward an applicant who shows sociopathic behaviors. Following this logic, they should not reward scientists who show behaviors that go against ethical science by practicing closed and profit-incentivized science. I need to be very clear here that I personally am still publishing regularly in journals published by for-profit publishers. I am not that famous, and I do not have tenure. This is a perfect example of why we need sweeping changes at the institutional level, and from the top of the scientific hierarchy.

With efforts to flip both  existing journals and efforts to reward new, ethical journals, we give science its greatest future chances for improvement and sustainability. There are many efforts underway and these should serve as guides, for example the Diamond Open Access Fund from the Dutch Research Council and MIT’s shift+OPEN initiative.

Diamond AI

Diamond open access has a second, equally important role for the future of science. It is necessary to inform Generative Artificial Intelligence (Gen AI). The LLMs that power popular Gen AI are trained heavily on corpora assembled from what is available via the Internet. Paywalled literature is missing from training corpora, not because it is unimportant, but because it is not accessible or legally usable42. Gen AI outputs are therefore based on a highly restricted sample of all scientific knowledge. Open access for all of science would ensure that anyone using Gen AI, would get the best possible information based on all that we know as humans. As essentially everyone is using Gen AI today43, this would mean that the public would be optimally informed.

It is not necessary for Gen AI development that all articles are diamond open access, they just need to be somewhere, e.g., green or gold open access. Yet, without diamond open access the entire knowledge enterprise will continue to favor the work of those with greater resources44. Resources are of course necessary to produce higher quality science because of research costs, but when it comes to publishing and dissemination, a resource advantage reproduces the already existing Global North-South disadvantages in science. This limits human capacity to tap resources in lower income societies. These are societies filled with potential contributions to science that could benefit both 1) their own societies’ development – because it enables them to gain status, resources and build stronger, more attractive and sustainable scientific institutions, and 2) all of scientific knowledge because there are brilliant minds waiting to be tapped that might otherwise give up on science because it is for them no sustainable in their region.

If the knowledge in Gen AI remains Global North and WEIRD biased (Western, educated, industrialized, rich and democratic), the result is that these countries remain culturally and economically advantaged beyond that which exists presently. This means that without diamond open access norms, we are willingly allowing Gen AI to increase cultural hegemony and economic domination of a minority. I am not directly arguing whether this is good or bad. I am in the Global North and profit from a stronger Global North science advantage. My argument is that science should be neutral, favoring the most optimal and reliable knowledge, rather than legacies or strategies that game the scientific system.

References

1.          Baker, M. 1,500 scientists lift the lid on reproducibility. Nature 533, 452–454 (2016).

2.          Metskas, A. How Much Do Academic Psychologists Trust Academic Psychology, and Is There Still a Replication Crisis? Transparent Replications https://replications.clearerthinking.org/how-much-do-academic-psychologists-trust-academic-psychology-and-is-there-still-a-replication-crisis/#survey-demographics (2025).

3.          Open Science Collaboration. Estimating the reproducibility of psychological science. Science 349, (2015).

4.          Breznau, N. et al. The reliability of replications: a study in computational reproductions. Royal Society Open Science 12, 241038 (2025).

5.          Engzell, P. & Rohrer, J. M. Improving Social Science: Lessons from the Open Science Movement. PS: Political Science & Politics 1–4 (2021) doi:10.1017/S1049096520000967.

6.          Breznau, N. Legacy of Jon Tennant, “Open science is just good science”. Crowdid https://crowdid.hypotheses.org/548 (2022) doi:10.58079/ne8h.

7.          Merton, R. K. The Sociology of Science: Theoretical and Empirical Investigations. (University of Chicago press, 1973).

8.          Wittenburg, P. Open Science and Data Science. Data Intelligence 3, 95–105 (2021).

9.          NCSES. Publication Output by Region, Country, or Economy and by Scientific Field. https://ncses.nsf.gov/pubs/nsb202333/publication-output-by-region-country-or-economy-and-by-scientific-field#utm_source=chatgpt.com (2023).

10.        Serenko, A. & Bontis, N. A critical evaluation of expert survey‐based journal rankings: The role of personal research interests. Asso for Info Science & Tech 69, 749–752 (2018).

11.        Mancoridis, M., Sumers, T. & Griffiths, T. Publish or Perish: Simulating the Impact of Publication Policies on Science. Proceedings of the Annual Meeting of the Cognitive Science Society 46, (2024).

12.        Breznau, N. Questionable research practices from the practitioners’ perspectives. Crowdid https://crowdid.hypotheses.org/1666 (2025) doi:10.58079/14f3u.

13.        Lawrence, S. Free online availability substantially increases a paper’s impact. Nature 411, 521–521 (2001).

14.        Fleck, C. The Impact Factor Fetishism. European Journal of Sociology / Archives Européennes de Sociologie 54, 327–356 (2013).

15.        Rushforth, A. & De Rijcke, S. Practicing responsible research assessment: Qualitative study of faculty hiring, promotion, and tenure assessments in the United States. Res Eval 33, (2024).

16.        Dougherty, M. R. & Horne, Z. Citation counts and journal impact factors do not capture some indicators of research quality in the behavioural and brain sciences. R Soc Open Sci. 9, 220334 (2022).

17.        Gopalakrishna, G. et al. Prevalence of questionable research practices, research misconduct and their potential explanatory factors: A survey among academic researchers in The Netherlands. PLOS ONE 17, e0263023 (2022).

18.        John, L. K., Loewenstein, G. & Prelec, D. Measuring the Prevalence of Questionable Research Practices With Incentives for Truth Telling. Psychol Sci 23, 524–532 (2012).

19.        Muñoz, J. & Young, C. We Ran 9 Billion Regressions: Eliminating False Positives through Computational Model Robustness. Sociological Methodology 48, 1–33 (2018).

20.        Borjas, G. J. & Breznau, N. Ideological bias in the production of research findings. Science Advances 12, eadz7173 (2026).

21.        Rainero, V., Stolz, J. & Luijkx, R. The Faith Factor. How Scholars’ Religiosity Biases Research Findings on Secularization. Sociological Science 13, 154–177 (2026).

22.        Aronson, J. K. When I use a word . . . “Publish or perish”: adverse effects. https://doi.org/10.1136/bmj.r1577 (2025) doi:10.1136/bmj.r1577.

23.        Paruzel-Czachura, M., Baran, L. & Spendel, Z. Publish or be ethical? Publishing pressure and scientific misconduct in research. Research Ethics 17, 375–397 (2021).

24.        Carey, J. Scientific Communication Before and After Networked Science. Information & Culture 48, 344–367 (2013).

25.        Larivière, V., Haustein, S. & Mongeon, P. The Oligopoly of Academic Publishers in the Digital Era. PLOS ONE 10, e0127502 (2015).

26.        Rossello, G. & Martinelli, A. The effect of lobbies’ narratives on academics’ perceptions of scientific publishing: A survey experiment. Information Economics and Policy 71, 101148 (2025).

27.        Butler, L.-A., Matthias, L., Simard, M.-A., Mongeon, P. & Haustein, S. The oligopoly’s shift to open access: How the big five academic publishers profit from article processing charges. Quantitative Science Studies 4, 778–799 (2023).

28.        Else, H. Dutch publishing giant cuts off researchers in Germany and Sweden. Nature 559, 454–455 (2018).

29.        Lancet, T. & Board, T. L. I. A. Reed Elsevier and the arms trade. The Lancet 366, 868 (2005).

30.        Jureidini, J. & Clothier, R. Elsevier should divest itself of either its medical publishing or pharmaceutical services division. The Lancet 374, 375 (2009).

31.        Scholastica. New fully-OA publishing toolkit and stakeholder reflections on 20 years of the BOAI. https://blog.scholasticahq.com/post/oa-publishing-toolkit-and-stakeholder-reflections-BOAI20/ (2022).

32.        Normand, S. Is Diamond Open Access the Future of Open Access? The iJournal: Student Journal of the Faculty of Information 3, (2018).

33.        Kumari, M. & A, S. Perceptions of open access publishing: A comparative study of gold and diamond models among global researchers. Alexandria 35, 55–73 (2025).

34.        Plan S. Working collectively towards an equitable, community-driven and academic-led scholarly publishing model. Action Plan for Diamond Open Access https://www.coalition-s.org/action-plan-for-diamond-open-access/ (2022).

35.        UNESCO. Diamond Open Access. Public Service Press Release vol. Online Report (2026).

36.        UNESCO. UNESCO Recommendation on Open Science. https://unesdoc.unesco.org/ark:/48223/pf0000379949.locale=en (2021).

37.        Simard, M.-A., Basson, I., Hare, M., Larivière, V. & Mongeon, P. The Value of a Diamond: Understanding Global Coverage of Diamond Open Access Journals in Web of Science, Scopus, and OpenAlex to Support an Open Future. Proceedings of the Annual Conference of CAIS / Actes du congrès annuel de l’ACSI https://doi.org/10.29173/cais1845 (2023) doi:10.29173/cais1845.

38.        Trueblood, J. S. et al. The misalignment of incentives in academic publishing and implications for journal reform. Proc. Natl. Acad. Sci. U.S.A. 122, e2401231121 (2025).

39.        Why I’m leaving the American Sociological Association. Family Inequality https://familyinequality.wordpress.com/2021/11/06/why-im-leaving-the-american-sociological-association/ (2021).

40.        Philip Cohen’s ASA Publications Committee platform. Family Inequality https://familyinequality.wordpress.com/2018/01/14/philip-cohens-asa-publications-committee-platform/ (2018).

41.        Singh Chawla, D. Open-access row prompts editorial board of Elsevier journal to resign. Nature https://doi.org/10.1038/d41586-019-00135-8 (2019) doi:10.1038/d41586-019-00135-8.

42.        Szkalej, K. The Paradox of Lawful Text and Data Mining? Some Experiences from the Research Sector and Where We (Should) Go from Here. GRUR Int 74, 307–319 (2025).

43.        Breznau, N. & Nguyen, H. H. V. An Introduction to Generative Artificial Intelligence for Academics. F1000 Research Preprint Status, (2025).

44.        Kwon, D. Open-access publishing fees deter researchers in the global south. Nature https://doi.org/10.1038/d41586-022-00342-w (2022) doi:10.1038/d41586-022-00342-w.

Ethical Statement

I have no conflict of interest to report. I used Google search which includes Gemini by default, ChatGPT, NotebookLM and Nano Banana to support my literature review, check arguments, suggest words or phrases and to extract facts. No writing was copied from Gen AI, it is all my own. Any mistakes or opinions are also my own.


[1] https://plos.org/fees/

To respectfully and ethically decline to support Elsevier

Just over 5 years ago, I decided I would no longer peer review for or publish with Elsevier. In addition, I would try not to support any of their products. This was a choice enabling me to be more consistent with my own ethical values and support of open science.

This is was a difficult decision to make, because for-profit publishing companies have all engaged in questionable behaviors in pursuing economic gains. Also, it means saying ‘no’ to colleagues who are editors and could really use my expertise for a peer review or book chapter contribution. Therefore, I have created a brief letter template to decline such offers. It provides some informative links, encourages the recipient to pursue their own opinion and avoids overreaction via defamation or fear-mongering. Elsevier is known to attack with lawyers, so it is also important that I protect myself and my associates when declining such offers.

Dear [person],

I apologize but I cannot in good faith [write a peer review / book chapter] for an Elsevier publication.

I realize that publishers are ‘for profit’ enterprises, and balance the needs of science with their own pursuit of economic success; however, in my experience Elsevier has consistently proven to be an exceptionally unethical company with practices antithetical to the pursuit of knowledge. Some of the more awful practices have been to sponsor arms fairs[1], advocate intensively against open access[2], create seemingly scientific journals and then sell space in those journals for companies to print bogus articles that support their products[3], purchase and copyright or paywall scientific tools[4], charge astronomical fees for journal subscriptions[5], and over-harvest my personal information[6]. Therefore, if I can help it, I do not peer review or publish with Elsevier anymore.

Please do not take this personally. This letter expresses my own personal opinion based on my own experiences. My opinion and resulting preferences do not reflect that of my employer or any colleagues with whom I have collaborated. I encourage you and each person to investigate for profit publishers on their own and make their own informed decisions.

Sicerely,

[Author]

[1] https://www.thelancet.com/journals/lancet/article/PIIS0140-6736(07)60488-7/fulltext

[2] https://en.wikipedia.org/wiki/Elsevier#Criticism_of_academic_practices, https://www.theguardian.com/science/political-science/2018/jun/29/elsevier-are-corrupting-open-science-in-europe, https://www.lub.lu.se/en/find/lubsearch/epublications/new-agreement-elsevier/sweden-stands-open-access-cancels-agreement-elsevier (attacks on ResearchGate and Sci-Hub) https://www.science.org/content/article/publishers-take-researchgate-court-alleging-massive-copyright-infringement, https://www.nature.com/articles/nature.2017.22196

[3] https://www.the-scientist.com/elsevier-published-6-fake-journals-44160, https://www.thelancet.com/journals/lancet/article/PIIS0140-6736(09)61404-5/fulltext

[4] Social Science Research Network – purchased and turned into a place for paywalls; Mendeley – purchased and converted into a product that is not (easily) interoperable with open source products like Zotero; tries to patent a peer review tool to a payment step into the scientific process https://www.eff.org/deeplinks/2016/08/stupid-patent-month-elsevier-patents-online-peer-review.

[5] https://www.science.org/content/blog-post/california-tells-elsevier-take-hike, https://www.thenation.com/article/society/neuroimage-elsevier-editorial-board-journal-profit

[6] https://eiko-fried.com/welcome-to-hotel-elsevier-you-can-check-out-any-time-you-like-not/

A fake journal, algorithmic plagiarism and tricking Google Scholar

A case of a fake journal that passed Google Scholars’ bots and algorithm checking to become indexed. The journal contained articles that were completely plagiarized but assigned new titles, abstracts and authors.

See updates at end of blog for table of plagiarized papers, and a small win for open science after the domain host revoked the content of the plagiarized paper’s URLs.

A colleague of mine was searching Google Scholar using certain words and the number two hit appeared with the following title and direct download link.

The PDF looks like a journal article. The journal has an ISBN instead of DOIs, but this is not unheard of. It even has an impact factor listed: 7.11, it must be good.

It turns out that this seemingly interesting paper with an alleged author who is an expert in this area, is plagiarized.

How do I know it is plagiarized? Just ask the real authors of the paper which appeared with a different title and abstract in Social Science Research one year earlier. Oh and with an identical word-for-word text, with some mistakes relating to equations – possibly a failed scraping algorithm.

The journal title is Journal of Current Research and Review, but it is extremely difficult to find through an internet search. This requires endless scrolling (as of 10.08.2023), or removing the ending on the download URL to arrive at the index page.

Scrolling down on the home page reveals the “Latest Articles”. Here I find the same title that appeared in Google Scholar. However there is no download link, only the title and abstract. In other words, the URL that Google Scholar found to download the paper is not accessible through any link on the journal’s website.

How is it possible that Google Scholar would pick up and index such garbage?

I asked ChatGPT to tell me how Google Scholar works, and the results suggest that someone or something had advanced knowledge and programming skills that would provide the right meta-data and materials on the journal webpage to pass the automated Google Scholar bots which are constantly scraping the web for academic papers and publications.

From ChatGPT prompted with “How does a journal get indexed in Google Scholar?” (09.08.2023). I received the following answer. Note that I removed the descriptions of each point, as they are mostly self-evident.

Google Scholar indexes scholarly literature, including journal articles, conference papers, theses, and other academic resources. In order for a journal to get indexed in Google Scholar, there are several steps and criteria to consider:

  1. Quality Content
  2. Digital Format
  3. Open Access and Accessibility
  4. Consistent Publishing Schedule
  5. Meta Tags and Structured Data
  6. Proper Article Metadata
  7. Citations and References
  8. Author Profiles

It’s important to note that Google Scholar’s indexing process is automated, and there is no formal application process for journals to get indexed. Google Scholar’s algorithms discover and index content based on various factors. However, journals can follow best practices to increase their chances of being indexed and improving their visibility within Google Scholar’s search results.

The journal website and meta-data had to pass Google Scholar’s algorithm. This is no lucky feat. Although Google does not reveal any usage of AI to evaluate journals for indexing, it clearly has an advanced system of scraping, filtering and crawling. Who or whatever designed this journal, or really this website that looks like a journal, understood how to pass the test. For example, the “Latest Articles” appear to be from “Volume 14” suggesting to potential crawling bots that the journal may have been published for 14 years now.

This is of course false. There are no other articles readily accessible from the journal’s website other than the two that appear. The other article is also plagiarized, in case you were wondering. But, a closer look at the URL reveals that the article I am discussing has the number “800” at the end of its URL.

https://zapjournals.com/Journals/index.php/jcrr/article/view/800

Changing this number yields other papers. At least as far back as number 795; prior to that yields 404 errors. Moreover, I cannot find the papers from 795-798, only their titles and abstracts.

The journal website’s main page yields something even stranger. It looks like what a computer science student might create as an example page to try and sell as a template or to demonstrate the services on offer for a website construction gig. The content has absolutely nothing to do with journals, academia or publishing.

It even comes with fake testimonials. I wonder whose pictures were stolen for this…

The real question for me is, ‘what to do about this?’. Clearly this is fraud and constitutes both ethical and legal infringements on science. Looking up the host of the domain of the journal using ICANN reveals that it is provided by a Lithuanian company Hostinger.

Looking at this company’s website suggests they are legitimate. The provide contact information to report abuse. So that is what I did by sending them the info in this blog post.

Hostinger replied within two days and told me they take abuse of their services very seriously and asked me for more evidence so they could pursue the case. I gave them the two links for the plagiarized papers and their original versions published one year earlier in Social Science Research.

https://zapjournals.com/Journals/index.php/jcrr/article/download/799/1237

https://zapjournals.com/Journals/index.php/jcrr/article/download/800/1238

Word for word plagiarized from:

https://www.sciencedirect.com/science/article/pii/S0049089X22000321

https://www.sciencedirect.com/science/article/pii/S0049089X22000333

For all the negative publicity surrounding Elsevier, they are still a player in the academic world. In 2020 they owned roughly 16% of the academic publishing market. If there is anyone who would want to prevent abuse of their services, including plagiarism of their work, it is Elsevier. They have a trove of lawyers fighting and winning battles to protect their content across the world. As the two plagiarized papers that I can download are from the Elsevier journal Social Science Research, it made sense to contact them as well. Elsevier’s due process suggests contacting the journal editor first. Thus, I have sent this information to the lead editor.

The journal also lists an ISBN number. But this number is a fake, and returns an invalid search with the ISBN lookup tool.

Google certainly would not want its products indexing fake journals with plagiarized papers, so I took the liberty of contacting them as well.

What is striking about the journal is that they have a long editorial board list. Internet searching reveals that these are real scholars. I also contacted them. I will continue to report on this case as it unfolds.

The question for me is: ‘what motivated someone to create this site?’. There is clearly no profit associated with it. There is also no status gain, because the plagiarized papers have authors assigned to them who are not the original authors and clearly not players supporting the fake journal. They are highly established scholars in their respective sub-fields. These fakely assigned authors are perfect examples of what an AI might choose to assign to a certain topic. I tested this by asking Chat GPT if Seamus McGuinness could have written the abstract. The response points at AI as a source for potentially re-writing abstracts of the plagiarized papers and for finding suitable authors to assign to them. ChatGPT said:

Yes, Seamus McGuinness could be a potential author who might have written this abstract. Seamus McGuinness is known for his research on labor market issues, including education-job mismatches, gender disparities, and remote work. His expertise aligns with the themes discussed in the abstract, making him a plausible candidate as one of the authors who could have written it.

It remains a mystery for now what is up with this website and the fake journal. Was it a computer science project that accidentally got picked up by Scholar, one that was never intended for public consumption? Was it an attempt to create a journal, but try and hide the real content of the journal until it could pick up real submissions? Was it entirely AI generated, to showcase the power of an AI?

[Update 14/08/2023]

Thanks to Random Cat on Twitter, I learned that there are more papers than just two. Using Google Scholar they searched by journal.

This allowed me to compile a table of plagiarized papers.

Table of Plagiarized Papers in JCRR, found via Google Scholar

Original TitleOriginal Author(s)Original JournalJCRR TitleJCRR AuthorsLink to Original ArticleLink to Plagiarized Article
The motherhood wage gap and trade-offs between family and work: A test of compensating wage differentialsNick Wuestenenk & Katia BegallSocial Science ResearchCOMPENSATING WAGE DIFFERENTIALS AND THE MOTHERHOOD WAGE GAP: A COMPARATIVE ANALYSISHK Kleven & CL Landaislinklink
Conflicting signals: Exploring the socioeconomic implications of gender discordant namesAndrew Francis-Tan & Aliya SapersteinSocial Science ResearchBREAKING BOUNDARIES: EXAMINING THE INTERSECTION OF GENDER DISCORDANT NAMES AND SOCIOECONOMIC ATTAINMENTAL Roberts & M Rosariolinklink
Gender overeducation gap in the digital age: Can spatial flexibility through working from home close the gap?Ana Santiago-Vela & Alexandra MergenerSocial Science ResearchBRIDGING THE GENDER OVEREDUCATION GAP: EXPLORING THE ROLE OF WORKING FROM HOME IN THE DIGITAL ERASMG McGuinnesslinklink
How the Great Recession changed class inequality: Evidence from 23 European countriesJad MoawadSocial Science ResearchCLASS INEQUALITIES IN THE WAKE OF THE GREAT RECESSION: A STUDY OF 23 EUROPEAN COUNTRIESPD Allisonlinklink
Higher education and high-wage gender inequalityNatasha Quadlin, Tom VanHeuvelen & Caitlin E. AhearnSocial Science ResearchASSESSING THE CONTRIBUTION OF EDUCATION TO GENDER WAGE DISPARITIES IN HIGH-EARNING PROFESSIONSJ Jacobslinklink
TRANSFORMING RESEARCH: EXPLORING THE INTERPLAY OF DATA MINING, MACHINE LEARNING, AND KNOWLEDGE DISCOVERRobert M. Bond & Christopher J. FarissSocial Science ResearchsKnowledge Discovery: Methods from data mining and machine learning [b]Xiaoling Shu & Yiwan Yelinklink
SOCIAL MEDIA ADDICTION: PROPOSED INDICATORS AND STAGESZakaria I. Saleh & Omar Zakaria SalehInternational Journal in Commerce, IT and Social SciencesUNPACKING THE CYCLE OF SOCIAL MEDIA ADDICTION: UNDERSTANDING SYMPTOMS, PROGRESSION, AND RECOVERY [a]OZ Salehlinklink

[a] Published twice in JCRR with different authors

[b] Not indexed in Google Scholar

All are from Social Science Research except one. There are also other papers where I cannot find an original. These papers might actually be original papers, or stolen working papers that are not easy to find online.

Interestingly, some papers that appear that they may not have been plagiarized have a different display in PDF form. This includes contact details for the journal. I thus emailed them as well to inform them that they are committing ethical and legal fraud.

[Update 23.08.2023]

Hostinger investigated the reported problem and determined that the user of their domain had violated the terms of ethical/legal usage and removed the plagiarized papers. However, the journal website still appears. In other words, a journal that obviously intentionally plagiarized several articles, possibly to boost its reputation and encourage others to submit and thus pay the 45 dollar fee, is still out there lurking. Moreover, this journal is part of a larger company called Zenodo Publishing. On their main page the list dozens of journals. If I had to guess I would assume these journals are also predatory, and may contain plagiarized content. Only investigating this will prove if this is true.

More to come.