arXivopen accessscientific preprintsPaul Ginsparge-prints

ArXiv: The Evolution of Open-Access Scientific Preprints

arXiv: The Evolution of Open-Access Scientific Preprints In the rapidly advancing world of scientific research, the speed at which information is shared can determine the pace of global d...

arXiv: The Evolution of Open-Access Scientific Preprints

In the rapidly advancing world of scientific research, the speed at which information is shared can determine the pace of global discovery. At the heart of this information exchange lies arXiv (pronounced "archive," where the 'X' represents the Greek letter chi), an independent, open-access repository of electronic preprints and postprints. These documents, often referred to as e-prints, allow researchers to share their findings with the global community before they undergo the formal, often lengthy, peer-review process.

Since its inception, arXiv has become a cornerstone of modern science, hosting a vast array of papers in fields ranging from mathematics and physics to computer science and economics. By providing a platform for self-archiving, it has played a pivotal role in the movement toward open access in scientific publishing.

A screenshot of the arXiv taken in 1994,[9] using the browser NCSA Mosaic. At the time, HTML forms were a new technology.
A screenshot of the arXiv taken in 1994,[9] using the browser NCSA Mosaic. At the time, HTML forms were a new technology.
: A screenshot of the arXiv taken in 1994,[9] using the browser NCSA Mosaic. At the time, HTML forms were a new technology.

Key Facts

  • Founder: Paul Ginsparg
  • Launched: August 14, 1991
  • Primary Fields: Physics, Mathematics, Computer Science, Astronomy, and more.
  • Current Scale: Approximately 24,000 new articles submitted per month as of late 2024.
  • Status: Independent nonprofit organization (as of July 2026).

The History of a Scientific Revolution

The origins of arXiv can be traced back to the early 1990s. Before a centralized system existed, researchers like Joanne Cohn shared physics preprints via email using the TeX file format—a specialized typesetting system that allowed scientific documents to be easily transmitted and rendered. However, as the volume of research grew, email inboxes quickly reached capacity.

Recognizing the need for central storage, Paul Ginsparg established a central repository mailbox at the Los Alamos National Laboratory (LANL) in August 1991. Over the following years, access methods evolved from FTP and Gopher to the World Wide Web. Originally known as the LANL preprint archive, the service expanded its scope beyond physics to include disciplines like astronomy and quantitative biology. In 2001, to ensure continued growth, the repository moved to Cornell University and adopted the name arxiv.org.

arXiv's yearly submission rate growth over 30 years since its beginning with topics labeled by the standard abbreviations used on arxiv.org[10]
arXiv's yearly submission rate growth over 30 years since its beginning with topics labeled by the standard abbreviations used on arxiv.org[10]
: arXiv's yearly submission rate growth over 30 years since its beginning with topics labeled by the standard abbreviations used on arxiv.org[10]

From Cornell to Independence

For much of its history, arXiv was managed by the Cornell University Library, which took over administrative and financial responsibility in 2011. However, in a significant structural shift, arXiv announced in March 2026 that it would separate from Cornell to become an independent nonprofit organization. This move was designed to diversify and increase the repository's funding streams.

How arXiv Works: Moderation and Identification

While arXiv is not a peer-reviewed journal, it is not unmoderated. To maintain the relevance and quality of its content, the platform utilizes an endorsement system. For certain categories, new authors must be endorsed by established researchers to ensure their work is appropriate for the intended subject area. While this system has faced criticism for potentially restricting inquiry, it serves as a primary filter for content suitability.

Each paper on the platform is assigned a unique identifier based on its submission date or a specific archival format. For example, modern identifiers follow a YYMM.NNNNN pattern (such as 1507.00123). To ensure researchers can track updates, different versions of the same paper are distinguished by a version number at the end of the identifier.

A screenshot of viewing a paper's abstract on arxiv.org in 2021
A screenshot of viewing a paper's abstract on arxiv.org in 2021
: A screenshot of viewing a paper' abstract on arxiv.org in 2021

Addressing Modern Challenges: AI and Quality Control

As scientific publishing faces new hurdles, arXiv has adapted its policies. In November 2025, the platform announced it would no longer accept computer science review articles or position papers that had not been previously vetted by an academic journal or conference. This decision was driven by an increase in AI-generated research, ensuring that the repository remains a trusted source of high-quality scientific discourse.

Summary of arXiv Statistics and Evolution

Key Milestones and Growth of arXiv
Milestone / Metric Date / Value
Launch Date August 14, 1991
500,000 Article Milestone October 3, 2008
1 Million Article Milestone End of 2014
2 Million Article Milestone End of 2021
Monthly Submission Rate ~24,000 articles (as of Nov 2024)

Frequently Asked Questions

Is arXiv a peer-reviewed journal?

No. arXiv is a repository of preprints and postprints. While the content is moderated to ensure it is relevant to the specified disciplines, it does not undergo the formal peer-review process used by traditional academic journals.

What is the difference between a preprint and a postprint?

A preprint is a version of a scientific paper that precedes formal peer review and publication in a journal. A postprint is the version of the paper after it has been peer-reviewed and accepted, but before the publisher's final formatting.

How is arXiv funded?

Historically, arXiv was funded by Cornell University, the Simons Foundation, and annual voluntary contributions from member institutions. Following its transition to an independent nonprofit in 2026, the organization seeks to further diversify its funding sources.

Why do some papers get withdrawn from arXiv?

According to reports from late 2024, approximately 14,000 preprints have been withdrawn. The most common reason for withdrawal is the discovery of "crucial errors" within the research, though some are withdrawn because they have been officially subsumed by other publications.

Can I use arXiv papers for my own research?

Yes, but copyright status varies. Some papers are available under Creative Commons licenses, while most remain the copyright of the author, with arXiv holding a non-exclusive license to distribute them.