Scholarly communication is the field of study and practice concerned with how research and other academic knowledge is created, evaluated, distributed, preserved, and used. It encompasses the formal and informal channels through which scholars share their work with each other and with broader publics, as well as the systems, policies, and economics that shape those channels. As a subfield of information science, it sits at the intersection of library science, publishing studies, the sociology of science, and science and technology studies, but its focus is specifically on the life cycle of scholarly outputs—from the moment a researcher conceives a project to the long-term archiving of its results.
The central questions of the field are deceptively simple: How does knowledge gain credibility? How should it be made accessible? Who pays for the process, and who profits? And how do the answers to these questions change as technologies, institutional structures, and social norms evolve? These questions matter because scholarly communication is not merely a neutral conduit for facts. The systems that govern it determine which research gets funded, which findings are seen, which voices are amplified or silenced, and ultimately what counts as reliable knowledge in society.
To understand scholarly communication, it helps to break it into its constituent functions, each of which carries its own history and its own set of contemporary problems.
Registration is the process by which a piece of work is formally claimed by its author and assigned a place in the scholarly record. In the modern system, this usually means publication in a journal or by a press, but it can also mean deposition in a preprint server or a repository. Certification is the quality-control function, most commonly performed through peer review, in which independent experts assess the validity and significance of a submission. Dissemination covers the distribution of the work to its intended audience. Preservation ensures that the scholarly record remains accessible and intact over time. Finally, reward is the system of credit—citations, promotions, grants, prizes—that incentivizes scholars to participate in the first place.
These functions are not always aligned. The most persistent tension in the field is between certification and dissemination. Traditional journal publishing bundles them together: a journal's brand, built on rigorous peer review, is what makes its articles credible, and that credibility is what justifies the subscription price. But the bundling also means that access to certified knowledge is gated by payment. The open access movement, which emerged in the 1990s and gained momentum in the 2000s, argues that these functions can and should be separated—that certification can be preserved while dissemination is made free to read. The resistance to this separation, and the complex business models that have arisen to navigate it, form the central economic drama of the field.
The roots of scholarly communication lie in the seventeenth century, with the emergence of the first scientific journals. The Philosophical Transactions of the Royal Society and the Journal des sçavans, both founded in 1665, established a template that would endure for over three centuries: a periodic publication, edited by a learned society, containing letters and reports from researchers. These journals served a dual purpose that was then novel: they established priority of discovery (registration) and they subjected claims to a form of community scrutiny (certification). The letter format meant that communication was relatively fast, and the audience was a small, literate, largely aristocratic community of natural philosophers.
The nineteenth century brought the professionalization of science and the rise of the disciplinary journal. As fields became more specialized, journals multiplied, and the volume of research outpaced the capacity of any single editor to judge it. The formal system of peer review—sending manuscripts to external referees—emerged gradually in this period, though its adoption was uneven and its procedures varied widely. By the mid-twentieth century, the basic architecture was in place: commercial publishers and learned societies produced journals, university libraries bought subscriptions, and academics served as unpaid editors and reviewers, compensated in kind through the prestige of association.
The digital revolution of the late twentieth century disrupted this stable arrangement. The transition from print to electronic distribution did not, at first, change the underlying economics; publishers simply moved their subscription models online. But it created the technical possibility of near-zero-cost distribution, which in turn made the subscription model seem increasingly arbitrary. The physicist Paul Ginsparg's arXiv preprint server, launched in 1991, demonstrated that scholars could share research without publishers at all, at least in physics. The open access movement, formalized in the Budapest Open Access Initiative of 2002, articulated a vision in which research literature would be freely available online, with costs shifted from readers to authors, institutions, or funders.
This period also saw the rise of quantitative metrics as a form of scholarly evaluation. The Science Citation Index, developed by Eugene Garfield in the 1960s, made it possible to measure the impact of journals and articles through citation counts. The journal impact factor, originally designed as a library acquisition tool, became a proxy for research quality, with profound and often criticized consequences for academic careers. The later development of the h-index for individual researchers and the rise of altmetrics—measuring attention through social media, downloads, and news mentions—extended the logic of quantification while also complicating it.
The field of scholarly communication is not organized around a single paradigm but rather around several overlapping approaches, each with its own assumptions and methods. These approaches coexist and often inform one another, though they can also come into conflict.
The oldest and most quantitatively oriented approach is bibliometrics, the statistical analysis of publications and citations. Rooted in the mid-twentieth century, this tradition treats the scholarly literature as a data source from which patterns of scientific activity can be inferred. Its methods include citation analysis, co-authorship network mapping, and the measurement of research productivity and impact. Bibliometricians study the structure of scientific fields, the diffusion of ideas, and the emergence of new specialties.
The bibliometric approach has been enormously influential, but it carries significant limitations. Citation counts are an imperfect measure of quality, confounded by self-citation, negative citation (citing work to refute it), and field-specific differences in publication rates. The approach tends to privilege quantity and visibility over the content or significance of research. Moreover, its reliance on commercial databases like the Web of Science and Scopus means that its picture of scholarship is skewed toward English-language, high-prestige journals, potentially marginalizing research from the Global South, non-English traditions, and the humanities.
A second tradition, drawing on the sociology of science and science and technology studies, examines scholarly communication as a social process. Rather than treating publications as neutral data points, this approach asks how credibility is constructed, how scientific communities establish consensus, and how power operates within them. Key concepts include Robert K. Merton's norms of science—communalism, universalism, disinterestedness, and organized skepticism—and their later critique by scholars who argued that actual scientific practice often deviates from these ideals.
This tradition has produced influential theories of how scientific knowledge is validated. Bruno Latour and Steve Woolgar's laboratory studies showed that scientific facts are not discovered but constructed through a complex process of inscription, argument, and negotiation. The concept of "invisible colleges"—informal networks of researchers who communicate outside formal channels—explains how scientific communities actually coordinate. This approach is less concerned with optimizing the system than with understanding it, and it has been particularly useful in explaining why reforms to scholarly communication often fail: because they underestimate the deeply social nature of academic trust and reputation.
A third approach treats scholarly communication as a market and a policy problem. Its practitioners analyze the economics of journal publishing, the structure of the academic publishing industry, and the effects of different funding and access models. This tradition has documented the phenomenon of the "serials crisis"—the sustained rise in journal subscription prices outpacing library budgets—and has modeled the trade-offs between subscription, open access, and hybrid models.
The economic approach is closely tied to the open access movement, though it is not identical to it. Its key insight is that the traditional publishing model creates a market failure: researchers produce content for free, peer review it for free, and then buy it back through their institutions' library budgets. The result is a system in which a small number of commercial publishers control access to a large share of the world's research, earning high profit margins. Policy interventions—funding mandates, institutional repositories, transformative agreements with publishers—are the practical output of this approach.
A fourth, more recent approach focuses on the technical and organizational infrastructure that underlies scholarly communication. This includes the standards, protocols, and systems that make the scholarly record interoperable: digital object identifiers (DOIs), metadata schemas, repository software, persistent identifiers for authors (such as ORCID), and the emerging infrastructure for research data. This approach is less concerned with the economics or sociology of communication than with its plumbing.
The infrastructure approach has become increasingly important as the scholarly record has diversified beyond journal articles to include datasets, software, preprints, and other research objects. Its practitioners argue that the reliability of the scholarly record depends on the robustness of this underlying infrastructure, and that failures in infrastructure—link rot, data loss, incompatible formats—pose a greater threat to scholarship than any single business model. This perspective has gained urgency with the rise of "predatory publishing" and other forms of scholarly misconduct, which exploit gaps in the infrastructure of trust.
The current state of scholarly communication is best described as a period of transition and contestation. The traditional journal article remains the dominant form of scholarly output, and the major commercial publishers remain powerful. But the certainties of the print era have eroded, and no new settlement has yet replaced them.
Open access has moved from the margins to the mainstream, but its implementation has taken forms that its early advocates did not anticipate. The "gold" model, in which authors pay article processing charges (APCs) to make their work free, has created new inequities, as researchers without institutional funding may be unable to pay. The "green" model, in which authors deposit versions of their work in repositories, has been hampered by publisher embargoes and inconsistent compliance. The "diamond" model, in which journals are free to both authors and readers, often supported by institutions or societies, has grown but remains underfunded. Meanwhile, "transformative agreements" between libraries and publishers—contracts that bundle subscription payments with open access publishing fees—have been criticized as simply shifting money from one pocket to another while preserving publisher profits.
The preprint has become a significant force in several fields, particularly physics, computer science, and more recently the life sciences, where the COVID-19 pandemic accelerated its adoption. Preprints offer speed and openness, but they also raise questions about certification, since they have not undergone peer review. The pandemic also highlighted the dangers of unreliable information circulating in preprint form, leading to calls for more careful labeling and screening.
The evaluation of research is another site of active contestation. The San Francisco Declaration on Research Assessment (DORA), issued in 2012, called for an end to the use of the journal impact factor as a proxy for the quality of individual articles. The Leiden Manifesto, published in 2015, offered a set of principles for the responsible use of metrics. These initiatives reflect a broader recognition that the quantitative tools developed in the bibliometric tradition are being used in ways that distort scholarly behavior, encouraging "salami slicing" (dividing research into the smallest publishable units), citation gaming, and a focus on quantity over quality.
The scholarly record itself is expanding in scope. Research data, software, and other digital objects are increasingly recognized as legitimate scholarly outputs, requiring their own infrastructure for preservation and citation. The FAIR principles—findable, accessible, interoperable, reusable—have become a widely adopted framework for data management. This expansion raises fundamental questions about what counts as a "publication" and how the certification and reward systems should adapt.
Several debates run through the field and show no signs of resolution. The most fundamental concerns the proper relationship between the functions of scholarly communication. Should certification and dissemination be bundled or separated? The open access movement has largely won the argument that dissemination should be free, but the question of how to pay for certification remains unresolved. The APC model has created a two-tier system in which well-funded researchers can publish anywhere while others are relegated to less prestigious venues.
A second debate concerns the role of commercial publishers. Some argue that publishers provide essential services—editorial management, typesetting, platform infrastructure—that justify their costs. Others contend that the core functions of scholarly communication could be performed by the academic community itself, as demonstrated by the success of community-led initiatives like arXiv and the Public Library of Science (PLOS). The empirical evidence is mixed, and the debate is as much ideological as it is economic.
A third debate concerns the relationship between scholarly communication and the broader information ecosystem. The rise of social media, the decline of traditional gatekeepers, and the spread of misinformation have blurred the boundaries between scholarly and non-scholarly communication. Some scholars argue that the academy must engage more actively with public audiences, while others worry that such engagement comes at the cost of rigor and independence.
Finally, there is the question of equity. The scholarly communication system has historically been dominated by the Global North, English-language publishing, and elite institutions. Efforts to make it more inclusive—through multilingual publishing, support for research from the Global South, and attention to the needs of marginalized scholars—have made progress but face structural barriers. The field's own tools of analysis, such as citation databases, often reproduce these biases, making them difficult to study and correct.
Scholarly communication is thus a field defined by its problems as much as by its methods. It is not a settled discipline with a fixed canon, but a dynamic area of inquiry and practice that must continually adapt to changes in technology, economics, and the social organization of research. Its enduring value lies in its insistence that the way knowledge is shared is not a trivial detail but a fundamental determinant of what knowledge is produced, who can access it, and whose voices are heard.