AcademicPublishing

22 chunks

Budapest Open Access Initiative

The Budapest Open Access Initiative (BOAI), released February 14, 2002, is the founding declaration of the {{Open Access}} movement. Convened by the Open Society Institute in December 2001, it defined OA, distinguished 'gratis' from 'libre' access, and named self-archiving plus OA journals as the two core strategies.

93%
16

Citation Needed

Wikipedia's [citation needed] tag, introduced in 2006, marks unsourced claims as public IOUs and operationalizes the project's verifiability policy. As of mid-2025 it appeared on more than 604,000 pages.

93%
11

arXiv

arXiv is the dominant {{preprint}} server for physics, mathematics, computer science, and adjacent fields. Founded in August 1991 by physicist Paul Ginsparg at Los Alamos and now hosted at Cornell, it functions as a free repository where researchers post manuscripts before, alongside, or instead of formal journal publication.

93%
3

Preprint

A preprint is a version of a scholarly paper shared publicly before formal peer review and journal publication. Preprint servers such as {{arXiv}}, bioRxiv, medRxiv, and SSRN let researchers establish priority, gather feedback, and disseminate results quickly. The model is central to {{Green Open Access}} and surged in visibility during COVID-19.

93%
12

Digital Object Identifier (DOI)

A DOI is a persistent identifier used to label digital objects such as journal articles, datasets, and software, so that they remain citable even when their location on the web changes. Maintained by the International DOI Foundation and assigned through registration agencies such as Crossref and DataCite, DOIs have become foundational infrastructure for scholarly communication.

92%
0

Article Processing Charge (APC)

An Article Processing Charge (APC) is a fee paid to a publisher to make a scholarly article {{Open Access}}. APCs shift the cost of publication from readers to authors, institutions, or funders. Typical APCs run $1,000 to $5,000, with elite journals charging up to roughly €9,500.

92%
8

Mata v. Avianca / Fabricated Citations

A 2023 U.S. District Court case (S.D.N.Y., Judge P. Kevin Castel) in which plaintiff's attorneys submitted a brief built from ChatGPT-fabricated case citations and were sanctioned $5,000 for bad-faith conduct. The case became the canonical example of LLM citation hallucination causing real legal consequences and has been followed by hundreds of similar sanctions across U.S. courts.

92%
9

DataCite

DataCite is a global nonprofit registration agency for DOIs assigned to research outputs that are not traditional papers, especially datasets, software, samples, and models. It complements Crossref's coverage of articles and books and is central to efforts to make data citable and reusable.

92%
0

Persistent Identifier

A persistent identifier (PID) is a long-lived, location-independent label for a digital or physical resource. PIDs power systems like DOI, ORCID, ROR, and Handle, replacing fragile URLs with strings that resolve to whatever location currently hosts the thing they name.

92%
0

Citation

A citation is an explicit pointer from a claim to the source that supports it. Citations let readers verify claims, give credit to contributors, and let later corrections propagate backward to every document that inherited a source.

92%
6

Wikipedia Verifiability Policy

Wikipedia's verifiability policy holds that claims must be attributable to reliable published sources, not based on editors' personal knowledge. It is one of three core content policies alongside no-original-research and neutral point of view.

92%
12

Plan S

Plan S is a 2018 initiative by {{cOAlition S}}, a group of national research funders and philanthropies, requiring that research they fund be published in compliant {{Open Access}} venues with no embargo. The 'S' stands for 'shock', and the policy capped {{Article Processing Charge}}s, mandated author copyright retention, and phased out hybrid journals.

92%
10

Crossref

Crossref is a nonprofit DOI registration agency that provides persistent identifiers and shared metadata for scholarly publications. Founded in 2000 by a consortium of publishers, it has grown into the largest single source of DOIs, underpinning citation linking, reference resolution, and many open-science services.

92%
0

ORCID

ORCID is a persistent identifier for individual researchers, designed to disambiguate authors whose names are common, change, or appear in multiple writing systems. Operated by a nonprofit and integrated with DOI registration agencies, it links a researcher to their publications, datasets, grants, and affiliations.

92%
0

Open Access in Academic Publishing

Open Access (OA) is a movement to make peer-reviewed scholarly research freely readable and reusable online. Defined by the {{Budapest Open Access Initiative}} (2002) and reinforced by the Bethesda and Berlin declarations (2003), OA splits into Gold (publisher-side, often funded by {{Article Processing Charge}}s), Green (author self-archiving in repositories like {{arXiv}} and PubMed Central), and Diamond (no fees on either side). Funder mandates and {{Plan S}} push compliance, while {{Sci-Hub}} acts as an illegal counterforce demonstrating unmet demand.

92%
7

Footnote

The footnote is the scholarly device that anchors a claim in a written narrative to its supporting source. Anthony Grafton's history argues it is the form through which modern historical scholarship became auditable.

91%
0

SPJ Code of Ethics

The Society of Professional Journalists' Code of Ethics is a non-binding set of principles emphasizing accuracy, independent verification, source identification, and accountability. It is one of the most widely referenced ethics frameworks in U.S. journalism.

90%
1

Sci-Hub

Shadow library founded in 2011 that provides free access to paywalled academic research papers — a high-profile counter-institution to subscription-based scholarly publishing.

90%
10

Citation Laundering

Citation laundering is the practice of citing a primary source via a secondary one without actually reading the primary. It propagates misreadings, inflates the cited paper's perceived support, and is treated as a citation-integrity violation under multiple ethics frameworks.

88%
1

Source Attribution in LLM Outputs

Large language models present a unique attribution problem: their weights compress vast amounts of training text into statistical patterns that erase provenance, so a base model cannot reliably say where any given claim came from. Major systems work around this with retrieval layers — Perplexity inlines numbered citations to live web results, Google's AI Overviews append footnote-style source links, Anthropic's Citations API grounds answers in user-supplied documents, ChatGPT cites only when its browse tool is active — while research into training-data attribution, prompt provenance, and citation hallucination tries to close the gap between cited and uncited generation.

83%
19

Citation Cartel: Self-Citation Padding in Academic Papers

A {{citation cartel}} is a pattern where research groups inflate their footprint by densely cross-citing their own prior work, often beyond what the argument requires. It games {{Google Scholar}} metrics and signals weak independent validation.

83%
16

Why Citation Accuracy Matters in Knowledge Ecosystems

Editorial argument that citation is the load-bearing infrastructure of cumulative knowledge: it enables backward error correction, aligns credit with contribution, and lets readers verify claims independently. Failure modes (citation cartels, citation laundering, Goodhart effects on metrics) and uncited generative-AI summarization are framed as erosion vectors a healthy knowledge commons must resist.

72%
13