Academic indexing is the invisible infrastructure behind almost every rule a researcher is told to follow — "publish in an indexed journal," "check if it's indexed," "your paper isn't indexed yet." Almost nobody explains what any of that actually means before handing out the instruction. This guide does.
By the end, you will be able to explain what academic indexing is, how it actually works inside the databases that matter most (Scopus, Web of Science, PubMed/MEDLINE/PMC, Google Scholar, Crossref, OpenAlex, DOAJ), what an "indexed journal" genuinely guarantees and does not guarantee, and — most importantly — how to independently verify any indexing claim yourself, without having to trust a journal's marketing. This is a foundations article, deliberately written before any single-database deep dive.
Picture a library holding several million books with no catalogue. To find anything, you would have to walk every aisle. A catalogue solves that problem by recording where each book is and what it is about, so the library can be searched instead of wandered.
Academic indexing performs the same function for research literature — at a scale no human catalogue could ever manage by hand. A database such as Scopus, Web of Science, or PubMed receives new articles from thousands of journals, records structured information about each one — title, authors, subject terms, and often citation links — and makes that record searchable.
When someone says a journal is "indexed" in one of these systems, the precise meaning is: that specific database has evaluated the journal against its own published criteria, accepted it, and is now cataloguing what it publishes going forward.
Two things follow immediately from that definition, and most of the confusion around indexing comes from missing them.
First, indexing is never a universal status. It is always relative to one named database. A journal can be indexed in one system and simply absent from another — not because anything is wrong, but because each database runs an entirely separate evaluation against its own criteria.
Second, indexing is an evaluation of the journal as a system — its editorial process, its regularity, its transparency — not a re-review of the scientific validity of every individual article it publishes. Those are related but genuinely different questions, and conflating them is the single most common mistake in how "indexed" gets used in everyday academic conversation.
Getting these words right matters more than it looks like it should, because loose usage is exactly what makes indexing claims easy to misrepresent.
| Term | What it actually means |
|---|---|
| Index / indexing service | An organization or system that systematically catalogues publications so they can be searched. |
| Bibliographic database | Records structured metadata (title, authors, journal, date, subject terms) about a document — with or without holding the full text. |
| Citation database | A bibliographic database that additionally tracks which papers cite which, making citation counts and citation-based metrics possible. |
| Abstracting service | Produces a condensed summary of a work. Historically distinct from indexing; now usually delivered together as "abstracting and indexing." |
| Full-text archive / repository | Stores the actual article text itself, not just a record pointing to it (PubMed Central is the clearest example). |
| Discovery system / search engine | Finds and surfaces scholarly content, sometimes through automated crawling rather than curated evaluation (Google Scholar). |
| Metadata infrastructure | Registers and manages structured publication data — such as DOIs — without necessarily evaluating content quality (Crossref). |
| Inclusion / selection criteria | The published standards a database uses to decide which journals it will index. |
| Coverage | How much of a field, region, or language a database actually captures. No database covers everything published anywhere. |
None of these words are synonyms for each other, even though everyday conversation treats them that way. A "database" is not automatically an "index" in the curated sense; a "search engine" is not automatically a "database" with published selection criteria; a "repository" is not automatically an "index" at all. Sections 5–8 make this concrete with real systems.
The exact mechanics differ between databases (Sections 5–8 cover Scopus, Web of Science, and PubMed/MEDLINE/PMC individually, because each genuinely works differently), but the general shape of the process is shared:
Two details in that chain matter more than they might seem to at first glance. The evaluation happens before most future content is indexed, but databases keep watching afterward — indexing is a relationship, not a one-time medal. And when a journal is discontinued, what happens to already-indexed content varies by database and is a genuinely separate question from what happens to future content — covered precisely in Section 5, because Scopus is unusually explicit about it in its own public policy.
There is no single master list of "all academic indexes." Wikipedia's own list of academic databases and search engines runs to roughly 180 entries and explicitly presents itself as representative rather than exhaustive — and even that page acknowledges that the line between "database" and "search engine" is blurry for many of the systems on it. This is not a gap someone forgot to fill in; the ecosystem is genuinely fragmented because different systems solve different problems for different audiences.
| Category | What it fundamentally does | Curated? | Example |
|---|---|---|---|
| Curated citation index | Evaluates journals against published criteria; tracks citations between indexed items | Yes — human/editorial | Scopus, Web of Science Core Collection |
| Curated bibliographic database (discipline-specific) | Applies its own field-specific selection criteria; catalogues metadata | Yes | MEDLINE (within PubMed) |
| Automated discovery search engine | Crawls content meeting technical criteria; no human review of individual titles | No — algorithmic | Google Scholar |
| Metadata registration infrastructure | Registers persistent identifiers (DOIs) and associated metadata for members | No — membership registration | Crossref |
| Open scholarly graph | Aggregates and links data from many sources into one open dataset | Partially — inherits sources' curation | OpenAlex |
| Open-access journal directory | Vets open-access journals against transparency and publishing-practice criteria | Yes | DOAJ |
| Full-text archive / repository | Stores complete article text, sometimes independent of journal-level curation | Varies | PubMed Central, institutional repositories |
The takeaway: calling every one of these an "index" without qualification erases real, functional differences that change what a researcher should actually do with each one.
| System | Operator | Type | Selection model | Citation tracking | Full text? | Access | What inclusion does NOT mean |
|---|---|---|---|---|---|---|---|
| Scopus | Elsevier | Curated citation database | CSAB evaluation | Yes | No (links out) | Commercial | Every article inside is error-free |
| Web of Science Core Collection | Clarivate | Curated citation database | 28-criteria evaluation | Yes | No (links out) | Commercial | Automatic Impact Factor |
| PubMed | NLM/NIH | Search interface | Includes MEDLINE + more | Limited | No (links out) | Free | The article is MEDLINE-indexed |
| MEDLINE | NLM | Curated bibliographic database | NLM selection + MeSH | No | No | Free (via PubMed) | NLM peer-reviewed the article |
| PubMed Central | NLM | Full-text archive | NLM selection + funder mandates | No | Yes | Free | The journal is MEDLINE-indexed |
| Google Scholar | Automated search engine | Automated crawling | Yes (less precise) | Links out | Free | Any curated evaluation occurred | |
| Crossref | Crossref (non-profit) | Metadata infrastructure | Membership-based | Via Cited-by | No | Free / membership | Any quality assessment occurred |
| OpenAlex | OurResearch (non-profit) | Open scholarly graph | Aggregates open sources | Yes (derived) | Links out | Free / open | Scopus/WoS-equivalent curation |
| DOAJ | DOAJ (non-profit) | Curated OA directory | Published OA criteria | No | Links out | Free | High citation impact or an Impact Factor |
What it is. Scopus is Elsevier's abstract-and-citation database — a curated, multidisciplinary index of journals, conference proceedings, books, and patents, alongside author and institution profiles built from that indexed content.
Who decides what gets in. Scopus is owned and operated by Elsevier, but journal-inclusion decisions are made by the Content Selection and Advisory Board (CSAB): an international body of subject-matter experts organized around roughly 17 subject chairs, responsible for reviewing every title suggested for inclusion. Four regional Expert Content Selection and Advisory Committees (ECSAC) — covering China, Thailand, Russia, and South Korea — screen regional journals before CSAB review. Board members recuse themselves from evaluations where they have a personal or consulting interest.
The technical gate. Before substantive review, a journal must show: peer review with a publicly available description, a registered ISSN, a demonstrable and regular publication history, English-language titles and abstracts, and a clear public statement on publication ethics and malpractice. Current policy also expects journals to disclose their generative-AI usage policies for content creation and peer review.
| Dimension | What it assesses |
|---|---|
| Journal policy | Editorial quality, peer-review type, diversity of editors/authors |
| Content | Academic contribution, abstract clarity, fit with stated scope |
| Journal standing | Citation metrics and editor reputation |
| Publishing regularity | Consistent, uninterrupted publication schedule |
| Online availability | Full accessibility, English homepage, site quality |
Re-evaluation, suspension, and delisting. Inclusion is not permanent. Scopus re-evaluates indexed journals when triggered by community concern or by Elsevier's own anomaly-detection models (unusual publication spikes, citation patterns, or co-authorship networks). During re-evaluation, new content stops being indexed until CSAB decides. If discontinued, already-indexed content is preserved as part of the scientific record — removal happens only in exceptional, proven cases of severe unethical practice.
Metrics from Scopus data — not interchangeable:
None of these three is the Journal Impact Factor — that belongs exclusively to Clarivate's Journal Citation Reports (Section 6).
How to verify: use Scopus's own official Sources search tool — never a journal's self-description or a third-party aggregator. Search the exact title and confirm active, current coverage.
What it is. Web of Science is Clarivate's citation-indexing platform. Its Core Collection — the curated component most people mean by "Web of Science" — covers more than 22,000 peer-reviewed journals indexed cover-to-cover, across 254 subject areas, with more than 2.4 billion cited references and over 97 million total records (Clarivate, official Core Collection page, accessed September 2026).
The evaluation. Clarivate applies a published set of 28 criteria: 24 quality criteria (editorial rigor, board composition, ethical publishing standards) and 4 impact criteria (citation-based influence). Evaluation runs through staged triage — ISSN/publisher checks, editorial triage, quality evaluation, and (for journals that clear the quality bar) impact evaluation.
Journal Citation Reports and the Impact Factor. The JCR is a separate Clarivate product built from Web of Science citation data — the sole source of the Journal Impact Factor (citations in year Y to items published in the two preceding years, divided by citable items published in those two years). Per the 2026 Journal Citation Reports (released 17 June 2026), coverage spans 22,643 journals, and 521 journals received a Journal Impact Factor for the first time, across 47 countries and regions, with 58% based outside the US and Western Europe (CASRAI summary of Clarivate's 2026 release).
How to verify: check Clarivate's official Master Journal List — not a journal's own claim of "SCI-indexed," a phrase frequently used loosely outside Clarivate's own terminology.
This is one of the most consistently confused relationships in the entire ecosystem — with an unusually clean, authoritative answer directly from the U.S. National Library of Medicine (NLM), which operates all three systems.
PubMed is a free search interface covering more than 40 million citations and abstracts (NIH/NLM official "About PubMed" page). PubMed itself does not contain full text; it links out where available.
MEDLINE is, in NLM's own words, "the largest component of PubMed" — a curated bibliographic database of citations from journals NLM has specifically selected, indexed with MeSH (Medical Subject Headings).
PubMed Central (PMC) is a separate full-text archive holding articles from NLM-selected journals plus individual funder-mandated deposits, regardless of whether the source journal is otherwise indexed anywhere.
How to verify: MEDLINE status is checked in the NLM Catalog, filtered specifically for "currently indexed for MEDLINE" — not a general name search, since the Catalog also lists journals indexed only in the past.
These four get grouped with Scopus and Web of Science constantly — calling all six "indexes" without qualification erases real differences.
Applies no curated selection comparable to Scopus or Web of Science. Content becomes discoverable through automated crawling of sites meeting technical and format requirements — no human editorial review of individual submissions. Google's own documentation acknowledges it cannot guarantee every reference is identified correctly. Use it for: broad, fast discovery and citation tracking — never as evidence of curated indexing status.
Not a quality index — scholarly metadata infrastructure. A non-profit membership organization registering DOIs and managing associated metadata. Its own documentation describes no journal-quality evaluation function. Use it for: verifying a DOI resolves to a real, registered publication — not as a legitimacy signal on its own.
A comprehensive, open index of the global research system, built from Crossref and other open sources, covering works, sources, authors, and institutions. It aggregates rather than gatekeeps — no independent quality-evaluation layer on top of what it ingests. Use it for: free, open bibliometric work and cross-checking commercial-database figures.
A genuinely curated index, but for a narrower purpose: separating legitimate open-access journals from illegitimate ones. Criteria include a minimum publishing history, a real editorial board, peer review by at least two independent reviewers, a registered ISSN, and limits on "endogeny" (no more than 25% of articles from a journal's own editors/reviewers). DOAJ explicitly disclaims using Impact Factor or bibliometrics to judge quality. Use it for: checking whether an open-access journal meets a baseline transparency standard.
"Is this journal indexed?" is, on its own, an incomplete question. The complete, useful version is always:
Consider three entirely ordinary, non-contradictory situations:
None of these are contradictions or red flags. They're the ordinary result of a fragmented ecosystem where different systems evaluate different things. The mistake is assuming "indexed" ever meant "indexed everywhere, permanently, by every standard."
No source examined for this article — official or independent — supports the claim that indexing guarantees quality. The distinction worth holding onto: database-level selection versus article-level scientific validity. A database evaluates a journal's process; it does not re-review whether any specific paper's methodology or conclusions hold up. That's what peer review, post-publication scrutiny, and — when something goes wrong — corrections and retractions are for.
This is exactly why a long-indexed, reputable journal can still contain papers with genuine methodological weaknesses, papers that later received a correction, papers carrying an expression of concern, or papers that were ultimately retracted. None of this means the indexing system "failed." A retraction is the integrity system working as intended.
| Concept | What it evaluates | Who determines it |
|---|---|---|
| Indexing | Editorial process, technical practice, standing | The database's own evaluators |
| Peer review | Whether manuscripts are independently reviewed | The journal's own editors and reviewers |
| Journal quality | Holistic reputation, rigor, transparency | No single authority — community judgment |
| Journal Impact Factor | A specific citation ratio | Exclusively Clarivate, via JCR |
| Indexing DOES indicate | Indexing does NOT indicate |
|---|---|
| The journal met a database's published criteria at evaluation | Every article is methodologically sound |
| The journal has a described peer-review process | Each review was rigorous or error-free |
| The journal is discoverable through that database | It's discoverable through every other database too |
| (For some databases) the journal met citation-impact criteria | The journal automatically has an Impact Factor |
| The journal was in good standing at last evaluation | It's in good standing today, without a fresh check |
Illegitimate journals frequently misrepresent indexing status because "indexed" carries real institutional weight and most readers never check claims directly. Common patterns:
Use this every time a journal's indexing status matters to a real decision.
| # | Step | Why it matters |
|---|---|---|
| 1 | Record the exact journal title | Similarly named journals are a known impersonation tactic |
| 2 | Record the ISSN / eISSN | More reliable than title searching |
| 3 | Identify the publisher | A mismatch is a warning sign |
| 4 | Identify the exact database claimed | "Indexed" alone is not enough |
| 5 | Go to that database's own official lookup tool | Only the database's own record is authoritative |
| 6 | Confirm current, active status | A past acceptance isn't a current guarantee |
| 7 | Check coverage start/end dates | Shows exactly which years are covered |
| 8 | Note the specific collection/category | ESCI vs. SCIE, MEDLINE vs. PubMed-only carry different meanings |
| 9 | Check for discontinued/delisted status | A journal can advertise a withdrawn acceptance |
| 10 | Record the date of verification | Status can change; an undated check ages fast |
| 11 | Save evidence (screenshot or saved record) | Protects you if the record later changes |
| 12 | Repeat for every additional database claimed | Claims compound; each needs its own check |
Why ISSN beats title-only searching: titles can be nearly identical between a legitimate and an impersonating journal, can change when a journal rebrands, and can be entered inconsistently across systems. An ISSN is a standardized identifier tied to one continuing publication.
Indian academic regulation has repeatedly tied career and degree outcomes directly to indexing status, making independent verification a practical skill for Indian researchers.
One documented example: earlier UGC PhD regulations (2016-era) required a scholar to present at two conferences and publish at least one paper in a refereed journal before thesis submission. Independent Indian education-news reporting from November 2022 describes UGC subsequently withdrawing that mandatory pre-submission publication requirement. That same reporting notes quality publications still carry weight elsewhere in the system — for instance, in recruitment scoring for academic posts.
The durable lesson: a requirement to publish in "an indexed journal" is never self-verifying. This is precisely why "indexed somewhere" proved too vague in Indian regulatory history — the since-discontinued UGC-CARE list was created in response. It should not be presented or relied upon as an active, currently operational resource today.
None of this fragmentation is a flaw waiting for a "unified" fix. It reflects real, different missions no single system could serve equally well.
Each figure states what it measures, its source, and access date — deliberately short rather than padded.
These figures are not directly comparable to each other — they measure different things and were published at different times.
Yes. Every major curated database reserves the right to re-evaluate and discontinue coverage.
Depends on the database. Scopus's published policy preserves already-indexed content as scientific record while halting future indexing — removal happens only in exceptional, proven cases.
Each database applies different selection criteria and counts citations only from sources it itself indexes.
Databases generally track the new title linked to the old one via ISSN. A publisher change can trigger a fresh re-evaluation, since ownership changes have been associated with quality or integrity shifts.
Because indexing is a status with a shelf life, not a permanent fact — a check today says nothing certain about a year from now.
The process by which a database evaluates a journal against its own published criteria and, if accepted, catalogues its content on an ongoing basis.
No. Indexing evaluates process and standing at a point in time — it does not certify the scientific validity of every article (Section 10).
Yes — curated and multidisciplinary, with inclusion decided by the CSAB (Section 5).
Yes — Clarivate's Core Collection is a curated set of citation indexes: ESCI, SCIE, SSCI, AHCI (Section 6).
No. PubMed is the search interface; MEDLINE is NLM's separately curated database supplying most of what's searchable there (Section 7).
No. PMC is a separate full-text archive; PubMed holds no full text itself (Section 7).
No — an automated search engine with no human editorial evaluation of individual titles (Section 8).
No. A Crossref-registered DOI confirms a metadata relationship, not a quality evaluation (Section 8).
A genuinely predatory journal is very unlikely to pass a real database's evaluation — but illegitimate journals routinely falsely claim indexing they never had (Section 11).
Yes — in every major database discussed here (Section 16).
No. Impact Factor is a specific Clarivate/JCR metric; a journal can be legitimately indexed elsewhere and simply not carry one (Section 6).
Indexing is infrastructure, not a verdict. It exists to solve a real, mechanical problem — too much published research for anyone to track by browsing — and every major system solves that problem differently, for different audiences, under different published criteria. None of them evaluates the scientific validity of every individual article they carry, and none of them is a stand-in for the others.
The one habit worth carrying forward from this entire guide is the discipline of asking the complete question — indexed where, under what criteria, as of when — and then actually going to check, using the workflow in Section 12, instead of trusting a claim made anywhere except the database's own official record.
Every rule in Indian and global academia that mentions "an indexed journal" is really asking a researcher to trust a claim. The entire argument of this guide is that trust should be replaced with a two-minute check: name the database, go to its own official record, and look for yourself. That habit — verification over marketing, evidence over reputation — is the same habit this guide asks you to bring to every other claim in the research ecosystem: a metric, a citation, a retraction status, an AI-generated reference. Understanding what indexing actually is isn't trivia. It's the first rehearsal of the exact skill every researcher eventually needs for everything else.