CiteOwl
CiteOwl
Blog/Finding and reading sources

Finding and reading sources

How to find sources for a paper, tell credible from unreliable, confirm a journal is peer-reviewed, and read a research paper without reading all of it.

Most students start a paper by searching for something that sounds like their topic, then keep whatever appears first. That is why reference lists fill up with sources that sit next to the argument without supporting any specific sentence in it.

Better order: write the question first, then search for work that answers it. A research question tells you which keywords to use, which databases to search, and, once you find a paper, whether it is actually relevant.

Six search surfaces, and the hole in each one

"Search the literature" sounds like one activity. It is at least six, run over indexes that were built for different purposes and that miss different things. Picking the wrong one is not a small mistake: a surface's blind spot is invisible from inside it, because a search engine cannot show you what it never indexed.

Counts below marked as measured are what that index's own public API returned when we queried it on 11 August 2026. Counts marked as claimed are the vendor's own published figure, which is a different kind of fact and is treated as one.

SurfaceWhat it indexesSizeWhat it misses
Your library The discovery layer, usually Primo, Summon or EBSCO Discovery: everything your institution licenses, merged with open content, in one search box Depends entirely on what your library buys Anything your library does not license. It is also the only surface that knows what you can actually open, which is why it should usually be first
Google Scholar Its own description: "articles, theses, books, abstracts and court opinions, from academic publishers, professional societies, online repositories, universities and other web sites" Google publishes none, and offers no API Unknowable, and that is the problem. No coverage list means nobody outside Google can audit what is absent. It also mixes preprints, predatory journals and several versions of one paper without labelling them
OpenAlex An open index of works, authors, venues and citation links, built on Crossref plus repositories 323,998,640 works, measured Full text it has no open copy of. It can tell you a paper exists and cannot tell you what is on page 7
Crossref Metadata publishers deposited themselves, keyed by DOI 185,346,103 records, measured Anything nobody registered: many books, much of the humanities, a lot of non-English regional publishing
Semantic Scholar Papers from publisher partnerships, data providers and web crawls, with extracted citation context "over 200 million academic papers", claimed The same deposit gaps, plus whatever its crawl did not reach
Scopus and Web of Science Curated title lists. Clarivate says the Web of Science Core Collection indexes "22k+ peer-reviewed journals" and holds "99m+ records"; Elsevier says Scopus holds "over 217,000 book titles" from "more than 7,000 publishers" Vendor figures, claimed Everything off the list, on purpose. Both are subscriptions, so you have them only while you are enrolled

Sources for the vendor figures: Google Scholar's about page, Semantic Scholar, Clarivate and Elsevier, whose Scopus page states its numbers are "current as of July 2025". The measured counts came from OpenAlex and Crossref.

Field databases are missing from the table on purpose. PubMed for the life sciences, ERIC for education, IEEE Xplore for engineering and their equivalents are not competing on size, they are competing on a boundary: everything in them belongs to one field, and that is the whole product. If your topic has one, it belongs in your search alongside a general surface, never instead of one.

Which one to open first, by what you are actually doing

Where each guide fits

The twelve guides below run in the order the work does: search, judge, read, keep.

10 guides

All topics
How to find sources for a research paper

Where to look for sources, how to judge them fast with the CRAAP test, and how to confirm each one is real before you rely on it. The step most students skip.

15 min read
How to find peer-reviewed articles: 6 indexes counted

Where to find peer-reviewed work, from library databases to Google Scholar, JSTOR, Scopus, and PubMed, and how to confirm a journal is genuinely peer-reviewed.

19 min read
How to read a research paper: we measured 80

Stop reading top to bottom. The order researchers actually read in, the three-pass method, what to extract, and how to keep notes that survive until you write.

12 min read
How to write an annotated bibliography (with an example)

The two parts of every entry, the three things a good annotation does (summarize, evaluate, reflect on relevance), and a worked example you can copy the shape of.

12 min read
Literature review example (with a template you can copy)

A worked literature review example: a real thematic synthesis paragraph, taken apart to show why it works, plus a structure template and how to organize one.

9 min read
How to find a source for a claim you already wrote

Reverse-search a sentence you already wrote to find real backing, using free tools. And what to do when no source supports it.

11 min read
How many sources does a research paper need? 8 guides read

Nobody agrees and nobody shows their working. What university guides really say, where they contradict each other, and the rule that actually decides it.

12 min read
The paper I need is behind a paywall. 89 of 120 were free

The legal ladder, fastest first, with how often each rung works. Plus the tool half these guides still recommend that shut down in 2025.

16 min read

More on the blog

Real papers, found and read for you

Free to start. No card needed.

Start writing