The Analyst's Evidence Capture and Due-Diligence Archive
The analyst's evidence capture and due-diligence archive explained — a practical Capture→Connect→Create workflow for analysts, lawyers, and finance
Persona Playbooks
The academic researcher's knowledge workflow explained — a practical Capture→Connect→Create system for PhD candidates and researchers who are drowning in
If you're a PhD candidate or active researcher, you've felt the specific pain of academic knowledge work: you've read thousands of papers, but you can't find the one that made the argument you need. You've saved hundreds of articles, but half the links are dead and the other half are buried in a folder you can't navigate. You've spent weeks on a literature review only to discover in revision that you missed a key paper in the same subfield.
The academic researcher's knowledge workflow is the system for ending that pain — a deliberate Capture→Connect→Create loop designed for the specific challenges of academic research: high source volume, precise citation requirements, long project timelines, and the need to synthesize not just collect.
The academic knowledge workflow has three stages:
Capture: Collect sources before they disappear, with enough metadata to find them again.
Connect: Organize your sources into a navigable knowledge base where synthesis is possible.
Create: Turn your accumulated knowledge base into the outputs you actually need: literature reviews, papers, grant applications, presentations.
Most researchers are good at Capture (they save things) and bad at Connect (they can't find what they saved) and Create (the synthesis step is painful because the Connect step was never done). The fix is building the Connect infrastructure at capture time, not as a retroactive project.
The capture problem for academic researchers isn't volume — it's specificity. You're not just saving things; you're building an evidence base for claims you'll make years later. A PDF with a meaningful filename in a topic folder is not enough.
What to capture:
Tools for capture:
The capture habit: When you open any source you're actually reading (not just scanning), capture it before you close it. A PDF goes into Zotero. A web source goes into WebSnips. An idea from a seminar goes into your fleeting notes. This 30-second habit prevents the "I read something about this somewhere" problem from compounding.
Saved sources become a knowledge base only when they're connected — to each other, to your arguments, and to the structure of your research.
The three Connect actions:
1. Write literature notes After reading each source, write a brief note (150–300 words) in your own words: the main argument, the method, the key finding, and — most importantly — why it's relevant to your research. Store these with the citation in Zotero (notes field) or in a dedicated PKM tool.
This is the step most researchers skip. They save the paper but don't write the note. Then six months later, when they need to cite it, they can't remember what it said and have to re-read it. Literature notes prevent re-reading.
2. Add tags and connections Tag each source by topic, method, period, argument type, and relevance to your specific research questions. Cross-reference: "this paper's methodology is critiqued by Smith 2019." These connections are what make later synthesis possible.
3. Build your argument structure Don't wait until the writing phase to organize your sources into arguments. As you read, note which papers support which claims, which ones are in tension, which ones you'll need to engage with. A simple outline with linked citations is more useful than an unstructured library.
Tools for Connect:
The Create stage is where the investment in Capture and Connect pays off. If you've done the first two stages well, the literature review, discussion section, or grant application is largely an assembly task — selecting and sequencing the evidence you already have — rather than a research task.
Literature reviews A literature review written from a well-organized Zotero library with literature notes is a different task than one written from a folder of PDFs. You're selecting and synthesizing what you already know, not re-discovering it. The notes you wrote when you read the papers become the raw material of the review.
Research papers By the time you're writing a paper, your argument structure (built in the Connect stage) becomes the outline. Each claim has already been tagged with the sources that support it. The writing is about articulating the argument; the sourcing is already done.
Grant applications Grants require synthesizing the state of a field and your contribution to it. A well-maintained research library makes this synthesis much faster than starting from memory.
Dr. A is a third-year PhD candidate in public health studying maternal health outcomes.
Morning — capturing a new paper:
She finds a relevant 2025 paper in PubMed through her Google Scholar alert. One click with Zotero's browser extension imports the citation and the PDF (full text available through her institution). She opens it, reads it, and writes a 200-word literature note: "Chen et al. 2025 argues that prenatal care access disparities account for 40% of the Black-white maternal mortality gap in US urban areas. Key limitation: cross-sectional design. Relevant to: Chapter 2 (structural drivers), may conflict with Gupta 2023 on insurance access as primary driver." She tags it: maternal-mortality, structural-racism, urban-US, cross-sectional.
Afternoon — writing a literature review section:
She opens her Zotero collection for "Structural Drivers" (the section she's writing) and filters by tag cross-sectional to see methodological limitations she should note. She opens her Obsidian notes, searches "insurance access as primary driver," and finds her note connecting Gupta 2023 to Chen 2025 — a point she wrote when she first read Gupta but which surfaced now because her tags created the connection. She writes the paragraph engaging both papers with this methodological tension.
Total research overhead for capturing and noting the new paper: 20 minutes.
What she avoided: Re-reading Chen 2025 in six months when writing this section. Not knowing about the connection to Gupta. Missing the methodological limitation in her review.
| Stage | Tool | What it does |
|---|---|---|
| Capture (academic) | Zotero + browser plugin | Imports citations and PDFs from databases |
| Capture (web) | WebSnips | Saves government reports, news, preprints, gray literature with full content |
| Literature notes | Zotero notes or Obsidian | Stores your synthesis of each source |
| Argument structure | Notion or Obsidian | Organizes sources into claims and outline |
| Writing | Word or Google Docs with Zotero plugin | Auto-inserts formatted citations |
| Citation formatting | Zotero | One-click style changes (APA/Chicago/Vancouver) |
WebSnips is specifically useful for the web source layer — the category of sources most vulnerable to link rot and most absent from academic database importers. Policy documents, preprints, government data, news coverage of research events, and agency reports need a different tool than academic PDF management.
Saving without noting. A folder of 500 PDFs is not a knowledge base. Without literature notes, you'll re-read everything when you write.
Under-tagging. "2026 papers" and "Chapter 2" are not useful tags. Tag by topic, method, argument, relevance, and — if relevant — quality flag. Tags you write at capture time are the index you'll search at write time.
Assuming library links will work. PDFs change. DOIs that resolve today may 404 in two years (institutional access changes, journal transfers, publisher mergers). Save full text at capture time. This applies especially to preprints, datasets, and web sources.
Waiting for the Connect step. Most researchers wait until they're "ready to write" to organize their sources. By then, they're doing this organizational work under deadline pressure with incomplete memory of what they read. The Connect step should happen at capture time, not just before writing.
Skipping the argument structure. Saving sources into a flat library and hoping the structure will emerge during writing is the source of the "I have everything but I don't know what to say" feeling. The argument structure — what sources support which claims — should be built as you read.
The academic researcher's knowledge workflow is not a complex system — it's a set of deliberate habits applied consistently at capture time. The investment is 20–30 extra minutes per source. The return is literature reviews that take days instead of weeks, writing sessions that flow because the synthesis is already done, and a research library that remains useful across the full arc of a PhD rather than becoming a graveyard of forgotten PDFs.
See also: Clip Articles for Later Reading.
More WebSnips articles that pair well with this topic.
The analyst's evidence capture and due-diligence archive explained — a practical Capture→Connect→Create workflow for analysts, lawyers, and finance
The consultant's knowledge workflow explained — a practical Capture→Connect→Create system for knowledge workers and consultants who need to manage
The course creator's source library explained — a practical Capture→Connect→Create system for educators and course creators who want to curate sources
The developer's second brain explained — a practical Capture→Connect→Create knowledge system for engineers managing snippets, documentation, ADRs, and
The founder's knowledge stack explained — a practical Capture→Connect→Create workflow for founders and solo operators juggling competitive intelligence
The marketer's swipe file and competitor monitoring system explained — a practical Capture→Connect→Create workflow for marketers, SEOs, and growth folks