Skip to main content

Extract structured data from academic papers, faster.

Turn collections of academic PDFs into structured datasets using the extraction questions that matter to your review.

Provide your papers and your extraction framework, and receive an Excel workbook with an answer and a short reasoning note for every question, ready for you to check.

Academic PDFs

Extraction questions

Structured extraction

Excel output

Researcher review

Spend less time extracting data from papers by hand.

Data extraction for a systematic review usually means locating the same types of information (study design, sample size, outcomes) across dozens or hundreds of papers, one at a time.

Systematically Reviewed lets you define your extraction framework once, in a spreadsheet, and applies it to your full set of papers. Instead of a blank table to fill in, you start from a completed one to check and correct.

How it works

Your extraction questions, applied consistently across your papers.

  1. 01

    Define what you need

    The template workbook has two sheets, paper_level_data and variable_level_data. You decide whether to use one or both.

    • Paper level (one row per paper)

      Questions asked once about the whole paper. If this is all you need, fill in this sheet alone.

      e.g. “What study design was used?

    • Variable level (one row per item)

      Use this sheet when a paper reports several models, treatments, or exposures and you want a row for each. Leave it empty to keep one row per paper.

      e.g. “What sample size was used for {variable_name}?

    Download the template workbook (.xlsx)
  2. 02

    Provide your papers

    Supply your papers as PDF files. Extraction covers the main text as well as tables and figures, and records only information explicitly stated in each paper.

  3. 03

    Receive structured results

    You receive a single Excel workbook with one row per paper, or one row per item when you use variable-level questions. Each question appears as a column paired with a reasoning column explaining where the answer was found.

  4. 04

    Review and verify

    Use the provided cross-checking application to review results one paper at a time. Edit extracted values, view reasoning notes, with the source PDF side by side, and mark each field as correct or incorrect before export.

From extraction questions to a usable dataset.

Each question becomes a column in your workbook, paired with a reasoning note.

Both layouts come from the same template: paper-level questions alone, or with variable-level questions added so that each model, treatment, or exposure has its own row.

Output row layout

Showing 2 rows.

Illustrative example
pdf_namestudy_designvariable_abstraction
Rodriguez_2026Retrospective cohort (target trial emulation)ADM trial; AOM trial; MAUD-T2D trial; MAUD-obesity trial
Chen_2024Propensity-score matched cohortSemaglutide; Tirzepatide

The variable_abstraction question returns a semicolon-separated list, and each listed item becomes its own row with the paper-level answers repeated. Every question is paired with a reasoning column; one is shown here.

Designed to be checked.

Extraction is AI-assisted, the program allows you to easily verify and update results before relying on them.

After delivery, you review results in a web application. You will receive an email with a link to the application with your results. Load the Excel workbook, work through one paper at a time, with the matching source PDF side by side. Each extracted field shows the value alongside a read-only reasoning note from the extraction, and shortcut to see exactly where the answer was found.

Edit any value, mark fields as correct or incorrect, and export your reviewed data with a parallel correctness grid. All processing happens locally in your browser; nothing is uploaded during verification.

Paper extraction review

Review one paper at a time, edit cells, mark each field correct or incorrect, then export.

Rodriguez_2026

Correct: 1 · Incorrect: 0 · Pending: 1 · Rows complete: 0/2

1 / 2 variables reviewed

← Previous paperPaper 1 of 3 · Variable 1 of 2Next paper →

Viewing: Rodriguez_2026.pdf

Source PDF opens alongside extracted fields on larger screens.

Variable-level data

hazard_ratio

0.73

Reasoning

Hazard ratio explicitly in results section, and represented in figures 1 and 2.
Agree (save as-is)CorrectIncorrect

outcome_of_interest

Alcohol related hospitalisation

Reasoning

Outcome defined by authors in methods section on page 2, confirmed in results section and table 1.
Agree (save as-is)CorrectIncorrect

Who I am

I'm Jackson, I am a medical doctor halfway through Australian physicians training currently completing a PhD in population health at the University of Oxford. I built this tool as I was doing my own systematic review because I thought it should exist, but just couldn't find it available anywhere on the internet.

My main aim for this tool is to assist in reproducible, reliable data extraction from academic papers. The tool uses both traditional and AI methods to extract data from PDF documents, including tables and figures, which are the passed through a harnsed LLM to generate results. It has helped speed up my own and my colleagues research, without compromising on quality.

I genuinely want this tool to produce as high quality results as possible. Whilst I have iterated and tried to consider edge cases, there are almost certainly failure modes or design flaws I have missed. As such, if you choose to use this tool and would either be willing to share your final checked data extraction table, or meaningful feedback on your experience with the tool, I will give you a 50% refund on the initial payment.

I see this as part of a broader project in reducing residual time spent in research on repetitive and automatable tasks. I ramble on my blog here if you'd like to read more. Thanks for supporting the project.

I offer a no questions asked full refund if for any reason the tool does not meet expectations.

Built for systematic review workflows

Consistent

Apply the same extraction framework across every paper in your review, with the same question definitions and answer formats throughout.

Flexible

Define questions relevant to your specific review: text, numbers, categories from your own list, or lists of items extracted per paper.

Reviewable

Inspect and amend every extracted value before analysis. Reasoning notes help you locate answers in the source paper.

Practical

Receive a structured Excel workbook that fits existing evidence-synthesis workflows, with an export path for reviewed results and correctness labels.

Choose the size of your extraction

Select the package that best matches your review. After payment, you will be directed to submit your papers and extraction framework.

Pilot

Up to 10 papers

£25

For small reviews, pilot projects or testing the workflow.

  • Up to 10 academic PDFs
  • Up to 20 researcher-defined extraction questions per paper
  • Structured Excel workbook with reasoning notes
  • Access to cross-checking application for review
Choose Pilot

Standard

Up to 30 papers

£50

For typical systematic review extraction projects.

  • Up to 30 academic PDFs
  • Up to 20 paper-level extraction questions per paper, with optional variable-level rows
  • Structured Excel workbook with reasoning notes
  • Access to cross-checking application for review
Choose Standard

Large

Up to 60 papers

£75

For larger systematic reviews and evidence-synthesis projects.

  • Up to 60 academic PDFs
  • Up to 20 paper-level extraction questions per paper, with optional variable-level rows
  • Structured Excel workbook with reasoning notes
  • Access to cross-checking application for review
Choose Large

Larger or more complex review?

Contact us to discuss projects that fall outside the standard packages.

Contact Us

What happens next?

  1. 1

    Pay securely

    Complete checkout through Stripe. You will receive a confirmation with next steps.

  2. 2

    Submit your project

    After payment, you are directed to a submission page. Upload your PDF papers and your completed extraction framework workbook. If you have not filled in the template yet, download it here.

    Download the template workbook (.xlsx)
  3. 3

    Receive and review

    We process your papers and deliver a structured Excel workbook. Use the cross-checking application to review, edit, and export your verified results.

Frequently asked questions

What do I provide?
Two things: your academic papers as PDF files, and one Excel workbook (.xlsx) containing your extraction framework. The workbook has a paper_level_data sheet for questions asked once per paper and an optional variable_level_data sheet for questions asked per item. A template is available in the How it works section.
Can I define my own extraction questions?
Yes. Every extraction question is defined by you. Each row of the workbook is one question, with a short name, the question text, an answer format (text, number, category from a list, or semicolon-separated list), and optional allowed categories. The same set of questions is applied to every paper in your project.
Can information be extracted from tables and figures?
Yes, where the content is visible in the paper. Tables and figures are processed alongside the document text. Extraction is limited to what is explicitly stated in the paper — the system does not infer values that are not reported.
What do I receive?
You receive a single Excel workbook (.xlsx). Each extraction question appears as a column, paired with a reasoning column that briefly explains where the answer was found. By default you get one row per paper. If you use variable-level questions, you get one row per item with a variable_name column and paper-level answers repeated on each row.
Can one paper contain multiple things I need extracted?
Yes. Add a variable_abstraction question on the paper sheet to list the items (for example, separate prediction models or target trials), then write variable-level questions using {variable_name}. The workbook expands to one row per item. If you only need study-wide answers, leave variable_level_data empty and you get one row per paper.
How do I cross-check results?
After delivery, you use the cross-checking application in your browser. Load the Excel workbook, review one paper at a time, and optionally open the matching source PDF side by side. Each extracted field shows the value and a read-only reasoning note. You can edit values and mark each field as correct or incorrect.
Can I edit extracted results?
Yes. In the cross-checking application you can edit any extracted value directly. You can also add new data points, add or remove variable rows, and copy values between rows. Reasoning columns are read-only reference material from the extraction.
Can I save or download reviewed results?
Yes. The cross-checking application autosaves your progress in the browser. You can also download a combined Excel workbook with two sheets: Data (your edited values) and Labels (correct/incorrect marks for each field). Session files can be downloaded and reloaded for backup.
Should I check the results?
Yes. AI-assisted extraction should always be verified before it informs analysis, review conclusions, or publication. The cross-checking application exists for exactly this step, and reasoning notes point you to where each answer was found in the source paper.
How long will it take?
Typical turnaround is 2-3 days depending on the number of papers and complexity of your extraction framework. We will confirm an estimated delivery date when you submit your project.
What is your refund policy?
If the extraction is not accurate enough, does not meaningfully speed up your work, or is otherwise not what you needed, we will refund your payment in full. This applies after you have received your results, so you can judge them before deciding. We also offer a 50% refund if you share detailed, constructive feedback on your experience: what went wrong, what was missing, or how the service could be better. This feedback directly shapes the product. To request a refund, email jackson@systematicallyreviewed.com. Refunds are processed to the original payment method.
What happens to my papers?
Uploaded materials are handled according to our privacy and data-retention policies. The cross-checking application runs entirely in your browser — your reviewed data does not leave your machine during verification.

Ready to test it on your review?

Choose a package and turn your extraction framework into structured, reviewable results.

View Pricing