Sunday, October 11, 2026
Science
No Result
View All Result
  • Login
  • HOME
  • SCIENCE NEWS
  • CONTACT US
  • HOME
  • SCIENCE NEWS
  • CONTACT US
No Result
View All Result
Scienmag
No Result
View All Result
Home Science News Technology and Engineering

New Open-Source R Workflow Aims to Make Systematic Literature Reviews Reproducible

October 11, 2026
in Technology and Engineering
Denise Maddox
By Denise Maddox Scienmag Editorial Profile - Mechanical Engineering
Reading Time: 5 mins read
0
New Open-Source R Workflow Aims to Make Systematic Literature Reviews Reproducible

New Open-Source R Workflow Aims to Make Systematic Literature Reviews Reproducible

65
SHARES
587
VIEWS
Share on FacebookShare on Twitter
ADVERTISEMENT

Systematic literature reviews are among the most labor-intensive exercises in modern science. Researchers must collect thousands of bibliographic records, deduplicate them, screen them against eligibility criteria, run bibliometric analyses, and then somehow convert citation networks and keyword maps into theory-oriented evidence. In practice, this usually means juggling a patchwork of separate tools, each with its own data formats and settings, and the methodological trail between them often breaks down. A newly published open-source project called SLR Suite, described in the journal SoftwareX by developer Metin Akbulut, proposes a different approach: a single, modular R workflow that carries a review from raw bibliographic data all the way to structured, theory-oriented evidence tables, with every step logged, versioned, and repeatable.

The core problem the software targets is reproducibility. When researchers rely on separate applications for data collection, screening, bibliometric analysis, and visualization, transferring data and methodological decisions between them makes analytical provenance and configuration management difficult to maintain. A second, subtler gap concerns the divide between quantitative bibliometrics and theory-oriented synthesis. Frameworks such as Theory–Context–Characteristics–Methodology (TCCM) and Structure–Process–Analysis–Research Agenda (SPAR) offer structured ways to organize evidence, but applying them has typically been a manual exercise conducted far from the computational pipeline. SLR Suite attempts to bridge that divide by creating a traceable connection between selected bibliometric outputs and configurable, inspectable TCCM evidence for researcher-led synthesis.

Architecturally, the suite is organized as a pipeline of nine ordered modules: bibliographic acquisition and deduplication, rule-based screening support, bibliometric analysis, VOSviewer export, dictionary-based TCCM matrix construction, thematic evolution analysis, citation-impact assessment, SPAR-oriented evidence aggregation, and generation of PRISMA flow-count artifacts. A shared launcher orchestrates execution but contains no analytical logic itself; each module reads its own declared configuration and input files and writes standardized outputs to designated data layers. Modules exchange data through documented file contracts rather than direct function calls, a design choice that reduces coupling and allows any individual module to be rerun independently when its upstream inputs are available.

The data flow is organized into three structured layers—raw, interim, and processed—so that every transformation is inspectable and intermediate datasets are preserved for auditing or debugging. Configuration lives in version-controlled YAML files: screening criteria in one file, and researcher-defined TCCM dictionaries of labels and case-insensitive regular-expression patterns in another. Because the classification procedure is rule-based rather than inferential, researchers can inspect exactly which patterns produced each label and revise the dictionaries for their particular domain. The direct input format is a Web of Science plain-text export; Scopus CSV records are handled upstream by a companion tool called BibexPy, which merges and harmonizes mixed-source data before SLR Suite consumes the combined output.

Reproducibility controls run through the entire project. The suite uses renv to lock all 130 package dependencies, ships with automated tests and quality gates, and runs continuous integration on Ubuntu, macOS, and Windows to verify installation and execution on bundled data. The current release, version 2.3.2, is permanently archived on Zenodo with a specific Git commit identifier, and the developers report that the analytical source matches the benchmark source exactly. Seven successive public releases have progressively added modular project structure, dependency locking, end-to-end tests, software–documentation traceability, and static Sankey diagram export, documenting steady improvements in maintainability and technical reproducibility.

The TCCM coding module illustrates the software’s philosophy in detail. It combines titles, abstracts, author keywords, and Keywords Plus fields, normalizes the text, and tests every record against every configured pattern. Multiple matched labels are retained as semicolon-separated values, while records with no match are reported as missing rather than assigned an inferred label. The method is fully deterministic: identical text and dictionary versions always produce identical matches, and no probability or confidence scores are generated. Dictionary additions are expected to come with domain justification, version control, and regression checks against known examples, with researchers reviewing the results for false positives, false negatives, and ambiguous assignments.

To verify that the machinery actually works, the developer ran the full nine-module pipeline on an illustrative corpus of 754 publications from 2015 to 2026 covering artificial intelligence, tourism, and customer experience—a dataset spanning 390 sources, 2,519 unique authors, and more than 2,700 author keywords. A separate technical benchmark then executed the pipeline five times on each of three collections containing 20, 208, and 500 records. All 15 runs completed successfully, with all 135 module executions logged as successful and required outputs, including thematic Sankey visualizations, verified. Median execution time rose from about 19 seconds for the smallest collection to roughly 53 seconds for the largest, while approximate peak memory increased from about 474 to 544 MiB on a Windows 11 test machine.

The developers are notably careful about what these results do and do not demonstrate. Software verification and scientific validation are treated as strictly separate concerns: automated tests confirm that the code installs and executes its declared operations, but they do not establish that a literature search is complete, that eligibility decisions are correct, or that the generated classifications are scientifically valid. A blinded dual-AI pilot on 60 records, used only to dry-run the coding protocol, showed perfect agreement between the two AI coders and 96.67 percent agreement with the software, but the authors explicitly decline to present this as evidence of classification accuracy, since AI agents are not independent human researchers and no independently human-coded reference set exists. Notably, an earlier benchmark of version 2.3.1 exposed a genuine bug—duplicate year boundaries caused the thematic-evolution calculation to fail silently while logs indicated success—which motivated corrections to interval construction and error propagation in the current release.

How does SLR Suite compare with the crowded ecosystem of existing review tools? The accompanying comparison table spans bibliometrix, VOSviewer, CiteSpace, Rayyan, ASReview, DistillerSR, EPPI-Reviewer, Parsifal, Colandr, Nested Knowledge, SWARM-SLR, BibexPy, Covidence, ARC, and INRA SLR Copilot, each excelling at particular stages such as screening, science mapping, or AI-assisted triage. What distinguishes SLR Suite, according to the assessment, is its combination of a fully reproducible versioned workflow with documented TCCM support and partial SPAR support—capabilities none of the compared tools document. The authors are careful to note that the comparison does not establish absolute uniqueness or superior performance, and that controlled comparative evaluation would be required for such claims.

The project’s limitations are stated with unusual candor. It does not formulate review protocols, execute database searches, replace independent eligibility judgments, assess risk of bias, or validate scientific interpretations. In domains without an established theory base, the workflow cannot discover or infer new theories; researchers must first build a provisional vocabulary from domain literature and refine it iteratively. Future evaluation priorities include two-researcher manual coding with agreement analysis, usability testing with real participants, controlled comparisons with alternative workflows, and cross-domain replication in health sciences, social sciences, and engineering. Positioned as a transparent support environment rather than a replacement for researcher judgment, SLR Suite represents a growing movement in metascience: treating the review process itself as an auditable, version-controlled computational pipeline whose every decision can be inspected, rerun, and improved.

Subject of Research: A reproducible open-source R software workflow for conducting systematic literature reviews

Article Title: SLR suite: A reproducible R workflow for systematic literature reviews

Article References: Akbulut, M. (2026). SLR suite: A reproducible R workflow for systematic literature reviews. SoftwareX, 36, Article 103101. https://doi.org/10.1016/j.softx.2026.103101

Image Credits: AI Generated

DOI: 10.1016/j.softx.2026.103101

Keywords: systematic literature review, SLR Suite, R programming, reproducibility, bibliometrics, TCCM framework, SPAR, PRISMA, open-source software, text classification, research synthesis, SoftwareX

Cite Scienmag News

Denise Maddox. (October 11, 2026). New Open-Source R Workflow Aims to Make Systematic Literature Reviews Reproducible. Scienmag. https://scienmag.com/new-open-source-r-workflow-aims-to-make-systematic-literature-reviews-reproducible/

Denise Maddox. "New Open-Source R Workflow Aims to Make Systematic Literature Reviews Reproducible." Scienmag, 11 October 2026, https://scienmag.com/new-open-source-r-workflow-aims-to-make-systematic-literature-reviews-reproducible/. Accessed 11 October 2026.

Denise Maddox. "New Open-Source R Workflow Aims to Make Systematic Literature Reviews Reproducible." Scienmag. October 11, 2026. https://scienmag.com/new-open-source-r-workflow-aims-to-make-systematic-literature-reviews-reproducible/

Tags: automation of evidence extractionbibliographic data management and deduplicationbibliographic data visualization and mappingbibliometricsbridging quantitative bibliometrics and qualitative synthesisintegrated bibliometric analysis toolsmethodological transparency in literature reviewsmodular software for literature screeningopen-source R workflow for literature reviewsopen-source softwarePRISMAR programmingreproducibilityreproducible research in systematic reviewsresearch synthesisSLR SuiteSoftwareXSPARsystematic literature reviewsystematic literature review reproducibilityTCCM frameworktext classificationtheory-oriented evidence synthesisversion control in research workflows
Share26Tweet16
Previous Post

Statistical Model Pinpoints How Big Lucknow’s Next Great Gomati Flood Could Get

Next Post

School Closures Did Not Curb Teen Aggression, Study Finds

Related Posts

Transient Dynamics Reveal Hidden Weaknesses in Standard Antibiotic Testing
Biology

Transient Dynamics Reveal Hidden Weaknesses in Standard Antibiotic Testing

October 11, 2026
Pineapple Peels Turned Into Glowing Nanoprobes That Track a Common Insecticide in Water
Technology and Engineering

Pineapple Peels Turned Into Glowing Nanoprobes That Track a Common Insecticide in Water

October 11, 2026
Carbon-Coated MoS2 Delivers Record Supercapacitor Power and Rapid Dye Breakdown
Technology and Engineering

Carbon-Coated MoS2 Delivers Record Supercapacitor Power and Rapid Dye Breakdown

October 11, 2026
Simple Regression Models Bring Predictive Socket Design Closer for Transradial Prostheses
Medicine

Simple Regression Models Bring Predictive Socket Design Closer for Transradial Prostheses

October 11, 2026
Motion Style Slider Gives Animators Precise Control Over AI Character Movement
Technology and Engineering

Motion Style Slider Gives Animators Precise Control Over AI Character Movement

October 11, 2026
New Defense System Shields Federated Learning From Hidden Backdoor Attacks
Technology and Engineering

New Defense System Shields Federated Learning From Hidden Backdoor Attacks

October 11, 2026
Next Post
School Closures Did Not Curb Teen Aggression, Study Finds

School Closures Did Not Curb Teen Aggression, Study Finds

  • Mothers who receive childcare support from maternal grandparents show more optimized

    Mothers who receive childcare support from maternal grandparents show more parental warmth, finds NTU Singapore study

    27656 shares
    Share 11059 Tweet 6912
  • University of Seville Breaks 120-Year-Old Mystery, Revises a Key Einstein Concept

    1061 shares
    Share 424 Tweet 265
  • Bee body mass, pathogens and local climate influence heat tolerance

    682 shares
    Share 273 Tweet 171
  • Researchers record first-ever images and data of a shark experiencing a boat strike

    546 shares
    Share 218 Tweet 137
  • Groundbreaking Clinical Trial Reveals Lubiprostone Enhances Kidney Function

    531 shares
    Share 212 Tweet 133
Science

Embark on a thrilling journey of discovery with Scienmag.com—your ultimate source for cutting-edge breakthroughs. Immerse yourself in a world where curiosity knows no limits and tomorrow’s possibilities become today’s reality!

RECENT NEWS

  • School Closures Did Not Curb Teen Aggression, Study Finds
  • New Open-Source R Workflow Aims to Make Systematic Literature Reviews Reproducible
  • Statistical Model Pinpoints How Big Lucknow’s Next Great Gomati Flood Could Get
  • MOTS-c: The Tiny Mitochondrial Peptide That Behaves Differently in Health and Cancer

Categories

  • Agriculture
  • Anthropology
  • Archaeology
  • Athmospheric
  • Biology
  • Biotechnology
  • Blog
  • Bussines
  • Cancer
  • Chemistry
  • Climate
  • Earth Science
  • Editorial Policy
  • Marine
  • Mathematics
  • Medicine
  • Pediatry
  • Policy
  • Psychology & Psychiatry
  • Science Education
  • Science News
  • Social Science
  • Space
  • Technology and Engineering

Subscribe to Blog via Email

Enter your email address to subscribe to this blog and receive notifications of new posts by email.

Join 5,150 other subscribers

© 2025 Scienmag - Science Magazine

Welcome Back!

Login to your account below

Forgotten Password?

Retrieve your password

Please enter your username or email address to reset your password.

Log In
No Result
View All Result
  • HOME
  • SCIENCE NEWS
  • CONTACT US

© 2025 Scienmag - Science Magazine

Discover more from Science

Subscribe now to keep reading and get access to the full archive.

Continue reading