Friday, September 11, 2026
Science
No Result
View All Result
  • Login
  • HOME
  • SCIENCE NEWS
  • CONTACT US
  • HOME
  • SCIENCE NEWS
  • CONTACT US
No Result
View All Result
Scienmag
No Result
View All Result
Home Science News Technology and Engineering

Ensemble deep learning model detects ChatGPT-generated text accurately

September 11, 2026
in Technology and Engineering
Blake Davidson
By Blake Davidson Scienmag Editorial Profile - Data Science
Reading Time: 5 mins read
0
Ensemble deep learning model detects ChatGPT-generated text accurately

Ensemble deep learning model detects ChatGPT-generated text accurately

65
SHARES
587
VIEWS
Share on FacebookShare on Twitter
ADVERTISEMENT

In a world where ChatGPT and its relatives are writing essays, news articles, emails and even academic papers, one question has become urgent for educators, journalists and policymakers alike: can we tell machine-written text from human-written text? A new study offers a powerful answer. Jawaher Alghamdi, Yuqing Lin and Suhuai Luo, researchers at the University of Newcastle in Australia, King Khalid University in Saudi Arabia and Jimei University in China, have built a deep learning framework that detects ChatGPT-generated text with up to 97 percent accuracy on long passages and 88 percent on short ones. The work, published in Multimedia Tools and Applications, is a significant step forward in the rapidly growing field of machine-generated text detection.

The rise of large language models has been extraordinary. ChatGPT, based on OpenAI’s GPT-3.5 architecture, has attracted enormous attention across education, business and entertainment for its ability to produce fluent, extended responses to human queries. But that fluency carries a risk. The authors of the study point out that the same capabilities can be misused for malicious purposes, including spreading disinformation, generating fraudulent academic work and producing deceptive content at scale. As the volume of synthetic text grows, the responsible and ethical use of ChatGPT has become paramount, and tools to distinguish original human-generated text from machine output are increasingly needed.

The new framework is built on an ensemble, a strategy in which several models are combined so that their individual strengths compensate for one another’s weaknesses. At its core are four transformer-based language models, each pre-trained on massive text corpora and each with a distinctive view of language. The first is BERT, Bidirectional Encoder Representations from Transformers, which reads text in both directions and builds deeply contextualized word representations. The second is RoBERTa, a robustly optimized variant of BERT that refines the pre-training recipe for stronger performance. The third is XLNet, a generalized autoregressive model that overcomes some limitations of standard masked language models by considering all permutations of the input sequence. The fourth is GPT-2, a generative pre-trained transformer that predicts the next token from left to right. Because these models differ in how they process context, the researchers reasoned that their combined judgment would be more reliable than any single model’s.

But the architecture does not stop there. The team recognized that while transformers excel at capturing rich semantic features, they are less sensitive to the fine-grained, local patterns that often distinguish machine text, such as characteristic word sequences and rhythm of expression. To address this, the framework passes the transformer outputs through two complementary components. A one-dimensional Convolutional Neural Network, or 1D-CNN, extracts high-level local features by sliding filters across the embedded text and picking up distinctive n-gram-like patterns. A Bidirectional Gated Recurrent Unit, or BiGRU, then considers the temporal flow of information in both directions, capturing how meaning unfolds across a passage from beginning to end and from end back to the beginning. Together, these components allow the model to process contextualized information and reveal meaningful patterns that enhance detection.

The choice to test the system on both short and long texts is one of the study’s most practical contributions. Detecting machine-generated text in short passages, such as social media posts, is notoriously difficult because there is little statistical signal to work with. Longer documents give classifiers far more evidence to examine. The researchers evaluated their model on datasets containing both human-written and ChatGPT-generated samples. Importantly, they controlled for a subtle confound: because ChatGPT’s output can be influenced by the conversation history, the team refreshed the chat thread for every text sample, ensuring that each generated passage was independent and not contaminated by prior exchanges.

The results were striking. On long texts, the ensemble achieved an accuracy rate of 97 percent, correctly identifying ChatGPT-generated text from human text in the vast majority of cases. On short texts, where the signal is weaker, the model still reached 88 percent accuracy, a level that compares favorably with earlier methods. Perhaps more importantly, the experiments demonstrated a clear finding: the transformer ensemble consistently outperformed individual transformer models when used alone. Combining BERT, RoBERTa, XLNet and GPT-2 into a single decision framework proved more robust than relying on any one of them, validating the ensemble approach for this task.

The study builds on a growing body of research into machine-generated text detection. Earlier efforts, such as work on detecting neural fake news, exposed the vulnerability of current classifiers to synthetic text and proposed defensive models. Other researchers have explored detection of short ChatGPT texts using explainable machine learning, and the field has since expanded with dedicated shared tasks on English and multilingual machine-generated content detection. The new work draws on that lineage but adds a distinctive hybrid architecture. The research team had previously developed BERT-CNN-BiLSTM models for detecting fake news on social media, and their experience with that problem clearly informed the design here. Both fake news detection and ChatGPT detection share a common structure: distinguishing content produced with deceptive intent from authentic content, a challenge the researchers note is related to established problems such as hate speech detection.

The implications of this research extend well beyond the laboratory. In education, institutions around the world are grappling with students submitting AI-generated essays, and detection tools could help preserve academic integrity. In journalism and information security, the ability to flag machine-generated disinformation could slow the spread of synthetic propaganda. The authors explicitly frame their results as a useful approach for policymakers and researchers concerned with detecting and preventing the malicious use of ChatGPT. By publishing their methodology, including comparisons against baseline models, the study gives developers a reference point for building practical detection systems.

The technical details also matter for the broader AI community. The study underscores why ensembles have become a recurring theme in modern natural language processing. Prior work, such as ensemble models combining BERT and RoBERTa for classifying idioms versus literal texts, has shown that different pre-trained models encode complementary information. This new study extends that insight to the AI-detection problem, showing that a model that effectively processes contextualized information from multiple transformer perspectives, then enriches it with convolutional and recurrent feature extraction, can reveal patterns in text that individual models miss. The generative-versus-discriminative distinction explored in classic machine learning research also echoes here, as GPT-2 and its generative cousins see language differently from encoder models like BERT and XLNet.

The road to publication was long. The paper was received in March 2024, revised in November 2025, and accepted in December 2025, reflecting the fast-moving nature of this field where the technology being studied evolves faster than the review cycle. The fact that the model achieves such strong performance on a text generated by GPT-3.5-era ChatGPT suggests the underlying patterns of machine fluency are stable enough to be learned reliably, even as chatbots grow more capable.

Of course, no detector is a permanent solution. As generative models improve, they may produce text that evades current detection methods, and the arms race between generation and detection is likely to continue. The researchers acknowledge the need for continued investigation into ways to distinguish human from machine text. But their results demonstrate that the problem is tractable: with the right combination of pre-trained transformers and sequence-aware deep learning, machine-written text leaves a detectable fingerprint. For now, this ensemble framework offers one of the most accurate available tools for finding that fingerprint, whether hidden in a long essay or a short post. The study was not funded by any organization, and the authors declare no conflicts of interest.

Subject of Research: Detection of ChatGPT-generated text versus human-generated text using a deep learning ensemble model

Subject of Research: Technology and Engineering

Article Title: Detecting ChatGPT-generated text: A deep learning ensemble for accurate differentiation from human-generated text

Article References: Alghamdi, J., Lin, Y., & Luo, S. (2026). Detecting ChatGPT-generated text: A deep learning ensemble for accurate differentiation from human-generated text. Multimedia Tools and Applications, 85(8), Article 686. https://doi.org/10.1007/s11042-026-21175-z

Image Credits: AI Generated

DOI: 10.1007/s11042-026-21175-z

Keywords: ChatGPT, text classification, BERT, RoBERTa, XLNet, GPT, deep learning, ensemble model, 1D-CNN, BiGRU, machine-generated text detection, natural language processing

Cite Scienmag News

Blake Davidson. (September 11, 2026). Ensemble deep learning model detects ChatGPT-generated text accurately. Scienmag. https://scienmag.com/ensemble-deep-learning-model-detects-chatgpt-generated-text-accurately/

Blake Davidson. "Ensemble deep learning model detects ChatGPT-generated text accurately." Scienmag, 11 September 2026, https://scienmag.com/ensemble-deep-learning-model-detects-chatgpt-generated-text-accurately/. Accessed 11 September 2026.

Blake Davidson. "Ensemble deep learning model detects ChatGPT-generated text accurately." Scienmag. September 11, 2026. https://scienmag.com/ensemble-deep-learning-model-detects-chatgpt-generated-text-accurately/

Tags: advancements in natural language processing for text verificationAI-generated content in education and journalismapplications of AI detection in academic integrityapplications of AI in journalism and educationchallenges in distinguishing human vs AI writingchallenges of identifying AI-generated academic papersChatGPT architecture and capabilitiesdeep learning frameworks for text analysisdeep learning frameworks for text classificationdisinformation and fraudulent content detectionEnsemble deep learning for ChatGPT-generated text detectionEnsemble deep learning models for detecting ChatGPT-generated textethical considerations of synthetic textethical implications of synthetic textimpact of AI-generated content on societyimpact of AI-generated text on academia and medialarge language models and AI-generated contentlarge language models and disinformationlong-passages vs short-passages in AI detectionmachine-generated text identification accuracymachine-written text detection accuracymultilingual AI text detectionmultilingual AI text detection research
Share26Tweet16
Previous Post

How a Single Chemokine Can Sabotage Radiotherapy and Shape the Immune Battlefield

Next Post

Medicinal Plants Offer Antibiotic Alternatives for Poultry Under One-Health Framework

Related Posts

Layered security approach protects ECG data in constrained IoT healthcare
Technology and Engineering

Layered security approach protects ECG data in constrained IoT healthcare

September 11, 2026
Graph Learning Method Detects Suspicious Citation Groups in Academic Networks
Technology and Engineering

Graph Learning Method Detects Suspicious Citation Groups in Academic Networks

September 10, 2026
Fractional tangent search-enhanced SqueezeNet monitors post-COVID heart health via federated learning
Technology and Engineering

Fractional tangent search-enhanced SqueezeNet monitors post-COVID heart health via federated learning

September 10, 2026
3D Vision System Teaches Robots to Repair Turbine Blades Without Human Programming
Technology and Engineering

3D Vision System Teaches Robots to Repair Turbine Blades Without Human Programming

September 10, 2026
Enhanced Genetic Algorithm Boosts Lifetime and Coverage in Underwater Sensor Networks
Technology and Engineering

Enhanced Genetic Algorithm Boosts Lifetime and Coverage in Underwater Sensor Networks

September 10, 2026
Multi-task framework fuses infrared and visible images for better semantics
Technology and Engineering

Multi-task framework fuses infrared and visible images for better semantics

September 10, 2026
Next Post
Medicinal Plants Offer Antibiotic Alternatives for Poultry Under One-Health Framework

Medicinal Plants Offer Antibiotic Alternatives for Poultry Under One-Health Framework

  • Mothers who receive childcare support from maternal grandparents show more optimized

    Mothers who receive childcare support from maternal grandparents show more parental warmth, finds NTU Singapore study

    27656 shares
    Share 11059 Tweet 6912
  • University of Seville Breaks 120-Year-Old Mystery, Revises a Key Einstein Concept

    1061 shares
    Share 424 Tweet 265
  • Bee body mass, pathogens and local climate influence heat tolerance

    682 shares
    Share 273 Tweet 171
  • Researchers record first-ever images and data of a shark experiencing a boat strike

    546 shares
    Share 218 Tweet 137
  • Groundbreaking Clinical Trial Reveals Lubiprostone Enhances Kidney Function

    531 shares
    Share 212 Tweet 133
Science

Embark on a thrilling journey of discovery with Scienmag.com—your ultimate source for cutting-edge breakthroughs. Immerse yourself in a world where curiosity knows no limits and tomorrow’s possibilities become today’s reality!

RECENT NEWS

  • Medicinal Plants Offer Antibiotic Alternatives for Poultry Under One-Health Framework
  • Ensemble deep learning model detects ChatGPT-generated text accurately
  • How a Single Chemokine Can Sabotage Radiotherapy and Shape the Immune Battlefield
  • Group Occupational Therapy Intervention Helps People With Mental Illness Rebuild Recovery

Categories

  • Agriculture
  • Anthropology
  • Archaeology
  • Athmospheric
  • Biology
  • Biotechnology
  • Blog
  • Bussines
  • Cancer
  • Chemistry
  • Climate
  • Earth Science
  • Editorial Policy
  • Marine
  • Mathematics
  • Medicine
  • Pediatry
  • Policy
  • Psychology & Psychiatry
  • Science Education
  • Social Science
  • Space
  • Technology and Engineering

Subscribe to Blog via Email

Enter your email address to subscribe to this blog and receive notifications of new posts by email.

Join 5,151 other subscribers

© 2025 Scienmag - Science Magazine

Welcome Back!

Login to your account below

Forgotten Password?

Retrieve your password

Please enter your username or email address to reset your password.

Log In
No Result
View All Result
  • HOME
  • SCIENCE NEWS
  • CONTACT US

© 2025 Scienmag - Science Magazine

Discover more from Science

Subscribe now to keep reading and get access to the full archive.

Continue reading