Friday, September 11, 2026
Science
No Result
View All Result
  • Login
  • HOME
  • SCIENCE NEWS
  • CONTACT US
  • HOME
  • SCIENCE NEWS
  • CONTACT US
No Result
View All Result
Scienmag
No Result
View All Result
Home Science News Earth Science

Machine learning improves classification of oceanic basalt in research

September 11, 2026
in Earth Science
Blake Davidson
By Blake Davidson Scienmag Editorial Profile - Data Science
Reading Time: 6 mins read
0
Machine learning improves classification of oceanic basalt in research

Machine learning improves classification of oceanic basalt in research

65
SHARES
587
VIEWS
Share on FacebookShare on Twitter
ADVERTISEMENT

For decades, geologists have relied on a handful of carefully drawn diagrams to decipher the origins of the ocean floor’s volcanic rocks. Now, a team of researchers in China has shown that machine learning can outperform these traditional tools by a wide margin, achieving classification accuracies of up to 97 percent when sorting oceanic basalts into the tectonic settings where they formed. The study, led by Haobin Xu and Juanjuan Kong of Shandong University of Science and Technology together with Yao Ma of Hebei Normal University of Science and Technology, appears in the journal Earth Science Informatics and offers a data-driven alternative to one of petrology’s most stubborn challenges.

Basalts are the most abundant rocks on the ocean floor, and their chemical fingerprints carry the memory of the mantle sources and geological processes that produced them. Mid-ocean ridge basalts, known as MORB, rise from the depleted upper mantle at spreading centers. Ocean island basalts, or OIB, tap enriched mantle plumes beneath volcanic islands such as Hawaii. Island arc basalts, abbreviated IAB, form above subduction zones, where descending oceanic slabs release fluids that modify the mantle wedge and introduce components from the crust. Telling these three families apart is fundamental to reconstructing plate tectonic histories, identifying ancient oceanic crust preserved in mountain belts, and understanding how the mantle has evolved over billions of years.

The classical approach uses discriminant diagrams, which plot ratios of two or three trace or major elements against one another and delineate fields corresponding to different tectonic settings. Geologists plot an unknown sample on the diagram and read off its likely origin from the field in which it falls. While elegant and easy to use, these diagrams have well-documented shortcomings. Their boundaries are drawn by eye around limited sample sets, the fields often overlap indistinctly, and a substantial fraction of real samples fall into ambiguous zones or outside the diagram’s coverage altogether. Previous large-scale audits of discrimination diagrams have found that misclassification rates can be uncomfortably high, and the new study quantifies that gap directly.

To build a more rigorous benchmark, the team turned to two of the largest open-access geochemical repositories in the Earth sciences: the GEOROC database and PetDB, maintained by EarthChem. From these archives they compiled 950 samples of MORB, OIB, and IAB, each characterized by a suite of major and trace element concentrations. This curated dataset formed the raw material for both a statistical exploration of the data and the training of machine learning classifiers.

The first analytical step was principal component analysis, a technique that compresses many correlated chemical variables into a small number of composite axes that capture the greatest variance in the data. The analysis revealed that the first principal component, PC1, is dominated by so-called enriched elements, most notably niobium, a trace element that behaves incompatibly during mantle melting and accumulates in melts derived from enriched mantle sources. The second principal component, PC2, corresponds mainly to silica dioxide and aluminum oxide, the major oxides that track the degree of melting and the mineralogical character of the source. Together, PC1 and PC2 explain more than 60 percent of the total chemical variance in the dataset, and, crucially, plots of samples in this two-dimensional space effectively separate the three basalt types. In other words, the essential information needed to distinguish MORB from OIB and IAB is genuinely present in the chemistry, and it can be extracted without hand-drawn boundaries.

With the feature structure established, the researchers trained three main classification models drawn from the standard machine learning toolbox: support vector machines, random forests, and k-nearest neighbors. They also included XGBoost, a gradient-boosting algorithm, as a supplementary benchmark. Each method approaches the problem differently. Support vector machines construct decision boundaries in a high-dimensional feature space, maximizing the margin between classes. K-nearest neighbors classifies a sample by asking which category dominates among its most chemically similar neighbors in the training set. Random forests, first formalized by Leo Breiman in 2001, build hundreds of decision trees, each trained on a random subset of the data and features, and then pool their votes. XGBoost instead builds trees sequentially, with each new tree trained to correct the errors of its predecessors.

Model evaluation followed best practices in machine learning. The team used five-fold cross-validation, in which the data are repeatedly split so that models are tested on samples they never saw during training, and they reported performance using confusion matrices and receiver operating characteristic curves, tools that capture not just overall accuracy but also how well each class is distinguished and how the trade-off between sensitivity and specificity behaves. Final performance figures were calculated on an independent test set, providing an unbiased estimate of real-world accuracy.

The results were striking. Random forests came out on top with an overall accuracy of approximately 0.97 on the independent test set. Support vector machines followed at about 0.95, the XGBoost benchmark reached roughly 0.86, and k-nearest neighbors trailed at about 0.83. For comparison, the traditional discrimination diagram managed only around 0.73 accuracy on the same task. The random forest’s advantage is likely rooted in its ensemble architecture: by averaging many decorrelated trees, it suppresses the noise from individual weak decision paths and captures nonlinear interactions among elements that no two- or three-axis diagram can represent.

Beyond raw accuracy, the study extracted geological insight from the models themselves. Feature-importance analysis of the random forest showed that both major-element differentiation and trace-element enrichment contribute jointly to the classification. This aligns with geochemical theory: major elements like silica reflect the degree and depth of partial melting, while trace elements such as niobium, thorium, and the rare earth elements record the nature of the mantle source and the imprint of subduction fluids. Machine learning, in effect, rediscovered the dual importance of source and process, but did so quantitatively and without human guidance on which elements should matter.

Perhaps the most provocative finding concerns the samples the models were least sure about. The researchers performed a probabilistic confidence analysis, examining cases where the classifier’s output probabilities were low or split between categories. Rather than representing random noise or data-entry errors, these low-confidence samples correspond to geochemically transitional or hybrid compositions, rocks whose chemistry sits between the canonical MORB, OIB, and IAB signatures. The authors interpret this as evidence that model uncertainty captures genuine tectonic information: ambiguous samples may record mixing of mantle sources, complex tectono-magmatic evolution, or settings where multiple processes overlap. In this reading, a machine learning model’s hesitation becomes a scientific signal, flagging samples that deserve closer scrutiny rather than discarding them as errors.

The geological coherence of the results reinforces this interpretation. The classification boundaries the algorithms learned map cleanly onto established mantle geochemistry. MORB samples cluster with depleted mantle signatures, reflecting the well-worn upper mantle beneath spreading ridges. OIB samples carry the enriched mantle fingerprints of deep plumes, with elevated incompatible element concentrations. IAB samples display the characteristic modifications wrought by subduction, including the involvement of crustal components and fluid-mobile element enrichment. The geology and the statistics tell the same story from independent directions.

The implications reach well beyond the ocean basins. Oceanic basalts are famously recycled into ophiolites, slices of ancient seafloor thrust onto continents, and distinguishing their original tectonic settings is essential for reconstructing supercontinent cycles and ancient plate configurations. Machine learning classifiers trained on modern samples can now be applied to these ancient rocks with far greater confidence than traditional diagrams allow. The approach also demonstrates a broader principle: by pairing the enormous public geochemical databases accumulated over decades with interpretable machine learning workflows, Earth scientists can turn archived measurements into quantitative tools that outperform the heuristic methods of the pre-digital era.

The study builds on a rapidly growing body of work applying machine learning to petrology, from neural networks that classify rocks from thin-section images to sparse-modeling approaches for tectono-magmatic discrimination. What distinguishes this contribution is its combination of a large curated training set, systematic comparison of multiple algorithms, honest evaluation on independent test data, and a deliberate effort to interpret not only the correct classifications but also the uncertainties. As the authors note, the integration of extensive geochemical databases with interpretable machine learning facilitates genuinely quantitative discrimination of oceanic basalt, a small but meaningful step toward making the Earth’s chemical archive fully searchable by algorithm, with geologists still supplying the questions.

Subject of Research: Machine learning classification of oceanic basalts (MORB, OIB, and IAB) using geochemical data from the GEOROC and PetDB databases

Subject of Research: Earth Science

Article Title: Machine learning enhances research on the classification of oceanic basalt

Article References: Xu, H., Kong, J., & Ma, Y. (2026). Machine learning enhances research on the classification of oceanic basalt. Earth Science Informatics, 19(8), Article 137. https://doi.org/10.1007/s12145-026-02190-y

Image Credits: AI Generated

DOI: 10.1007/s12145-026-02190-y

Keywords: Machine learning, Random forest, Oceanic basalt classification, MORB-OIB-IAB, Probabilistic confidence analysis, Geochemical databases, Support vector machine, Discriminant diagrams

Cite Scienmag News

Blake Davidson. (September 11, 2026). Machine learning improves classification of oceanic basalt in research. Scienmag. https://scienmag.com/machine-learning-improves-classification-of-oceanic-basalt-in-research/

Blake Davidson. "Machine learning improves classification of oceanic basalt in research." Scienmag, 11 September 2026, https://scienmag.com/machine-learning-improves-classification-of-oceanic-basalt-in-research/. Accessed 11 September 2026.

Blake Davidson. "Machine learning improves classification of oceanic basalt in research." Scienmag. September 11, 2026. https://scienmag.com/machine-learning-improves-classification-of-oceanic-basalt-in-research/

Tags: advances in geological research with artificial intelligenceadvances in oceanic crust researchand IAB basaltsand IAB differentiationchemical fingerprint analysis of ocean floor basaltsclassification accuracy in geosciencesdata-driven approaches in earth sciencesdata-driven petrology analysisdistinguishing MORBgeochemical analysis of oceanic basaltsgeological processes and volcanic rock typesimpact of machine learning on geoscience classificationmachine learning accuracy in rock classificationmachine learning in Earth sciencesmantle source characteristics and basalt typesmantle source signatures in basaltsMORBocean floor volcanic rock originsocean floor volcanic rock researchoceanic basalt classification using machine learningOIBpetrology and geochemistry data-driven methodspetrology data-driven methodstectonic setting identification of volcanic rocksvolcanic origin determination in oceanic crust
Share26Tweet16
Previous Post

Multi-stage growth-aware maize yield prediction using graph neural networks

Next Post

Greenland sea ice and Arctic winds shape Himalayan summer rainfall

Related Posts

Greenland sea ice and Arctic winds shape Himalayan summer rainfall
Earth Science

Greenland sea ice and Arctic winds shape Himalayan summer rainfall

September 11, 2026
How oblique bedding gives rockslides lateral resistance: Shanyang case study
Earth Science

How oblique bedding gives rockslides lateral resistance: Shanyang case study

September 11, 2026
Human activities strongly influence forest edge effects across China
Earth Science

Human activities strongly influence forest edge effects across China

September 11, 2026
Taxonomic groups differ in what drives richness–depth gradients
Earth Science

Taxonomic groups differ in what drives richness–depth gradients

September 11, 2026
Rice Rises as Sugarcane Fades: Water Scarcity Quietly Rewrites India’s Crop Map
Earth Science

Rice Rises as Sugarcane Fades: Water Scarcity Quietly Rewrites India’s Crop Map

September 11, 2026
Machine learning and landscape metrics track forest fragmentation in Naqamte City, Ethiopia
Earth Science

Machine learning and landscape metrics track forest fragmentation in Naqamte City, Ethiopia

September 11, 2026
Next Post
Greenland sea ice and Arctic winds shape Himalayan summer rainfall

Greenland sea ice and Arctic winds shape Himalayan summer rainfall

  • Mothers who receive childcare support from maternal grandparents show more optimized

    Mothers who receive childcare support from maternal grandparents show more parental warmth, finds NTU Singapore study

    27656 shares
    Share 11059 Tweet 6912
  • University of Seville Breaks 120-Year-Old Mystery, Revises a Key Einstein Concept

    1061 shares
    Share 424 Tweet 265
  • Bee body mass, pathogens and local climate influence heat tolerance

    682 shares
    Share 273 Tweet 171
  • Researchers record first-ever images and data of a shark experiencing a boat strike

    546 shares
    Share 218 Tweet 137
  • Groundbreaking Clinical Trial Reveals Lubiprostone Enhances Kidney Function

    531 shares
    Share 212 Tweet 133
Science

Embark on a thrilling journey of discovery with Scienmag.com—your ultimate source for cutting-edge breakthroughs. Immerse yourself in a world where curiosity knows no limits and tomorrow’s possibilities become today’s reality!

RECENT NEWS

  • Greenland sea ice and Arctic winds shape Himalayan summer rainfall
  • Machine learning improves classification of oceanic basalt in research
  • Multi-stage growth-aware maize yield prediction using graph neural networks
  • Cotton gene GhMYB102 fights Verticillium wilt by boosting lignin production

Categories

  • Agriculture
  • Anthropology
  • Archaeology
  • Athmospheric
  • Biology
  • Biotechnology
  • Blog
  • Bussines
  • Cancer
  • Chemistry
  • Climate
  • Earth Science
  • Editorial Policy
  • Marine
  • Mathematics
  • Medicine
  • Pediatry
  • Policy
  • Psychology & Psychiatry
  • Science Education
  • Social Science
  • Space
  • Technology and Engineering

Subscribe to Blog via Email

Enter your email address to subscribe to this blog and receive notifications of new posts by email.

Join 5,151 other subscribers

© 2025 Scienmag - Science Magazine

Welcome Back!

Login to your account below

Forgotten Password?

Retrieve your password

Please enter your username or email address to reset your password.

Log In
No Result
View All Result
  • HOME
  • SCIENCE NEWS
  • CONTACT US

© 2025 Scienmag - Science Magazine

Discover more from Science

Subscribe now to keep reading and get access to the full archive.

Continue reading