Friday, October 2, 2026
Science
No Result
View All Result
  • Login
  • HOME
  • SCIENCE NEWS
  • CONTACT US
  • HOME
  • SCIENCE NEWS
  • CONTACT US
No Result
View All Result
Scienmag
No Result
View All Result
Home Science News Technology and Engineering

AI Meets Human Judgment: A Governed Framework for Smarter Supplier Selection

October 2, 2026
in Technology and Engineering
Denise Maddox
By Denise Maddox Scienmag Editorial Profile - Mechanical Engineering
Reading Time: 5 mins read
0
AI Meets Human Judgment: A Governed Framework for Smarter Supplier Selection

AI Meets Human Judgment: A Governed Framework for Smarter Supplier Selection

AI Meets Human Judgment: A Governed Framework for Smarter Supplier Selection

65
SHARES
587
VIEWS
Share on FacebookShare on Twitter
ADVERTISEMENT

Choosing the right supplier has never been a simple matter of comparing prices. Modern procurement teams must weigh cost against quality, delivery reliability against sustainability credentials, and hard compliance rules against the messy, contradictory evidence that arrives in supplier documents, news reports, and certification records. A new study published in Machine Learning with Applications proposes a way to bring order to this complexity: a governed framework that combines the linguistic fluency of large language models with the unyielding precision of deterministic mathematics, all under the watchful eyes of human decision makers.

The research, conducted by Mohammadreza Rezaei, Sridhar Sreemulnath Iyer, and Omid Fatahi Valilai, tackles a gap that has long frustrated both researchers and practitioners. Multi-criteria decision-making methods such as TOPSIS can rank suppliers rigorously, but only after someone has formalised the criteria, weights, constraints, and evidence. Large language models, meanwhile, excel at reading natural-language procurement requests and extracting information from unstructured documents, yet they cannot be trusted to perform arithmetic or to rank suppliers autonomously. Existing approaches, the authors argue, address fragments of the problem while leaving the complete path from a spoken or written sourcing request to a traceable, defensible supplier decision largely uncharted.

The framework’s central design principle is a strict division of labour. Language model components handle the semantic work: interpreting what a procurement request actually means, formulating candidate evaluation scenarios, extracting claims from supplier documents, and explaining completed results in plain language. Deterministic services handle everything numerical: constructing weight vectors, applying constraints and evidence policies, calculating multi-criteria scores, resolving ties, and fixing the final supplier order. Crucially, the language models never touch the numbers. When a manager asks to place greater emphasis on geographic fit, the model proposes only a direction and a bounded strength category; a deterministic controller then converts that advice into a valid numerical weight vector through a constrained projection that enforces unit sums and policy-defined bounds.

Human authority is organised through four checkpoints. The data owner approves the structured supplier table and its quality record. The procurement analyst approves the scenario, its criteria, constraints, and the numerical weights. The evidence reviewer confirms whether extracted claims genuinely support their cited passages and what analytical role each claim may play. Finally, the decision owner reviews the complete package, including rankings, comparisons, explanations, and caveats, before any action is taken. Every consequential action generates an append-only audit event recording who acted, in what role, when, on what object, and why, creating a traceable chain from the original request to the final sourcing decision.

The empirical evaluation is unusually comprehensive for this field. Using more than 22,000 real procurement cases drawn from the European Union’s Tenders Electronic Daily database, the team first compared deterministic ranking methods. TOPSIS, which ranks suppliers by their distance from ideal and anti-ideal solutions after vector normalisation and weighting, outperformed frequency-based ranking, weighted-sum models, and VIKOR. But the more striking finding concerned the supplier representation itself. A generic four-feature baseline placed the historically awarded supplier in the top ten of only 5.259 percent of cases. When the researchers added procurement-specific feature groups capturing category fit, geographic relevance, and prior buyer relationships, that figure rose to 31.642 percent, and the median rank of the reference supplier improved from 914 to just 34.

The scenario experiments revealed how the framework responds when sourcing priorities change. Under a continuity profile that increases the weight on prior buyer relationships, the top-ten overlap with the balanced baseline remained high at 0.9369. Under a diversification profile that reduces reliance on existing relationships, overlap fell to 0.8304, yet suppliers entering the top ten had stronger geographic fit than those exiting in 96.984 percent of turnover cases. In other words, the rankings moved in traceable, explainable ways that followed the approved priorities, allowing procurement professionals to compare competing strategies and see exactly which suppliers enter or leave the shortlist under each.

The language model experiments were equally revealing. In intent formulation, the model correctly identified the scenario in 146 of 150 procurement requests and made the right proceed-or-clarify decision in all but one, with constraint and evidence-request extraction achieving F1 scores of 0.990 and 0.996 respectively. Evidence extraction proved robust: across 40 supplier evidence packets containing 160 expected claims, the module achieved a mean claim F1 of 0.928 and grounded 99.5 percent of its quotations in the actual source text, with unsupported quotations appearing in only 0.5 percent of cases. Yet the study is candid about weaknesses. Temporal-status accuracy fell to 0.834 and evidence-role accuracy to 0.748, meaning a claim can be quoted perfectly while still being assigned the wrong analytical meaning. Under conflicting evidence, performance degraded further, with hallucination rising to 4.3 percent.

Perhaps the most instructive experiment examined what happens when extracted qualitative evidence feeds into a deterministic ranking. In a controlled benchmark combining profit, lead time, and ESG evidence, all six material compliance violations were detected and eligibility accuracy was perfect. However, all twelve suppliers with mixed sustainability evidence, those with an emissions-reduction plan whose achieved reductions remained unverified, were mapped to an overly favourable certified state. The consequence was subtle: shortlists remained largely intact, with top-three overlap between 0.889 and 1.000, but the first-ranked supplier changed in many cases. The lesson is clear and important. High grounding and perfect violation detection can coexist with incorrect semantic qualification of evidence, which is precisely why the framework routes consequential evidence through human review before it can affect eligibility or scores.

Explanation fidelity completed the picture. In 24 controlled cases, the explanation module preserved the exact deterministic supplier order, used only valid supplier and evidence identifiers, and correctly escalated every case involving constraints, close score margins, negative evidence, or missing data. Its main weakness was over-inclusiveness: explanations sometimes mentioned all available feature groups rather than isolating those that genuinely drove the ranking. The authors treat this as a review item rather than a failure, presenting the narrative alongside the fixed score tables so the decision owner can inspect both.

The broader significance of this work lies in its refusal to choose between automation and augmentation. Rather than asking whether language models can replace procurement professionals, the framework specifies exactly which tasks each participant may perform and where approval is required. It also supports adaptive supplier reassessment: when market conditions, regulations, or supplier circumstances change, a revised scenario passes through the same validated, approved, deterministic pathway rather than someone quietly editing an existing ranking. The authors acknowledge that organisational deployment, including review effort and decision-owner behaviour, still requires field evaluation, and that benchmarking across multiple model families remains future work. But as a blueprint for bringing generative AI into high-stakes sourcing decisions without surrendering accountability, the study offers something rare: a complete, tested architecture in which the flexibility of language meets the rigour of mathematics, and where the final word always belongs to a human being.

Subject of Research: A governed human-in-the-loop framework integrating large language models with deterministic multi-criteria supplier evaluation

Article Title: A governed framework combining large language models and human oversight for supplier selection

Article References: Rezaei, M., Iyer, S. S., & Fatahi Valilai, O. (2026). A governed framework combining large language models and human oversight for supplier selection. Machine Learning with Applications, 26, Article 101028. https://doi.org/10.1016/j.mlwa.2026.101028

Image Credits: AI Generated

DOI: 10.1016/j.mlwa.2026.101028

Keywords: large language models, supplier selection, procurement, multi-criteria decision-making, TOPSIS, human-in-the-loop, supply chain management, explainable AI, evidence extraction, AI governance, sustainable sourcing, auditability

Cite Scienmag News

Denise Maddox. (October 2, 2026). AI Meets Human Judgment: A Governed Framework for Smarter Supplier Selection. Scienmag. https://scienmag.com/ai-meets-human-judgment-a-governed-framework-for-smarter-supplier-selection/

Denise Maddox. "AI Meets Human Judgment: A Governed Framework for Smarter Supplier Selection." Scienmag, 2 October 2026, https://scienmag.com/ai-meets-human-judgment-a-governed-framework-for-smarter-supplier-selection/. Accessed 2 October 2026.

Denise Maddox. "AI Meets Human Judgment: A Governed Framework for Smarter Supplier Selection." Scienmag. October 2, 2026. https://scienmag.com/ai-meets-human-judgment-a-governed-framework-for-smarter-supplier-selection/

Tags: AI governanceAI-driven supplier selection frameworkauditabilitychallenges of trust and reliability in AI procurement toolscombining NLP and deterministic mathematics in procurementcriteria formalization in supplier selectionevidence extractionexplainable AIgoverned decision-making in procurementhandling unstructured procurement documents with AIhuman oversight in AI-powered supplier rankinghuman-in-the-loopimproving procurement accuracy with AI-human collaborationintegration of large language models in supplier evaluationlarge language modelsmulti-criteria decision makingmulti-criteria decision-making methods for sourcingprocurementsupplier selectionSupply Chain Managementsustainable and compliant supplier assessmentsustainable sourcingTOPSIStransparency and traceability in supplier decisions
Share26Tweet16
Previous Post

Ultra-Processed Food Debate: Nova Researchers Defend the Evidence on Human Health

Next Post

Assam’s Traditional Ahu Rice Landraces Reveal Hidden Genetic Keys to Drought Tolerance

Related Posts

Fuzzy Math Gets Sharper: New Decision Tool Tames Uncertainty in Supplier Choices
Technology and Engineering

Fuzzy Math Gets Sharper: New Decision Tool Tames Uncertainty in Supplier Choices

October 2, 2026
Self-Reporting Nanoparticle Turns Tumor Hypoxia Into a Weapon Against Cancer
Technology and Engineering

Self-Reporting Nanoparticle Turns Tumor Hypoxia Into a Weapon Against Cancer

October 2, 2026
Softer Airway Device Linked to Faster Breathing Recovery in Preterm Infants After Eye Injection
Technology and Engineering

Softer Airway Device Linked to Faster Breathing Recovery in Preterm Infants After Eye Injection

October 2, 2026
Smarter Algorithms Could Shield Billions of IoT Devices From Hackers
Technology and Engineering

Smarter Algorithms Could Shield Billions of IoT Devices From Hackers

October 2, 2026
Scientists Track the Mineral That Crumbles Concrete From the Inside Out
Technology and Engineering

Scientists Track the Mineral That Crumbles Concrete From the Inside Out

October 2, 2026
Ancient Herb Compound Purges Zombie Cells That Make Arteries Fragile
Technology and Engineering

Ancient Herb Compound Purges Zombie Cells That Make Arteries Fragile

October 2, 2026
Next Post
Assam’s Traditional Ahu Rice Landraces Reveal Hidden Genetic Keys to Drought Tolerance

Assam's Traditional Ahu Rice Landraces Reveal Hidden Genetic Keys to Drought Tolerance

  • Mothers who receive childcare support from maternal grandparents show more optimized

    Mothers who receive childcare support from maternal grandparents show more parental warmth, finds NTU Singapore study

    27656 shares
    Share 11059 Tweet 6912
  • University of Seville Breaks 120-Year-Old Mystery, Revises a Key Einstein Concept

    1061 shares
    Share 424 Tweet 265
  • Bee body mass, pathogens and local climate influence heat tolerance

    682 shares
    Share 273 Tweet 171
  • Researchers record first-ever images and data of a shark experiencing a boat strike

    546 shares
    Share 218 Tweet 137
  • Groundbreaking Clinical Trial Reveals Lubiprostone Enhances Kidney Function

    531 shares
    Share 212 Tweet 133
Science

Embark on a thrilling journey of discovery with Scienmag.com—your ultimate source for cutting-edge breakthroughs. Immerse yourself in a world where curiosity knows no limits and tomorrow’s possibilities become today’s reality!

RECENT NEWS

  • Assam’s Traditional Ahu Rice Landraces Reveal Hidden Genetic Keys to Drought Tolerance
  • AI Meets Human Judgment: A Governed Framework for Smarter Supplier Selection
  • Ultra-Processed Food Debate: Nova Researchers Defend the Evidence on Human Health
  • Vitamin Cocktail for Seeds Shields Rapeseed From Cadmium Damage

Categories

  • Agriculture
  • Anthropology
  • Archaeology
  • Athmospheric
  • Biology
  • Biotechnology
  • Blog
  • Bussines
  • Cancer
  • Chemistry
  • Climate
  • Earth Science
  • Editorial Policy
  • Marine
  • Mathematics
  • Medicine
  • Pediatry
  • Policy
  • Psychology & Psychiatry
  • Science Education
  • Social Science
  • Space
  • Technology and Engineering

Subscribe to Blog via Email

Enter your email address to subscribe to this blog and receive notifications of new posts by email.

Join 5,151 other subscribers

© 2025 Scienmag - Science Magazine

Welcome Back!

Login to your account below

Forgotten Password?

Retrieve your password

Please enter your username or email address to reset your password.

Log In
No Result
View All Result
  • HOME
  • SCIENCE NEWS
  • CONTACT US

© 2025 Scienmag - Science Magazine

Discover more from Science

Subscribe now to keep reading and get access to the full archive.

Continue reading