Saturday, September 12, 2026
Science
No Result
View All Result
  • Login
  • HOME
  • SCIENCE NEWS
  • CONTACT US
  • HOME
  • SCIENCE NEWS
  • CONTACT US
No Result
View All Result
Scienmag
No Result
View All Result
Home Science News Medicine

Gene Selection Gets Smarter: Co-expression Networks Meet Genetic Algorithms

September 12, 2026
in Medicine
Juliet Wilcox
By Juliet Wilcox Scienmag Editorial Profile - Human Genetics
Reading Time: 5 mins read
0
Gene Selection Gets Smarter: Co-expression Networks Meet Genetic Algorithms

Gene Selection Gets Smarter: Co-expression Networks Meet Genetic Algorithms

Gene Selection Gets Smarter: Co-expression Networks Meet Genetic Algorithms

65
SHARES
587
VIEWS
Share on FacebookShare on Twitter
ADVERTISEMENT

Modern medicine is drowning in data, and a new study argues that the way out is not more computing power but a smarter partnership between biology and evolutionary mathematics. In research published in the Journal of Advanced Research, a team led by Zhilin Wang, Weiping Ding, Jinquan Zhang, Ali Asghar Heidari, Mingjing Wang and Huiling Chen introduces a feature selection framework called CJWGA, a Weighted Non-dominated Sorting Genetic Algorithm that combines gene co-expression networks with information-theoretic genetic operators. The method is designed to tackle one of the most stubborn problems in computational biology: how to find the handful of genes that truly matter for disease classification inside datasets containing thousands of candidate features, most of which are noise, redundancy, or statistical distraction.

The scale of the problem is easy to underestimate. High-throughput sequencing and mass spectrometry now allow laboratories to measure the expression of every gene in the human genome across hundreds of samples at once. A dataset might record 10,000 genes while including fewer than a hundred patients. This imbalance creates what statisticians call the curse of dimensionality: the number of possible feature subsets grows as two to the power of n, so for a dataset with 10,000 genes the search space is astronomically larger than anything a brute-force enumeration could ever cover. Worse, adding features does not reliably improve a model. Extra genes can introduce redundancy and noise, causing classifiers to overfit the training data while performing poorly on patients they have never seen. Running times grow as well, because computational complexity rises steadily with the number of features examined.

Existing feature selection strategies fall into three broad families, each with well-known trade-offs. Filtering methods, which rank genes using statistical measures such as mutual information, are fast and scalable but blind to the interactions between features. Wrapper methods, which evaluate subsets by feeding them to a classifier, capture those nonlinear relationships but at a punishing computational cost. Embedded methods such as LASSO regression and tree-based models select features during training, but none of these approaches ask the deeper biological question: which genes actually work together, and which modules of co-regulated genes drive the disease being studied? The new framework was built precisely to fill that gap, treating the biology of gene cooperation as the starting point rather than an afterthought.

The first stage of CJWGA relies on Weighted Gene Co-expression Network Analysis, or WGCNA, a technique originally proposed by Zhang and Horvath that constructs a weighted network linking genes whose expression levels rise and fall together across samples. Genes are not loners; they participate in biological processes through intricate webs of interaction, and WGCNA captures those relationships from a systems perspective. The pipeline begins with Z-score normalization of expression values, followed by a Pearson correlation matrix that is then transformed into a weighted adjacency matrix using a soft thresholding exponent chosen so the network follows a scale-free topology, in which a few highly connected hub genes dominate while most genes have few connections. A Topological Overlap Measure, which accounts for shared neighbors, is then fed into hierarchical clustering to identify modules of functionally related genes, with module eigengenes derived by principal component analysis.

But the authors recognized that relying on a single eigengene per module throws away too much information. A lone principal component cannot reflect the diversity of functions within a module, and it can be biased by outlier expression patterns. Their answer is a preprocessing step called IMGCNet, which uses conditional mutual information to rank genes within each module by how much extra information they carry about the disease label, given the eigengene is already known. A higher conditional mutual information value means a gene retains a strong dependency on the phenotype even after controlling for what the module representative already explains. Larger modules are allowed to retain more genes and smaller modules fewer, through a descending allocation rule that preserves the biological representativeness of each module without letting small, specialized groups flood the analysis.

The second stage hands the modules to an enhanced version of NSGA-II, the classic multi-objective genetic algorithm that balances competing goals by evolving a population of candidate solutions toward a Pareto front. Here the two objectives are minimizing the number of selected genes and maximizing classification accuracy, formalized with a binary decision vector over features and evaluated with a K-Nearest Neighbor classifier on a 70-30 train-test split. Crucially, the researchers designed a hierarchical encoding scheme: the first layer of each chromosome encodes which modules are selected, and the second layer encodes which genes within each chosen module survive. This two-layer structure preserves the biological meaning of the modularization rather than flattening it back into a flat string of bits.

The heart of the contribution lies in two new operators. The Combined Information Entropy Crossover Operator, or CIECO, computes a joint mutual information score across the genes selected in both parents, those selected in neither, and those selected in only one. The resulting value, transformed through a probabilistic function, decides whether crossover should prune doubly-selected genes, promote single-selected ones, or hold steady. When the score is positive, unselected genes carry little information and conservative trimming is favored; when it is negative, redundant double selections are removed and a few unselected genes are introduced to seek greater information content. The Joint Adaptive Mutation Operator, or JAMO, then fine-tunes individual genes using an adaptive rate that depends on iteration progress, the proportion of genes already selected in the module, and the ratio of joint mutual information between selected and unselected genes, with an exponent parameter that keeps the balance under control.

The experimental evaluation covered eight publicly available gene expression datasets, including Brain_Tumor1, Brain_Tumor2, CNS, Leukemia, Leukemia1, Leukemia2, Lung_Cancer and Prostate_Tumor, all with more than 5,000 features and sample sizes between 50 and 203. Against three specialist algorithms, FQEISS, WMOSS and WQEISS, CJWGA achieved the lowest classification error on the CNS, Leukemia, Leukemia1, Leukemia2 and Prostate_Tumor datasets, while selecting the smallest feature subsets on six of the eight datasets. On Inverted Generational Distance, a standard measure of how well a computed Pareto front approximates the true optimum, CJWGA scored zero, meaning perfect overlap with the reference front, on six datasets. Ablation experiments confirmed that both new operators contribute measurably: removing the crossover operator or the mutation operator individually degraded either accuracy or subset compactness. Parameter sweeps established that a crossover proportion of 0.2 and a mutation exponent of 3 offered the most robust results. Because joint mutual information is computed only within compact modules rather than across the entire feature space, the framework retains reasonable scalability even as dataset dimensionality grows.

The implications reach beyond benchmark tables. A feature selection method that respects gene co-expression relationships can point clinicians toward biologically meaningful biomarkers rather than statistical artifacts, a prerequisite for precision medicine where a compact, interpretable gene panel must support diagnosis and treatment decisions. The authors caution, however, that systematic biological interpretation of the selected genes remains future work, and they note that the framework could be extended to dimensionality reduction problems well outside genomics. As high-throughput biology continues to generate data faster than medicine can absorb it, tools like CJWGA suggest that the path forward lies in algorithms that speak both languages fluently: the language of information theory and the language of biological networks. The study is available as open access, supported by the National Natural Science Foundation of China and several provincial research programs.

Subject of Research: Gene feature selection in high-dimensional medical gene expression data using co-expression networks and genetic algorithms

Article Title: Advancing Gene Feature Selection: A Synergistic Approach with Co-expression Networks and Genetic Algorithms

Article References: Wangy, Z., Ding, W., Zhang, J., Heidari, A. A., Wang, M., & Chen, H. (2026). Advancing Gene Feature Selection: A Synergistic Approach with Co-expression Networks and Genetic Algorithms. Journal of Advanced Research. https://doi.org/10.1016/j.jare.2026.08.064

Image Credits: AI Generated

DOI: 10.1016/j.jare.2026.08.064

Keywords: gene feature selection, co-expression networks, WGCNA, genetic algorithms, multi-objective optimization, NSGA-II, mutual information, bioinformatics, cancer classification, precision medicine, high-dimensional data, biomarker discovery

Cite Scienmag News

Juliet Wilcox. (September 12, 2026). Gene Selection Gets Smarter: Co-expression Networks Meet Genetic Algorithms. Scienmag. https://scienmag.com/gene-selection-gets-smarter-co-expression-networks-meet-genetic-algorithms/

Juliet Wilcox. "Gene Selection Gets Smarter: Co-expression Networks Meet Genetic Algorithms." Scienmag, 12 September 2026, https://scienmag.com/gene-selection-gets-smarter-co-expression-networks-meet-genetic-algorithms/. Accessed 12 September 2026.

Juliet Wilcox. "Gene Selection Gets Smarter: Co-expression Networks Meet Genetic Algorithms." Scienmag. September 12, 2026. https://scienmag.com/gene-selection-gets-smarter-co-expression-networks-meet-genetic-algorithms/

Tags: bioinformaticsbioinformatics feature selection methodsbiomarker discoverycancer classificationco-expression networkscomputational biology data challengesdimensionality reduction in genomicsdisease classification gene markersgene co-expression network analysisgene feature selectiongenetic algorithmsgenetic algorithms for feature selectionhigh-dimensional datahigh-throughput sequencing data analysisinformation-theoretic genetic operatorsintegrating biology and evolutionary mathematicsmachine learning in biomedical datamulti-objective optimizationmutual informationnoise reduction in genetic datasetsNSGA-IIPrecision medicineWeighted Non-dominated Sorting Genetic AlgorithmWGCNA
Share26Tweet16
Previous Post

AI Reads Raman Spectra to Measure Cysteine in Pea Cultivars

Next Post

Mind-Reading Skills Take a Major Hit in Users of Alcohol, Opioids and Stimulants, Landmark Analysis Finds

Related Posts

Mind-Reading Skills Take a Major Hit in Users of Alcohol, Opioids and Stimulants, Landmark Analysis Finds
Medicine

Mind-Reading Skills Take a Major Hit in Users of Alcohol, Opioids and Stimulants, Landmark Analysis Finds

September 12, 2026
AMPA Receptor Blockers Show Promise for Brain Injury Recovery in Animal Studies
Medicine

AMPA Receptor Blockers Show Promise for Brain Injury Recovery in Animal Studies

September 12, 2026
Childhood Stomach Bug Infections Fall in China, but Cancer Burden Refuses to Shrink
Medicine

Childhood Stomach Bug Infections Fall in China, but Cancer Burden Refuses to Shrink

September 12, 2026
A Year-Long Trial Finds Tiotropium Fails to Prevent Severe Asthma Attacks in Preschoolers
Medicine

A Year-Long Trial Finds Tiotropium Fails to Prevent Severe Asthma Attacks in Preschoolers

September 12, 2026
Weight Returns Fast After Stopping Ozempic-Style Drugs, Major Analysis Finds
Medicine

Weight Returns Fast After Stopping Ozempic-Style Drugs, Major Analysis Finds

September 12, 2026
Fatigue During Immunotherapy Does Not Track With Fitness, Study Finds
Medicine

Fatigue During Immunotherapy Does Not Track With Fitness, Study Finds

September 12, 2026
Next Post
Mind-Reading Skills Take a Major Hit in Users of Alcohol, Opioids and Stimulants, Landmark Analysis Finds

Mind-Reading Skills Take a Major Hit in Users of Alcohol, Opioids and Stimulants, Landmark Analysis Finds

  • Mothers who receive childcare support from maternal grandparents show more optimized

    Mothers who receive childcare support from maternal grandparents show more parental warmth, finds NTU Singapore study

    27656 shares
    Share 11059 Tweet 6912
  • University of Seville Breaks 120-Year-Old Mystery, Revises a Key Einstein Concept

    1061 shares
    Share 424 Tweet 265
  • Bee body mass, pathogens and local climate influence heat tolerance

    682 shares
    Share 273 Tweet 171
  • Researchers record first-ever images and data of a shark experiencing a boat strike

    546 shares
    Share 218 Tweet 137
  • Groundbreaking Clinical Trial Reveals Lubiprostone Enhances Kidney Function

    531 shares
    Share 212 Tweet 133
Science

Embark on a thrilling journey of discovery with Scienmag.com—your ultimate source for cutting-edge breakthroughs. Immerse yourself in a world where curiosity knows no limits and tomorrow’s possibilities become today’s reality!

RECENT NEWS

  • Scientists Are Scrubbing Their Own Vocabulary to Survive Federal Funding Pressure
  • Mind-Reading Skills Take a Major Hit in Users of Alcohol, Opioids and Stimulants, Landmark Analysis Finds
  • Gene Selection Gets Smarter: Co-expression Networks Meet Genetic Algorithms
  • AI Reads Raman Spectra to Measure Cysteine in Pea Cultivars

Categories

  • Agriculture
  • Anthropology
  • Archaeology
  • Athmospheric
  • Biology
  • Biotechnology
  • Blog
  • Bussines
  • Cancer
  • Chemistry
  • Climate
  • Earth Science
  • Editorial Policy
  • Marine
  • Mathematics
  • Medicine
  • Pediatry
  • Policy
  • Psychology & Psychiatry
  • Science Education
  • Social Science
  • Space
  • Technology and Engineering

Subscribe to Blog via Email

Enter your email address to subscribe to this blog and receive notifications of new posts by email.

Join 5,151 other subscribers

© 2025 Scienmag - Science Magazine

Welcome Back!

Login to your account below

Forgotten Password?

Retrieve your password

Please enter your username or email address to reset your password.

Log In
No Result
View All Result
  • HOME
  • SCIENCE NEWS
  • CONTACT US

© 2025 Scienmag - Science Magazine

Discover more from Science

Subscribe now to keep reading and get access to the full archive.

Continue reading