<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"
	xmlns:content="http://purl.org/rss/1.0/modules/content/"
	xmlns:wfw="http://wellformedweb.org/CommentAPI/"
	xmlns:dc="http://purl.org/dc/elements/1.1/"
	xmlns:atom="http://www.w3.org/2005/Atom"
	xmlns:sy="http://purl.org/rss/1.0/modules/syndication/"
	xmlns:slash="http://purl.org/rss/1.0/modules/slash/"
	>

<channel>
	<title>semi-supervised regression &#8211; Science</title>
	<atom:link href="https://scienmag.com/tag/semi-supervised-regression/feed/" rel="self" type="application/rss+xml" />
	<link>https://scienmag.com</link>
	<description></description>
	<lastBuildDate>Sun, 04 Oct 2026 13:49:43 +0000</lastBuildDate>
	<language>en-US</language>
	<sy:updatePeriod>
	hourly	</sy:updatePeriod>
	<sy:updateFrequency>
	1	</sy:updateFrequency>
	<generator>https://wordpress.org/?v=7.1.2</generator>

<image>
	<url>https://scienmag.com/wp-content/uploads/2024/07/cropped-scienmag_ico-32x32.jpg</url>
	<title>semi-supervised regression &#8211; Science</title>
	<link>https://scienmag.com</link>
	<width>32</width>
	<height>32</height>
</image> 
<site xmlns="com-wordpress:feed-additions:1">73899611</site>	<item>
		<title>New AI Framework Turns Every Unlabeled Data Point Into a Trustworthy Teacher</title>
		<link>https://scienmag.com/new-ai-framework-turns-every-unlabeled-data-point-into-a-trustworthy-teacher/</link>
		
		<dc:creator><![CDATA[Denise Maddox]]></dc:creator>
		<pubDate>Sun, 04 Oct 2026 13:49:43 +0000</pubDate>
				<category><![CDATA[Technology and Engineering]]></category>
		<category><![CDATA[AI in petroleum industry]]></category>
		<category><![CDATA[applications in environmental monitoring]]></category>
		<category><![CDATA[Applied Intelligence]]></category>
		<category><![CDATA[confidence weighting]]></category>
		<category><![CDATA[data similarity]]></category>
		<category><![CDATA[FullReg semi-supervised learning method]]></category>
		<category><![CDATA[handling noisy predictions in regression]]></category>
		<category><![CDATA[improving model accuracy with unlabeled data]]></category>
		<category><![CDATA[leveraging unlabeled data for machine learning]]></category>
		<category><![CDATA[loss function]]></category>
		<category><![CDATA[Machine learning]]></category>
		<category><![CDATA[neural networks]]></category>
		<category><![CDATA[new AI framework for unlabeled data]]></category>
		<category><![CDATA[pseudo-labels]]></category>
		<category><![CDATA[reducing reliance on labeled data]]></category>
		<category><![CDATA[residual connections]]></category>
		<category><![CDATA[semi-supervised regression]]></category>
		<category><![CDATA[semi-supervised regression challenges]]></category>
		<category><![CDATA[solar photovoltaic forecasting]]></category>
		<category><![CDATA[Southwest Petroleum University]]></category>
		<category><![CDATA[trustworthiness of AI models]]></category>
		<category><![CDATA[unlabeled data]]></category>
		<category><![CDATA[unlabeled data in machine learning]]></category>
		<guid isPermaLink="false">https://scienmag.com/?p=235218</guid>

					<description><![CDATA[Researchers in China have developed FullReg, a semi-supervised regression framework that weights pseudo-labels by data similarity and stabilizes training with cross-epoch residual connections, outperforming eight state-of-the-art algorithms on benchmarks spanning biomedical, business, ecological, physical, and solar energy data.]]></description>
										<content:encoded><![CDATA[<p>Machine learning models are famously hungry for labeled data, but in most real-world settings, labels are expensive, slow, and sometimes impossible to obtain at scale. A research team at Southwest Petroleum University in Chengdu, China, has now unveiled a new framework that promises to squeeze far more value out of the vast pools of unlabeled data that surround every labeled dataset. The method, called FullReg, is described in a study published in the journal Applied Intelligence, and it tackles one of the most persistent weaknesses in semi-supervised regression: what to do with the noisy, unreliable predictions that models generate for data they have never been taught to label.</p>
<p>Semi-supervised regression sits at the intersection of two worlds. In supervised learning, every training example comes with a known target value, such as the exact energy output of a solar panel or the measured concentration of a pollutant. In unsupervised learning, the algorithm must find structure without any answers at all. Semi-supervised methods try to have it both ways, using a small labeled set to anchor the model and a much larger unlabeled set to refine it. The promise is enormous, because collecting raw measurements is usually far cheaper than annotating them, and domains from biomedicine to finance to industrial manufacturing are drowning in unannotated numerical data.</p>
<p>The dominant strategies in this field have long followed a conservative philosophy. Early approaches selected only a small number of high-confidence unlabeled examples and folded them into the training data, effectively discarding the rest. This filtering kept the training signal clean, but it also threw away most of the information contained in the unlabeled pool. More recent methods took the opposite approach, using off-the-shelf semi-supervised regressors to generate pseudo-labels, which are the model&#8217;s own predictions treated as if they were ground truth, for every unlabeled example. The problem, as the Chinese team points out, is that these methods treat all pseudo-labels uniformly during training, ignoring the inherent quality differences among them and potentially injecting significant noise into the learning process.</p>
<p>FullReg addresses this weakness with two interlocking mechanisms. The first is a confidence-weighting scheme based on data similarity. Rather than accepting every pseudo-label at face value, the framework assigns each one a weight in the loss function that reflects how trustworthy it is likely to be. The intuition is geometric: if an unlabeled example sits close to labeled examples in the input space, its neighbors&#8217; known target values provide meaningful evidence about what its own target should be, so its pseudo-label earns a high weight. If an unlabeled point floats in a sparse region far from any labeled data, the model&#8217;s guess about its value is essentially unsupported, and the weight drops accordingly. By modulating the contribution of each pseudo-label to the overall loss, the mechanism dampens the influence of unreliable predictions and preserves the validity of the training signal.</p>
<p>This idea of weighting by similarity has deep roots in statistical learning, where kernel methods and locally weighted regression have long recognized that predictions are more reliable near observed data. What FullReg adds is a systematic way to translate that geometric intuition into the training dynamics of a neural network performing semi-supervised regression. The result is a framework that can exploit the entire unlabeled pool, as the newer generation of methods does, while retaining the noise resistance that made the older, selective approaches robust. In effect, the model no longer has to choose between using all of its data and trusting what that data tells it.</p>
<p>The second innovation is a residual-connection mechanism that operates across training epochs rather than across network layers. The name deliberately echoes the residual connections popularized by deep residual networks in computer vision, where skip links allow information to bypass layers and stabilize training. Here, the connection is temporal: at each epoch, the framework blends the model parameters inherited from previous epochs with the parameters being learned in the current one. Instead of letting the network lurch toward whatever solution the latest batch of data suggests, the residual mechanism anchors it to its own history, producing a trajectory of parameter updates that evolves progressively rather than erratically.</p>
<p>This temporal smoothing serves a similar purpose to techniques such as temporal ensembling and weight averaging, which have been shown in prior research to lead neural networks toward wider optima and better generalization. It also echoes the mean teacher paradigm, in which an averaged copy of a model provides steadier training targets than the model itself. By embedding that stabilizing principle directly into the parameter updates of a semi-supervised regression pipeline, FullReg gains resilience against the fluctuations that pseudo-label noise would otherwise introduce. The two mechanisms reinforce each other: confidence weighting reduces the noise entering the loss, while residual connections prevent whatever noise remains from destabilizing the learned parameters.</p>
<p>To test the framework, the researchers ran experiments on benchmark datasets drawn from five distinct domains, spanning biomedical, business, ecology, physical, and life sciences data, sourced from public repositories including the UCI Machine Learning Repository, the Delve repository, and the StatLib archive. They also evaluated the method on a real-world solar photovoltaic dataset, a setting where accurate regression matters for forecasting power generation from grid-connected installations. Across these benchmarks, FullReg was compared against eight state-of-the-art semi-supervised regression algorithms, and it compared favorably in most benchmark settings, suggesting that the combination of full data utilization and noise-aware training translates into measurable predictive gains rather than merely theoretical elegance.</p>
<p>The practical implications extend well beyond benchmark tables. Consider solar power forecasting, where weather stations, inverter readings, and satellite imagery generate torrents of measurements but ground-truth labels for every operating condition are scarce. A framework that can safely exploit all of that unlabeled data, rather than a hand-picked high-confidence subset, could sharpen the forecasts that grid operators rely on to balance supply and demand. Similar logic applies to air temperature mapping, stock price prediction during volatile periods, thermal error compensation in precision manufacturing, and clinical risk prediction, all of which are cited in the study&#8217;s bibliography as active application areas for semi-supervised regression. In each case, the bottleneck is the same: labeled examples are few, unlabeled examples are plentiful, and the quality of machine-generated labels varies wildly.</p>
<p>The work also contributes to a broader conversation in machine learning about how models should treat their own outputs. Pseudo-labeling has become a cornerstone of modern semi-supervised learning, powering influential techniques in image classification, semantic segmentation, and few-shot learning, yet the calibration of pseudo-label quality remains an open problem. FullReg&#8217;s answer, grounding confidence in data similarity and stabilizing learning through temporal residual connections, offers a template that other researchers may adapt to classification and other tasks. The authors have made their benchmark analysis transparent, drawing on publicly available datasets, with the solar photovoltaic dataset and code available from the corresponding author on reasonable request. As unlabeled data continues to accumulate faster than any labeling effort could match, methods like this one, which learn to distrust their own mistakes in a principled way, may define the next generation of practical machine learning.</p>
<p><strong>Subject of Research:</strong> Semi-supervised regression using confidence-weighted pseudo-labels and residual parameter connections</p>
<p><strong>Article Title:</strong> Semi-supervised regression via confidence-weighting and residual-connection</p>
<p><strong>Article References:</strong> Liu, L., Mao, Y., Lu, X., &amp; Min, F. (2026). Semi-supervised regression via confidence-weighting and residual-connection. <em>Applied Intelligence, 56</em>(15), Article 440. <a href="https://doi.org/10.1007/s10489-026-07489-3" rel="noopener noreferrer">https://doi.org/10.1007/s10489-026-07489-3</a></p>
<p><strong>Image Credits:</strong> AI Generated</p>
<p><strong>DOI:</strong> <a href="https://doi.org/10.1007/s10489-026-07489-3" rel="noopener noreferrer">10.1007/s10489-026-07489-3</a></p>
<p><strong>Keywords:</strong> semi-supervised regression, pseudo-labels, confidence weighting, data similarity, residual connections, neural networks, machine learning, unlabeled data, solar photovoltaic forecasting, Applied Intelligence, Southwest Petroleum University, loss function</p>
]]></content:encoded>
					
		
		
		<post-id xmlns="com-wordpress:feed-additions:1">235218</post-id>	</item>
	</channel>
</rss>
