<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"
	xmlns:content="http://purl.org/rss/1.0/modules/content/"
	xmlns:wfw="http://wellformedweb.org/CommentAPI/"
	xmlns:dc="http://purl.org/dc/elements/1.1/"
	xmlns:atom="http://www.w3.org/2005/Atom"
	xmlns:sy="http://purl.org/rss/1.0/modules/syndication/"
	xmlns:slash="http://purl.org/rss/1.0/modules/slash/"
	>

<channel>
	<title>garbage recognition using self-attention mechanisms &#8211; Science</title>
	<atom:link href="https://scienmag.com/tag/garbage-recognition-using-self-attention-mechanisms/feed/" rel="self" type="application/rss+xml" />
	<link>https://scienmag.com</link>
	<description></description>
	<lastBuildDate>Sun, 11 Oct 2026 00:29:51 +0000</lastBuildDate>
	<language>en-US</language>
	<sy:updatePeriod>
	hourly	</sy:updatePeriod>
	<sy:updateFrequency>
	1	</sy:updateFrequency>
	<generator>https://wordpress.org/?v=7.1.3</generator>

<image>
	<url>https://scienmag.com/wp-content/uploads/2024/07/cropped-scienmag_ico-32x32.jpg</url>
	<title>garbage recognition using self-attention mechanisms &#8211; Science</title>
	<link>https://scienmag.com</link>
	<width>32</width>
	<height>32</height>
</image> 
<site xmlns="com-wordpress:feed-additions:1">73899611</site>	<item>
		<title>Lightweight Transformer Brings Real-Time Road Garbage Detection to Edge Devices</title>
		<link>https://scienmag.com/lightweight-transformer-brings-real-time-road-garbage-detection-to-edge-devices/</link>
		
		<dc:creator><![CDATA[Blake Davidson]]></dc:creator>
		<pubDate>Sun, 11 Oct 2026 00:29:51 +0000</pubDate>
				<category><![CDATA[Technology and Engineering]]></category>
		<category><![CDATA[attention mechanism]]></category>
		<category><![CDATA[city-scale waste management automation]]></category>
		<category><![CDATA[cleanliness assessment]]></category>
		<category><![CDATA[computer vision]]></category>
		<category><![CDATA[computer vision for street cleanliness]]></category>
		<category><![CDATA[deep learning]]></category>
		<category><![CDATA[deep learning for environmental monitoring]]></category>
		<category><![CDATA[edge computing]]></category>
		<category><![CDATA[edge device artificial intelligence]]></category>
		<category><![CDATA[embedded deployment]]></category>
		<category><![CDATA[garbage recognition using self-attention mechanisms]]></category>
		<category><![CDATA[lightweight model]]></category>
		<category><![CDATA[lightweight transformer models]]></category>
		<category><![CDATA[low-power hardware AI applications]]></category>
		<category><![CDATA[real-time urban waste monitoring]]></category>
		<category><![CDATA[resource-efficient AI for smart cities]]></category>
		<category><![CDATA[road garbage detection]]></category>
		<category><![CDATA[SegFormer]]></category>
		<category><![CDATA[semantic segmentation]]></category>
		<category><![CDATA[Transformer]]></category>
		<category><![CDATA[transformer attention mechanism optimization]]></category>
		<category><![CDATA[trash segmentation on embedded systems]]></category>
		<category><![CDATA[waste management]]></category>
		<guid isPermaLink="false">https://scienmag.com/?p=260494</guid>

					<description><![CDATA[Researchers in China have developed a lightweight transformer model with a novel cross-axis attention module and a feature memory component that boosts road garbage segmentation speed and accuracy on embedded hardware, enabling real-time street cleanliness monitoring.]]></description>
										<content:encoded><![CDATA[<p>A team of researchers in China has unveiled a lightweight transformer-based artificial intelligence model that can identify and map garbage on roads in real time, even on the kind of modest, low-power hardware that fits inside a street-cleaning vehicle. The work, published in the journal Neural Computing and Applications, tackles one of the most persistent obstacles in practical computer vision: the gap between the impressive accuracy of large transformer models in the laboratory and their inability to run on the resource-constrained edge devices where they are actually needed. By redesigning the attention mechanism at the heart of the network and adding a novel module that remembers what categories of trash it has seen, the researchers report both faster inference and better segmentation accuracy on an embedded development board, suggesting a realistic path toward automated, city-scale street cleanliness monitoring.</p>
<p>Transformer architectures have transformed computer vision since their introduction, replacing or augmenting convolutional neural networks with self-attention mechanisms that let a model weigh relationships between distant parts of an image. For road garbage recognition, this global context is valuable: a crumpled plastic bag half-hidden at the edge of a curb can be distinguished from a patch of fallen leaves only by considering the surrounding scene. The problem is computational cost. Self-attention scales poorly with the number of image tokens, and full-size vision transformers demand memory and processing power far beyond what an embedded board mounted on a cleaning vehicle can supply. Earlier approaches to garbage detection relied on convolutional networks or heavyweight transformers that either sacrificed contextual understanding or required cloud-level hardware, neither of which suits a camera bolted to a truck sweeping a street at dawn.</p>
<p>The new model, developed by Suheng Peng, Jiacai Liao, Libo Cao, Cong Duan and Minghai Zhang at Hunan University and Changsha University of Science and Technology, builds on SegFormer, a transformer architecture already designed with efficiency in mind. SegFormer uses a hierarchical encoder that processes the image at multiple scales and a simple, all-MLP decoder, avoiding the most expensive parts of classic transformer designs. But even SegFormer&#8217;s attention blocks can strain an embedded processor. The team&#8217;s central contribution is a module they call Residual Multi-axis Hadamard Attention, which reorganizes how the network computes attention so that it remains expressive while dramatically cutting the arithmetic burden.</p>
<p>The key idea is to split the feature representation into groups along different axes of the image and then fuse information across those groups using element-wise multiplication, the operation mathematicians call the Hadamard product. Instead of letting every image token attend to every other token, which grows quadratically with image size, the module promotes what the authors describe as complementary feature interactions across multiple axes through residual cross-axis feature-weighted fusion. In practical terms, features computed along the horizontal axis of an image are weighted and combined with features computed along the vertical axis, and a residual connection preserves the original information through the operation. This enhances the sharing of information between groups without reconstructing the full attention matrix, keeping computational complexity low while still allowing the model to build an efficient representation of the scene&#8217;s context.</p>
<p>Attention redesign alone, however, does not address a second, subtler problem in garbage segmentation: class imbalance and category confusion. Roads are mostly clean surfaces, and the objects that litter them come in wildly varying shapes, sizes and frequencies. Small, thin items such as bottles or paper scraps occupy few pixels, while large debris dominates. A segmentation network can easily learn to favor the frequent background classes and blur the rare or visually similar garbage categories together. To counter this, the researchers propose a Feature Memory module, a component that stores and continuously updates a memory of category distribution information as the network processes images. By consulting this memory, the model can reinforce the characteristic features of each garbage category and improve its feature representations for classes that appear infrequently or look alike, effectively giving the network a running summary of what it has learned about each type of trash.</p>
<p>The two modules were integrated into a lightweight SegFormer backbone and evaluated on the RK3568, a low-power embedded development board of the sort that could realistically be installed on municipal cleaning vehicles. The results were striking on both fronts. Inference speed increased by 12.48 frames per second compared with the baseline, a meaningful margin for a real-time system that must keep pace with a moving vehicle, and the mean intersection-over-union score, the standard accuracy metric for semantic segmentation, improved by 4.54 percent. The model outperformed other semantic segmentation approaches tested under the same conditions, indicating that the gains were not simply a matter of trimming the network until it was fast but inaccurate. The combination of higher speed and higher accuracy on constrained hardware is precisely the trade-off that most lightweight architectures struggle to achieve.</p>
<p>Beyond segmentation itself, the team also revisited how street cleanliness should be measured. Cleanliness assessment has historically been a subjective exercise; earlier studies have compared municipal scorecards with citizen surveys and proposed quantitative indices for specific cities, but these approaches are labor-intensive and inconsistent. The researchers reformulated the road cleanliness assessment metric used in their pipeline to improve the accuracy and stability of the overall assessment system. With a segmentation model that reliably delineates garbage regions frame by frame, a well-defined metric can convert those pixel-level predictions into a stable, reproducible cleanliness score, turning a fleet of camera-equipped vehicles into a continuous environmental monitoring network rather than an occasional inspection tool.</p>
<p>The practical implications extend across urban management. Cities spend substantial budgets on street sweeping, and decisions about where to deploy cleaning resources are often guided by sparse manual inspections. A real-time, edge-deployed system that segments garbage, classifies it by category and produces a quantitative cleanliness index could allow routing decisions to be made on current data, direct vehicles to the dirtiest streets first, and provide objective evidence of service quality. The authors position the work as bridging the gap between high-accuracy transformer-based models and practical edge deployment, offering what they describe as a cost-effective solution for real-time environmental monitoring. Because the model runs on an inexpensive development board rather than a server, the marginal cost of adding intelligence to each vehicle is low.</p>
<p>The study also sits within a broader research movement applying deep learning to waste management, from convolutional detectors for recycling facilities to vision transformers for waste classification and multimodal systems for plastic identification. What distinguishes this contribution is its explicit focus on the deployment constraint: rather than chasing leaderboard accuracy on powerful GPUs, the authors optimized for the hardware that municipal systems can actually afford to install. The datasets generated during the study are not publicly available because they form part of an ongoing investigation, though the authors state they are available from the corresponding author on reasonable request. If the approach proves robust in field trials, the sight of a street-sweeping truck that not only cleans the road but also maps, counts and reports every piece of litter it passes may soon become a routine feature of the smart city.</p>
<p><strong>Subject of Research:</strong> Lightweight transformer-based semantic segmentation for real-time road garbage detection and cleanliness assessment on edge devices</p>
<p><strong>Article Title:</strong> Transformer-based lightweight model and category enhancement for road garbage segmentation</p>
<p><strong>Article References:</strong> Peng, S., Liao, J., Cao, L., Duan, C., &amp; Zhang, M. (2026). Transformer-based lightweight model and category enhancement for road garbage segmentation. <em>Neural Computing and Applications, 38</em>(19), Article 792. <a href="https://doi.org/10.1007/s00521-026-12554-6" rel="noopener noreferrer">https://doi.org/10.1007/s00521-026-12554-6</a></p>
<p><strong>Image Credits:</strong> AI Generated</p>
<p><strong>DOI:</strong> <a href="https://doi.org/10.1007/s00521-026-12554-6" rel="noopener noreferrer">10.1007/s00521-026-12554-6</a></p>
<p><strong>Keywords:</strong> transformer, semantic segmentation, road garbage detection, edge computing, lightweight model, SegFormer, attention mechanism, cleanliness assessment, computer vision, waste management, embedded deployment, deep learning</p>
]]></content:encoded>
					
		
		
		<post-id xmlns="com-wordpress:feed-additions:1">260494</post-id>	</item>
	</channel>
</rss>
