期刊文献+
共找到554篇文章
< 1 2 28 >
每页显示 20 50 100
RCTUnet:a deep learning model for crop-residue-soil image segmentation and crop residue cover extraction 认领 引用 被引量:1
1
作者 Ting LI Yang LIU +10 位作者 Haikuan FENG Meiyan SHU Hao YANG Yuanyuan FU Xin XU Yinghao LIN Hongbo QIAO Wei GUO Xinming MA Lei SHI Jibo YUE 《Journal of Zhejiang University-SCIENCE B》 SCIE CAS CSCD 2026年第5期517-536,共20页
Accurate quantification of crop residue cover(CRC)is crucial for monitoring and evaluating conservation tillage practices,yet it poses a significant image segmentation challenge.The subtle visual distinctions between ... Accurate quantification of crop residue cover(CRC)is crucial for monitoring and evaluating conservation tillage practices,yet it poses a significant image segmentation challenge.The subtle visual distinctions between fragmented residue and soil,compounded by variable illumination and shadows in field imagery,often lead to poor segmentation performance.To overcome these limitations,we introduce RCTUnet,a novel deep learning architecture designed for robust crop-residue-soil segmentation and precise CRC estimation.RCTUnet’s architecture synergistically integrates three key components:(1)a ResNet50 backbone for deep,multi-scale feature extraction;(2)a convolutional block attention module(CBAM)to adaptively focus on salient residue features across both channel and spatial dimensions;and(3)a transformer-based global context fusion module(GCFM)to model long-range spatial dependencies,which is critical for interpreting heterogeneous residue patterns.We evaluated RCTUnet on a dataset of 1220 field-acquired images spanning four typical crop rotations.Experimental results show that,compared to traditional models:(1)RCTUnet achieves significantly higher crop-residue-soil segmentation accuracy than classic models including Unet,Unet++,DeepLabV3,segmentation network(SegNet),and fully convolutional network(FCN),with improvements of 3.24%,3.42%,4.88%,8.28%,and 6.05%in overall accuracy,respectively;(2)RCTUnet yields superior residue-soil segmentation performance,with increases in residue recall of 7.67%,7.37%,14.09%,27.05%,and 16.91%,respectively;(3)RCTUnet shows enhanced CRC estimation accuracy,achieving a root mean square error(RMSE)of 4.875,representing a 45.5%improvement over Unet(RMSE=8.941).These results demonstrate the efficacy of our hybrid approach,which combines deep hierarchical features,dual-domain attention,and global context modeling.RCTUnet provides a robust and reliable tool for automated CRC assessment,advancing the capabilities of in-field agricultural monitoring. 展开更多
关键词 Deep learning Crop residue cover Image segmentation Conservation tillage
暂未订购 下载PDF
FMTNet:A Fourier-Mamba–Transformer Enhanced Network for Medical Image Segmentation 认领 引用
2
作者 Shaoqiang Wang Guiling Shi +5 位作者 Yuanyuan Zhang Sibo Qiao Yuchen Wang Yifan Wang Yawu Zhao Xiaochun Cheng 《CAAI Transactions on Intelligence Technology》 SCIE EI CSCD 2026年第3期798-815,共18页
Models based on U-shaped networks have achieved widespread success in the field of medical image segmentation,but their performance is generally limited by structural bottlenecks in the network.At this stage,feature m... Models based on U-shaped networks have achieved widespread success in the field of medical image segmentation,but their performance is generally limited by structural bottlenecks in the network.At this stage,feature maps experience a sharp decline in spatial resolution due to continuous downsampling,resulting in significant loss of critical boundaries and structural details.Additionally,the local receptive fields of convolutions limit the effective modelling of global context.To address this core issue,we propose a novel enhanced segmentation network called FMTNet.FMTNet fundamentally enhances the expressive power of deep features by integrating an innovative composite enhancement module at the bottleneck of the U-Net.This module consists of three synergistically working submodules:the Fourier spatial fusion module,which introduces a frequency-domain perspective to compensate for and reconstruct high-frequency structural information lost in the spatial domain;the hybrid mamba–transformer module,which efficiently captures cross-regional long-range dependencies to establish global context and the multi-scale context Aggregation module,which fuses features of different scales to adapt to objects of varying sizes.We conducted extensive experiments on multiple public multi-modal datasets,including colonoscopy polyps,dermatoscopy lesions,breast ultrasound and dental X-rays.The results demonstrate that FMTNet comprehensively outperforms SOTA methods across all key metrics,showcasing exceptional segmentation accuracy and generalisation capabilities.Our research study demonstrates that by synergistically enhancing deep features across three dimensions—frequency,global,and multi-scale—FMTNet provides a general and efficient solution to address the bottleneck issues of U-Net,significantly enhancing the accuracy and robustness of medical image segmentation.The source code and pre-trained weights are available at http://gffzz188fe103f8f1460asqkpuxpv06uqv6xxc.ffgz.tsg.suse.edu.cn/shiguiling0-has/FMTNet. 展开更多
关键词 bottleneck enhancement Fourier transform medical image segmentation multi‐scale feature fusion
暂未订购 下载PDF
A Novel Semi-Supervised Multi-View Picture Fuzzy Clustering Approach for Enhanced Satellite Image Segmentation 认领 引用
3
作者 Pham Huy Thong Hoang Thi Canh +2 位作者 Nguyen Tuan Huy Nguyen Long Giang Luong Thi Hong Lan 《Computers, Materials & Continua》 SCIE EI 2026年第3期1092-1117,共26页
Satellite image segmentation plays a crucial role in remote sensing,supporting applications such as environmental monitoring,land use analysis,and disaster management.However,traditional segmentation methods often rel... Satellite image segmentation plays a crucial role in remote sensing,supporting applications such as environmental monitoring,land use analysis,and disaster management.However,traditional segmentation methods often rely on large amounts of labeled data,which are costly and time-consuming to obtain,especially in largescale or dynamic environments.To address this challenge,we propose the Semi-Supervised Multi-View Picture Fuzzy Clustering(SS-MPFC)algorithm,which improves segmentation accuracy and robustness,particularly in complex and uncertain remote sensing scenarios.SS-MPFC unifies three paradigms:semi-supervised learning,multi-view clustering,and picture fuzzy set theory.This integration allows the model to effectively utilize a small number of labeled samples,fuse complementary information from multiple data views,and handle the ambiguity and uncertainty inherent in satellite imagery.We design a novel objective function that jointly incorporates picture fuzzy membership functions across multiple views of the data,and embeds pairwise semi-supervised constraints(must-link and cannot-link)directly into the clustering process to enhance segmentation accuracy.Experiments conducted on several benchmark satellite datasets demonstrate that SS-MPFC significantly outperforms existing state-of-the-art methods in segmentation accuracy,noise robustness,and semantic interpretability.On the Augsburg dataset,SS-MPFC achieves a Purity of 0.8158 and an Accuracy of 0.6860,highlighting its outstanding robustness and efficiency.These results demonstrate that SSMPFC offers a scalable and effective solution for real-world satellite-based monitoring systems,particularly in scenarios where rapid annotation is infeasible,such as wildfire tracking,agricultural monitoring,and dynamic urban mapping. 展开更多
关键词 Multi-view clustering satellite image segmentation semi-supervised learning picture fuzzy sets remote sensing
暂未订购 下载PDF
An APO Algorithm Based on Taguchi Methods and Its Application in Multi-Level Image Segmentation 认领 引用
4
作者 Jeng-Shyang Pan Yan-Na Wei +3 位作者 Ling-Da Chi Shu-Chuan Chu Ru-Yu Wang Junzo Watada 《Computers, Materials & Continua》 SCIE EI 2026年第5期814-837,共24页
Multilevel image segmentation is a critical task in image analysis,which imposes high requirements on the global search capability and convergence efficiency of segmentation algorithms.In this paper,an improved Artifi... Multilevel image segmentation is a critical task in image analysis,which imposes high requirements on the global search capability and convergence efficiency of segmentation algorithms.In this paper,an improved Artificial Protozoa Optimization algorithm,termed the two-stage Taguchi-assisted Gaussian–Levy Artificial Protozoa Optimization(TGAPO)algorithm,is proposed and applied tomultilevel image segmentation.The proposed algorithm adopts a two-stage evolutionary mechanism.In the first stage,Gaussian perturbation is introduced to enhance local search capability;in the second stage,Levy flight is incorporated to expand the global search range;and finally,the Taguchi strategy is employed to further refine the optimal solution.Consequently,the global optimization performance and robustness of the algorithm are significantly improved.To evaluate the effectiveness of the proposed TGAPO algorithm,comparative experiments are conducted with representative optimization algorithms,including the Grey Wolf Optimizer(GWO)and Particle Swarm Optimization(PSO),in the context ofmultilevel image segmentation.The segmentation quality is assessed using the minimum cross-entropy function as the performance metric.Experimental results demonstrate that the TGAPO algorithm outperforms the comparison algorithms in terms of segmentation accuracy and convergence speed,and exhibits superior stability in high-threshold segmentation tasks.Furthermore,the proposedmethod achieves excellentmulti-threshold segmentation performance for color images and shows strong potential for practical applications. 展开更多
关键词 Meta-heuristic algorithm multilevel image segmentation taguchi strategy minimum cross-entropy threshold artificial protozoa optimization(APO)
暂未订购 下载PDF
Advances in deep learning for bacterial image segmentation in optical microscopy 认领 引用
5
作者 Zhijun Tan Yang Ding +6 位作者 Huibin Ma Jintao Li Danrou Zheng Hua Bai Weini Xin Lin Li Bo Peng 《Journal of Innovative Optical Health Sciences》 SCIE EI CSCD 2026年第1期30-44,共15页
Microscopy imaging is fundamental in analyzing bacterial morphology and dynamics,offering critical insights into bacterial physiology and pathogenicity.Image segmentation techniques enable quantitative analysis of bac... Microscopy imaging is fundamental in analyzing bacterial morphology and dynamics,offering critical insights into bacterial physiology and pathogenicity.Image segmentation techniques enable quantitative analysis of bacterial structures,facilitating precise measurement of morphological variations and population behaviors at single-cell resolution.This paper reviews advancements in bacterial image segmentation,emphasizing the shift from traditional thresholding and watershed methods to deep learning-driven approaches.Convolutional neural networks(CNNs),U-Net architectures,and three-dimensional(3D)frameworks excel at segmenting dense biofilms and resolving antibiotic-induced morphological changes.These methods combine automated feature extraction with physics-informed postprocessing.Despite progress,challenges persist in computational efficiency,cross-species generalizability,and integration with multimodal experimental workflows.Future progress will depend on improving model robustness across species and imaging modalities,integrating multimodal data for phenotype-function mapping,and developing standard pipelines that link computational tools with clinical diagnostics.These innovations will expand microbial phenotyping beyond structural analysis,enabling deeper insights into bacterial physiology and ecological interactions. 展开更多
关键词 Bacterial image deep learning optical microscopy image segmentation artificial intelligence
暂未订购 下载PDF
DSSeg-FLHA: A Decentralized Secure Self-Adapting Image Segmentation Framework Using Federated Learning and Hybrid Architectures 认领 引用
6
作者 Rifat Sarker Aoyon Fahmid Al Farid +3 位作者 Ismail Hossain Mahe Zabin Sarina Mansor Jia Uddin 《Computers, Materials & Continua》 SCIE EI 2026年第8期1794-1818,共25页
This research introduces an innovative lightweight image segmentation framework where models of hybrid architectures work together to predict the output and also have self-adapting ability,along with maintaining data ... This research introduces an innovative lightweight image segmentation framework where models of hybrid architectures work together to predict the output and also have self-adapting ability,along with maintaining data privacy.In this framework,data is distributed and trained in a decentralized way using different deep learning architectures.That is how the advantages of all these models will be integrated into the system.Each trained model makes its own prediction,and the final output is determined through cooperation among these models.Here,the confidence-level and pixel-wise voting majority algorithms will be utilized for the co-operation-based output prediction system.Due to the efficient setup of the operations of these two algorithms,each input will get its accurate output.Additionally,the federated learning-based self-adapting feature facilitated the proposed framework for advancing its performance consistently by interacting with the inputs.Here,UNet,SegNet and FCNN models have been trained and integrated into the prediction framework.Here,the Oxford-IIT pet dataset was used.And all the data of this dataset is distributed among these three models.The framework’s effectiveness was measured using metrics like average pixel accuracy,IoU,F1 score,precision,and recall,which resulted in scores of 89.26%,71.48%,81.29%,83.96%and 81.16%,respectively.Another notable feature of this proposed framework is allocating comparatively fewer computational resources and taking less time.To validate these claims,the proposed system is compared with three other state-of-the-art models,and the proposed system delivered superior performance among all. 展开更多
关键词 Image segmentation decentralized federated learning UNet SegNet FCNN data privacy hybrid architecture self-adapting
暂未订购 下载PDF
A Control-based Transition Reinforced Optimization Process for Multi-level Threshold Image Segmentation 认领 引用
7
作者 Wei Wang Peiying Zhang +5 位作者 Saleh Ali Alomari Raed Abu Zitar Aseel Smerat Mohamed Sharaf Absalom E.Ezugwu Laith Abualigah 《Journal of Bionic Engineering》 SCIE EI CSCD 2026年第2期1061-1087,共27页
In this study,we present a novel approach to multi-threshold image segmentation using an adaptive method that combines the Ebola Optimization Search Algorithm(EOSA)with the Aquila Optimizer,termed the Integrated Enhan... In this study,we present a novel approach to multi-threshold image segmentation using an adaptive method that combines the Ebola Optimization Search Algorithm(EOSA)with the Aquila Optimizer,termed the Integrated Enhanced Ebola Optimization Search Algorithm(IEOSA).Our approach leverages this integration to produce high-quality segmented images.The IEOSA method introduces two distinct optimization mechanisms to identify optimal solutions.By blending the randomness of the Aquila Optimizer with the capabilities of EOSA,we enhance the exploration potential of the algorithm.Additionally,we incorporate a self-transition learning system within the IEOSA to further boost its performance.To tackle multi-level threshold image segmentation,we apply Kapur’s entropy between-class variance within the IEOSA framework.Our findings show that the IEOSA-based techniques outperform other comparable methods,offering faster convergence and more stable segmentation results.Through comparative analysis using standard test images,we demonstrate that IEOSA achieves higher solution accuracy than other methods.Ultimately,the proposed IEOSA methodologies effectively address multi-level threshold image segmentation challenges,accurately segmenting even the minor errors that are often overlooked in high-resolution images. 展开更多
关键词 Ebola optimization search algorithm(EOSA) Aquila optimizer(AO) Multi-level threshold Image segmentation Transition mechanism
RE-UKAN:A Medical Image Segmentation Network Based on Residual Network and Efficient Local Attention 认领 引用
8
作者 Bo Li Jie Jia +2 位作者 Peiwen Tan Xinyan Chen Dongjin Li 《Computers, Materials & Continua》 SCIE EI 2026年第3期2184-2200,共17页
Medical image segmentation is of critical importance in the domain of contemporary medical imaging.However,U-Net and its variants exhibit limitations in capturing complex nonlinear patterns and global contextual infor... Medical image segmentation is of critical importance in the domain of contemporary medical imaging.However,U-Net and its variants exhibit limitations in capturing complex nonlinear patterns and global contextual information.Although the subsequent U-KAN model enhances nonlinear representation capabilities,it still faces challenges such as gradient vanishing during deep network training and spatial detail loss during feature downsampling,resulting in insufficient segmentation accuracy for edge structures and minute lesions.To address these challenges,this paper proposes the RE-UKAN model,which innovatively improves upon U-KAN.Firstly,a residual network is introduced into the encoder to effectively mitigate gradient vanishing through cross-layer identity mappings,thus enhancing modelling capabilities for complex pathological structures.Secondly,Efficient Local Attention(ELA)is integrated to suppress spatial detail loss during downsampling,thereby improving the perception of edge structures and minute lesions.Experimental results on four public datasets demonstrate that RE-UKAN outperforms existing medical image segmentation methods across multiple evaluation metrics,with particularly outstanding performance on the TN-SCUI 2020 dataset,achieving IoU of 88.18%and Dice of 93.57%.Compared to the baseline model,it achieves improvements of 3.05%and 1.72%,respectively.These results fully demonstrate RE-UKAN’s superior detail retention capability and boundary recognition accuracy in complex medical image segmentation tasks,providing a reliable solution for clinical precision segmentation. 展开更多
关键词 Image segmentation U-KAN residual network ELA
暂未订购 下载PDF
Importance-Aware Image Segmentation-Based Semantic Communication for Autonomous Driving 认领 引用
9
作者 Lyu Jie Tong Haonan +4 位作者 Pan Qiang Zhang Zhilong He Xinxin Luo Tao Yin Changchuan 《China Communications》 SCIE EI CSCD 2026年第2期228-243,共16页
This article studies the problem of image segmentation-based semantic communication in autonomous driving.In real traffic scenes,the detecting of objects(e.g.,vehicles and pedestrians)is more important to guarantee dr... This article studies the problem of image segmentation-based semantic communication in autonomous driving.In real traffic scenes,the detecting of objects(e.g.,vehicles and pedestrians)is more important to guarantee driving safety,which is always ignored in existing works.Therefore,we propose a vehicular image segmentation-oriented semantic communication system,termed VIS-SemCom,focusing on transmitting and recovering image semantic features of high-important objects to reduce transmission redundancy.First,we develop a semantic codec based on Swin Transformer architecture,which expands the perceptual field thus improving the segmentation accuracy.To highlight the important objects'accuracy,we propose a multi-scale semantic extraction method by assigning the number of Swin Transformer blocks for diverse resolution semantic features.Also,an importance-aware loss incorporating important levels is devised,and an online hard example mining(OHEM)strategy is proposed to handle small sample issues in the dataset.Finally,experimental results demonstrate that the proposed VIS-SemCom can achieve a significant mean intersection over union(mIoU)performance in the SNR regions,a reduction of transmitted data volume by about 60%at 60%mIoU,and improve the segmentation accuracy of important objects,compared to baseline image communication. 展开更多
关键词 autonomous driving image segmentation semantic communication Swin Transformer
暂未订购 下载PDF
DenT:Dense-Transformer for Label-Free Microscopy Image Segmentation 认领 引用
10
作者 Chan-Min Hsu Shang-Ru Yang +1 位作者 Yi-Ju Lee An-Chi Wei 《Computers, Materials & Continua》 SCIE EI 2026年第7期278-293,共16页
U-Net,a fully convolutional neural network(FCNN)with U-shaped features,has demonstrated significant success in biomedical image segmentation.However,the locality of convolution operations in the U-Net limits its abili... U-Net,a fully convolutional neural network(FCNN)with U-shaped features,has demonstrated significant success in biomedical image segmentation.However,the locality of convolution operations in the U-Net limits its ability to learn long-range dependencies.Transformers,originally developed for natural language processing,have recently been adapted for image segmentation because of their global self-attention mechanisms.Inspired by the long-range feature learning capability of transformers,we propose Dense-Transformer(DenT),an architecture designed for volumetric microscopy image segmentation.DenT incorporates transformers as encoders within each convolutional layer to capture global contextual information.Additionally,dense skip connections at multiple resolutions enhance feature propagation,enabling precise localization.We evaluated DenT on mitochondrial segmentation using our confocal microscopy dataset and a public fluorescence microscope dataset from the Allen Institute for Cell Science.The experimental results demonstrate that DenT incrementally improves the segmentation of mitochondria and mitochondrial DNA substructures from transmitted light microscopy images.DenT offers a tool for visualization,measurement,and analysis of mitochondrial morphology and mitochondrial DNA in label-free microscopy. 展开更多
关键词 UNet transformer image segmentation mitochondrial organelle deep learning microscopy imaging
暂未订购 下载PDF
Multiple PointMedSAM Prompting for Enhanced Medical Image Segmentation 认领 引用
11
作者 Wasfieh Nazzal Ezequiel López-Rubio +1 位作者 Miguel A.Molina-Cabello Karl Thurnhofer-Hemsi 《Computers, Materials & Continua》 SCIE EI 2026年第5期2100-2115,共16页
Automatic and accurate medical image segmentation remains a fundamental task in computer-aided diagnosis and treatment planning.Recent advances in foundation models,such as the medical-focused Segment AnythingModel(Me... Automatic and accurate medical image segmentation remains a fundamental task in computer-aided diagnosis and treatment planning.Recent advances in foundation models,such as the medical-focused Segment AnythingModel(MedSAM),have demonstrated strong performance but face challenges inmanymedical applications due to anatomical complexity and a limited domain-specific prompt.Thiswork introduces amethodology that enhances segmentation robustness and precision by automatically generating multiple informative point prompts,rather than relying on single inputs.The proposed approach randomly samples sets of spatially distributed point prompts based on image features,enabling MedSAM to better capture fine-grained anatomical structures and boundaries.During inference,probability maps are aggregated to reduce local misclassifications without additional model training.Extensive experiments on various computed tomography(CT)and magnetic resonance imaging(MRI)datasets demonstrate improvements in Dice Similarity Coefficient(DSC)and Normalized Surface Dice(NSD)metrics compared to baseline SAM and Scribble Prompt models.A semi-automatic point sampling version based on the ground truth segmentations yielded enhanced results,achieving up to 92.1%DSC and 86.6%NSD,with significant gains in delineating complex organs such as the pancreas,colon,kidney,and brain tumours.The main novelty of our method consists of effectively combining the results of multiple point prompts into the medical segmentation pipeline so that single-point prompt methods are outperformed.Overall,the proposed model offers a straightforward yet effective approach to improve medical image segmentation performance while maintaining computational efficiency. 展开更多
关键词 Medical image segmentation deep learning test-time augmentation point prompt
暂未订购 下载PDF
TriLVM-UNet:Multi-Scale State Space Modeling with Cross-Channel Fusion Attention Mechanism for Precise Medical Image Segmentation 认领 引用
12
作者 Kexin Zhang Lihua Liu +3 位作者 Yuting Xue Tao Zhou Fengshuai Yue Ruifeng Du 《Computers, Materials & Continua》 SCIE EI 2026年第9期2223-2252,共30页
Traditional Mamba-UNet integrations employ four-stage architectures,replacing conventional five-stage UNets with VMamba blocks for global dependency modeling.Unlike Transformers,which suffer from quadratic complexity ... Traditional Mamba-UNet integrations employ four-stage architectures,replacing conventional five-stage UNets with VMamba blocks for global dependency modeling.Unlike Transformers,which suffer from quadratic complexity and high memory consumption in self-attention,Mamba-UNet achieves efficient global modeling through linear-complexity state space modeling.This paper proposes TriLVM-UNet,a lightweight three-stage architecture that integrates parameter-efficient VMamba blocks and enhances cross-stage feature interaction via an improved skip-attention bridge(SAB)module inspired by UltraLight VM-UNet.The model incorporates a Lightweight Vision Mamba(LVM)layer for high-resolution feature extraction,alongside multi-scale dilated convolution(MSDC)and convolutional block attention module(CBAM)for enhanced feature fusion.Evaluated on the 3D ACDC dataset against six baseline models,TriLVM-UNet achieves 98.57%accuracy.The GitHub repository is available at:http://gffzz188fe103f8f1460asqkpuxpv06uqv6xxc.ffgz.tsg.suse.edu.cn/730432ch/TriLVM-UNet. 展开更多
关键词 Medical image segmentation visual state space model TriLVM-UNet lightweight architecture
暂未订购 下载PDF
An Adaptive Hybrid SCSOWOA Algorithm for Generalized Multi-level Thresholding in Multi-organ Medical Image Segmentation 认领 引用
13
作者 M.Faruk Sahin Farzad Kiani 《Journal of Bionic Engineering》 SCIE EI CSCD 2026年第2期1240-1262,共23页
This study presents a novel hybrid optimization model that combines the complementary aspects of Sand Cat Swarm Optimization(SCSO)and Whale Optimization Algorithm(WOA)to solve the multi-level image thresholding proble... This study presents a novel hybrid optimization model that combines the complementary aspects of Sand Cat Swarm Optimization(SCSO)and Whale Optimization Algorithm(WOA)to solve the multi-level image thresholding problem.The proposed approach utilizes an adaptive two-stage mechanism that balances the high exploration capacity of SCSO with the concentrated local search capability of WOA,aiming to maximize inter-class variance in the histogram-based thresholding process.Various experiments are conducted on lung cancer,prostate,and mixed medical image datasets.Results demonstrate that modified SCSOWOA achieves superior performance across all datasets.For LC25000,it attains PSNR 27.9453 dB,SSIM 0.9340,FSIM 0.9542,Dice coefficient 0.8901,and Jaccard index 0.8034 at T=12.For prostate,PSNR reaches 28.3965 dB,SSIM 0.7532,FSIM 0.8170,Dice 0.9215,and Jaccard 0.8593.In the MSD dataset,SCSOWOA achieves PSNR 29.3244 dB,SSIM 0.7118,FSIM 0.7562,Dice 0.8901,and Jaccard 0.8034,indicating consistent performance across diverse organs and modalities.The method also demonstrates high computational efficiency,with an average execution time of 1.3221 s,offering up to 40%speed improvement over conventional metaheuristics such as PSO and GWO.Overall,proposed method provides high-accuracy,low-variance,and computationally efficient segmentation,preserving both structural and perceptual fidelity.These results confirm the method’s robustness,generalizability,and practical applicability for AI-assisted diagnostic systems across histopathological and medical imaging contexts,balancing precision,structural preservation,and speed for real-world clinical deployment. 展开更多
关键词 Hybrid meta-heuristics Image processing Multi-level image thresholding Medical image segmentation
MS-RWKV-UNet:Multi-Head Scan Receptance Weighted Key Value UNet for Medical Image Segmentation 认领 引用
14
作者 JIANG Dong JI Zhongping FANG Meie 《Wuhan University Journal of Natural Sciences》 CAS CSCD 2026年第1期1-9,共9页
The Transformer has achieved great success in the field of medical image segmentation,but its quadratic computational complexity limits its application in dense medical image prediction.Recently,the receptance weighte... The Transformer has achieved great success in the field of medical image segmentation,but its quadratic computational complexity limits its application in dense medical image prediction.Recently,the receptance weighted key value(RWKV)architecture has garnered widespread attention due to its linear computational complexity and its capability of parallel computation during training.Despite the RWKV model's proficiency in addressing long-range modeling tasks with linear computational complexity,most current RWKV-based approaches employ static scanning patterns.These patterns may inadvertently incorporate biased prior knowledge into the model's predictions.To address this challenge,we propose a multi-head scan strategy combined with padding methods to effectively simulate spatial continuity in 2D images.Within the Feature Aggregation Attention(FAA)module,asymmetric convolutions are designed to aggregate 1D sequence features along a single dimension,thereby expanding effective receptive fields while preserving structural sparsity.Additionally,panoramic token shift(P-Shift)effectively models local dependency relationships by moving tokens from a wide receptive field.Extensive experiments conducted on the ISIC17/18 and ACDC datasets demonstrate that our method exhibits superior performance in dense medical image prediction tasks. 展开更多
关键词 multi-head scan receptance weighted key value(RWKV) asymmetric convolution panoramic token shift(P-Shift) medical image segmentation
暂未订购 下载PDF
TransUNet framework improved computed tomography image segmentation for core pore evolution 认领 引用
15
作者 Shao-Hua Zhou Tian-Bao Liu +6 位作者 Yu Zhao Shao-Hao Yin Zhi-Yu Liu Ji-Jun Liu Ling-Wei Du Yue-Tong Zhao Wei-Guang Shi 《Petroleum Science》 SCIE EI CAS CSCD 2026年第6期3682-3697,共16页
The development of oil and gas is constrained by difficulties in dynamically characterizing pore structures.Traditional methods inadequately represent the complex interactions between mineral dissolution,precipitation... The development of oil and gas is constrained by difficulties in dynamically characterizing pore structures.Traditional methods inadequately represent the complex interactions between mineral dissolution,precipitation,and fluid flow.This study addresses these gaps by introducing a Transformer U-Neural Network(TransUNet)for computed tomography(CT)image segmentation.The integrated workflow combines conventional CT(Resolution of 5.4μm)and synchrotron radiation CT(Resolution of0.8μm)for dynamic flooding,imaging,segmentation,and precise 3D pore network extraction,overcoming resolution limits.TransUNet's strong global attention and feature extraction reduce overfitting and deliver high-accuracy segmentation of minerals,pores,and argillaceous microporous networks(AMN),achieving 74.92%intersection over union(IoU)for AMN.A porosity correction method improves conventional CT porosity accuracy to 94%of gas-measured values.Alkaline flooding experiments reveal:(1)initial clay swelling reduces small pore size by~50%as alkaline ions destabilize clay;(2)mineral dissolution,such as dolomite,creates secondary pores,increasing 80μm pores by 1.8 times;(3)silicate dissolution increases porosity and leads to a 93.7%rise in permeability.Clay reorganization enhances the AMN by 46.1%.The pore size distribution shifts to log-no rmal at steady state,and throat connectivity improves flow capacity.This work pioneers Transformer-based CT image segmentation,introduces cross-resolution prediction,and clarifies pore regulation by mineral phase changes,establishing a new paradigm for chemical flooding in sandstone reservoirs. 展开更多
关键词 Synchrotron radiation CT Image segmentation TransUNet Argillaceous microporous network Alkaline flooding porosity correction
暂未订购 下载PDF
Pixel to Parcel:Transformative Applications of Image Segmentation in Geospatial and Crop Research 认领 引用
16
作者 Hui Zeng 《Journal of Environmental & Earth Sciences》 CAS 2026年第3期112-125,共14页
The rising need for precision farming and sustainable land management has catalyzed the requirement for sophisticated means of deriving practical data from remote sensing images.Image segmentation,or the process of di... The rising need for precision farming and sustainable land management has catalyzed the requirement for sophisticated means of deriving practical data from remote sensing images.Image segmentation,or the process of dividing the image into semantically relevant parts,has become a groundbreaking technology that allows resolving the problem of transitioning the pixel-level data to a parcel-level analysis.This review is a synthesis of the segmentation methods and their use in crop research and geospatial science.The architectures of pixel-based,object-based,and deep learning(convolutional neural networks,U-Net,Mask R-CNN,and Transformer models)are considered in terms of principles,capabilities,and limitations.Multi-spectral,hyperspectral,LiDAR,and SAR data are integrated to improve the efficiency of segmentation,allowing the possible delineation of fields,the classification of crops,health monitoring,monitoring of yields,and stress identification.In addition to agriculture,segmentation helps in land use and land cover mapping,identification of temporal change,monitoring of the environment,and is used in combination with GIS-based spatial modeling.Nevertheless,issues related to data heterogeneity,mixed pixels,computational requirements,and inadequate availability of labelled data still exist despite the major progress.The future directions involve multi-source data fusion,pixel-to-parcel pipeline automation,and predictive models based on AI,which are used to enhance its scalability,robustness,and the ability to monitor in real-time.This review makes it clear that the use of image segmentation as a tool in generating precision agriculture,sustainable land use,and informed geospatial. 展开更多
关键词 Image Segmentation Precision Agriculture Geospatial Analysis Crop Monitoring Remote Sensing
暂未订购 下载PDF
DGFE-Mamba:Mamba-Based 2D Image Segmentation Network 认领 引用 被引量:2
17
作者 Junding Sun Kaixin Chen +4 位作者 Shuihua Wang Yudong Zhang Zhaozhao Xu Xiaosheng Wu Chaosheng Tang 《Journal of Bionic Engineering》 SCIE EI CSCD 2025年第4期2135-2150,共16页
In the field of medical image processing,combining global and local relationship modeling constitutes an effective strategy for precise segmentation.Prior research has established the validity of Convolutional Neural ... In the field of medical image processing,combining global and local relationship modeling constitutes an effective strategy for precise segmentation.Prior research has established the validity of Convolutional Neural Networks(CNN)in modeling local relationships.Conversely,Transformers have demonstrated their capability to effectively capture global contextual information.However,when utilized to address CNNs’limitations in modeling global relationships,Transformers are hindered by substantial computational complexity.To address this issue,we introduce Mamba,a State-Space Model(SSM)that exhibits exceptional proficiency in modeling long-range dependencies in sequential data.Given Mamba’s demonstrated potential in 2D medical image segmentation in previous studies,we have designed a Dual-encoder Global-local Feature Extraction Network based on Mamba,termed DGFE-Mamba,to accurately capture and fuse long-range dependencies and local dependencies within multi-scale features.Compared to Transformer-based methods,the DGFE-Mamba model excels in comprehensive feature modeling and demonstrates significantly improved segmentation accuracy.To validate the effectiveness and practicality of DGFE-Mamba,we conducted tests on the Automatic Cardiac Diagnosis Challenge(ACDC)dataset,the Synapse multi-organ CT abdominal segmentation dataset,and the Colorectal Cancer Clinic(CVC-ClinicDB)dataset.The results showed that DGFE-Mamba achieved Dice coefficients of 92.20,83.67,and 94.13,respectively.These findings comprehensively validate the effectiveness and practicality of the proposed DGFE-Mamba architecture. 展开更多
关键词 Medical image segmentation Mamba CNN Attention Mechanism
M2ANet:Multi-branch and multi-scale attention network for medical image segmentation 认领 引用 被引量:1
18
作者 Wei Xue Chuanghui Chen +3 位作者 Xuan Qi Jian Qin Zhen Tang Yongsheng He 《Chinese Physics B》 SCIE EI CAS CSCD 2025年第8期547-559,共13页
Convolutional neural networks(CNNs)-based medical image segmentation technologies have been widely used in medical image segmentation because of their strong representation and generalization abilities.However,due to ... Convolutional neural networks(CNNs)-based medical image segmentation technologies have been widely used in medical image segmentation because of their strong representation and generalization abilities.However,due to the inability to effectively capture global information from images,CNNs can easily lead to loss of contours and textures in segmentation results.Notice that the transformer model can effectively capture the properties of long-range dependencies in the image,and furthermore,combining the CNN and the transformer can effectively extract local details and global contextual features of the image.Motivated by this,we propose a multi-branch and multi-scale attention network(M2ANet)for medical image segmentation,whose architecture consists of three components.Specifically,in the first component,we construct an adaptive multi-branch patch module for parallel extraction of image features to reduce information loss caused by downsampling.In the second component,we apply residual block to the well-known convolutional block attention module to enhance the network’s ability to recognize important features of images and alleviate the phenomenon of gradient vanishing.In the third component,we design a multi-scale feature fusion module,in which we adopt adaptive average pooling and position encoding to enhance contextual features,and then multi-head attention is introduced to further enrich feature representation.Finally,we validate the effectiveness and feasibility of the proposed M2ANet method through comparative experiments on four benchmark medical image segmentation datasets,particularly in the context of preserving contours and textures. 展开更多
关键词 medical image segmentation convolutional neural network multi-branch attention multi-scale feature fusion
暂未订购 下载PDF
Multi-Consistency Training for Semi-Supervised Medical Image Segmentation 认领 引用 被引量:1
19
作者 WU Changxue ZHANG Wenxi +1 位作者 HAN Jiaozhi WANG Hongyu 《Journal of Shanghai Jiaotong university(Science)》 EI 2025年第4期800-814,共15页
Medical image segmentation is a crucial task in clinical applications.However,obtaining labeled data for medical images is often challenging.This has led to the appeal of semi-supervised learning(SSL),a technique adep... Medical image segmentation is a crucial task in clinical applications.However,obtaining labeled data for medical images is often challenging.This has led to the appeal of semi-supervised learning(SSL),a technique adept at leveraging a modest amount of labeled data.Nonetheless,most prevailing SSL segmentation methods for medical images either rely on the single consistency training method or directly fine-tune SSL methods designed for natural images.In this paper,we propose an innovative semi-supervised method called multi-consistency training(MCT)for medical image segmentation.Our approach transcends the constraints of prior methodologies by considering consistency from a dual perspective:output consistency across different up-sampling methods and output consistency of the same data within the same network under various perturbations to the intermediate features.We design distinct semi-supervised loss regression methods for these two types of consistencies.To enhance the application of our MCT model,we also develop a dedicated decoder as the core of our neural network.Thorough experiments were conducted on the polyp dataset and the dental dataset,rigorously compared against other SSL methods.Experimental results demonstrate the superiority of our approach,achieving higher segmentation accuracy.Moreover,comprehensive ablation studies and insightful discussion substantiate the efficacy of our approach in navigating the intricacies of medical image segmentation. 展开更多
关键词 semi-supervised learning(SSL) multi-consistency training(MCT) medical image segmentation intermediate feature perturbation
暂未订购 下载PDF
Stochastic Augmented-Based Dual-Teaching for Semi-Supervised Medical Image Segmentation 认领 引用
20
作者 Hengyang Liu Yang Yuan +2 位作者 Pengcheng Ren Chengyun Song Fen Luo 《Computers, Materials & Continua》 SCIE EI 2025年第1期543-560,共18页
Existing semi-supervisedmedical image segmentation algorithms use copy-paste data augmentation to correct the labeled-unlabeled data distribution mismatch.However,current copy-paste methods have three limitations:(1)t... Existing semi-supervisedmedical image segmentation algorithms use copy-paste data augmentation to correct the labeled-unlabeled data distribution mismatch.However,current copy-paste methods have three limitations:(1)training the model solely with copy-paste mixed pictures from labeled and unlabeled input loses a lot of labeled information;(2)low-quality pseudo-labels can cause confirmation bias in pseudo-supervised learning on unlabeled data;(3)the segmentation performance in low-contrast and local regions is less than optimal.We design a Stochastic Augmentation-Based Dual-Teaching Auxiliary Training Strategy(SADT),which enhances feature diversity and learns high-quality features to overcome these problems.To be more precise,SADT trains the Student Network by using pseudo-label-based training from Teacher Network 1 and supervised learning with labeled data,which prevents the loss of rare labeled data.We introduce a bi-directional copy-pastemask with progressive high-entropy filtering to reduce data distribution disparities and mitigate confirmation bias in pseudo-supervision.For the mixed images,Deep-Shallow Spatial Contrastive Learning(DSSCL)is proposed in the feature spaces of Teacher Network 2 and the Student Network to improve the segmentation capabilities in low-contrast and local areas.In this procedure,the features retrieved by the Student Network are subjected to a random feature perturbation technique.On two openly available datasets,extensive trials show that our proposed SADT performs much better than the state-ofthe-art semi-supervised medical segmentation techniques.Using only 10%of the labeled data for training,SADT was able to acquire a Dice score of 90.10%on the ACDC(Automatic Cardiac Diagnosis Challenge)dataset. 展开更多
关键词 Semi-supervised medical image segmentation contrastive learning stochastic augmented
暂未订购 下载PDF
上一页 1 2 28 下一页 到第
在线咨询 使用帮助 返回顶部 意见反馈