Computer Vision and Multimedia Computation articles for China

Time frame: 1 May 2025 - 30 April 2026
Count: 151

Article ‘Count’ for Computer Vision and Multimedia Computation.

Journal Count Share
2 2.00
A Biomimetic Fisheye for High-Accuracy Underwater Recognition Enabled by Oxide Semiconductor Retina 1.00
Deep Learning Enables Pixel-Level Nanoparticle Distribution Mapping in Routine Histological Sections by Integrating Cancer Associated Fibroblasts Features 1.00
2 2.00
Falcon Vision-Inspired Ultrafast Traffic Obstacle Avoidance Based on 2D Edge-Rich van der Waals Heterostructures 1.00
High‐Fidelity and Low‐Latency Machine Vision via Manipulation of Contact Potential Behaviors at Perovskite Hetero‐Interfaces 1.00
4 3.89
YOLO-Drop: A Deep Learning Model Enabling Accurate, High-Throughput Image Analysis for Droplet Digital Immunoassay at Attomolar Concentrations 1.00
MSInet: A Self-Supervised CNN Framework Integrating Global and Local Context for Robust Mass Spectrometry Imaging Segmentation 0.89
UPBAS-Net: An Upsampling-Powered Boundary-Aware Segmentation Network for Fluorescent Spots in Microscopy Images 1.00
GA-HIDMS-PSO-BPNN Model-Based Suppression of Cross-Interference in Absorption Spectra for Dual-Gas Sensing 1.00
1 1.00
Differentially enhanced parallel rotating neuron reservoir computing for dynamic gas mixture identification 1.00
1 1.00
Towards large-scale chemical reaction image parsing via a multimodal large language model 1.00
8 7.60
Comprehensive Temperature Perception for Battery Energy Storage Cluster with Sparse Sensing and Mixture-of-Experts 1.00
A multi-scale photovoltaic (PV) panel dust accumulation simulation dataset based on physical consistency modeling 1.00
Online optimization and dynamic prediction of an intermediate-circuit thermal management system based on physical-data dual-driven modeling and experimental validation 1.00
A novel dual-knowledge embedded graph convolutional network method combining knowledge quantification for gas turbine gas path analysis 1.00
Monitoring Scour Pits around Offshore Wind Turbine Foundations for Enhanced Maintenance 0.60
Machine learning-based simulation and experiment of a concentration sensor-free control strategy for direct methanol fuel cells 1.00
Prediction of NOx concentration based on interpretable convolutional gated recurrent unit with clustering-extracting features 1.00
A wearable driver gesture recognition system enabled AI application in virtual reality interaction for smart traffic 1.00
1 1.00
Polyomino reconstructs spatial transcriptomic profiles with single-cell resolution via a region-allocation method 1.00
123 118.33
EFFOcc: Learning Efficient Occupancy Networks from Minimal Labels for Autonomous Driving 1.00
PCMF2-Net: A Pyramid Cross-Modal Feature Fusion Network for Off-Road Freespace Detection 1.00
TCNet: A Temporally Consistent Network for Self-supervised Monocular Depth Estimation 1.00
SKT: Integrating State-Aware Keypoint Trajectories with Vision-Language Models for Robotic Garment Manipulation 0.89
Diff-IP2D: Diffusion-Based Hand-Object Interaction Prediction on Egocentric Videos 1.00
Improved Calibration for Panoramic Annular Lens Systems with Angular Modulation 1.00
Rejecting Outliers in 2D-3D Point Correspondences from 2D Forward-Looking Sonar Observations 1.00
EdgeSpotter: Multi-Scale Dense Text Spotting for Industrial Panel Monitoring 1.00
ETO+: Revisit the Refinement Stage in Efficient Feature Matching 1.00
CoDifFu: Diffusion-Based Collaborative Perception with Efficient Heterogeneous Feature Fusion 1.00
A Multi-modal Hand Imitation Dataset for Dexterous Hand 1.00
Multi-Cali Anything: Dense Feature Multi-Frame Structure-from-Motion for Large-Scale Camera Array Calibration 0.09
Improved 2D Hand Trajectory Prediction with Multi-View Consistency 1.00
STAGE: A Stream-Centric Generative World Model for Long-Horizon Driving-Scene Simulation 1.00
Dense Semantic Bird-Eye-View Map Generation from Sparse LiDAR Point Clouds via Distribution-aware Feature Fusion 0.67
Layer Decomposition and Morphological Reconstruction for Task-Oriented Infrared Image Enhancement 1.00
Noise Fusion-based Distillation Learning for Anomaly Detection in Complex Industrial Environments 1.00
Novel Diffusion Models for Multimodal 3D Hand Trajectory Prediction 0.83
PCGE: Boosting 3D Visual Grounding via Progressive Comprehension and Geometric-topology Perception Enhancement 1.00
Adaptive Visuo-Tactile Fusion with Predictive Force Attention for Dexterous Manipulation 1.00
UAV-MaLO: Mamba-Augmented YOLO Hybrid Architecture for UAV Micro-Object Detection in Autonomous Robotics 1.00
Edge-Guided Lighting Adaptation: Real-Time Detection of Transparent Objects for Cell Culture Robot 1.00
Rotation-Equivariant Robot Vision: A Perspective via Correspondence-Matching and Pre-training 1.00
BoRe-Depth: Self-Supervised Monocular Depth Estimation with Boundary Refinement for Embedded Systems 1.00
RAG-6DPose: Retrieval-Augmented 6D Pose Estimation via Leveraging CAD as Knowledge Base 0.86
FUSE: Label-Free Image-Event Joint Monocular Depth Estimation via Frequency-Decoupled Alignment and Degradation-Robust Fusion 1.00
Diffusion Suction Grasping with Large-Scale Parcel Dataset 1.00
Crouch Gait Recognition of Children with Cerebral Palsy Based on CNN-LSTM Hybrid Model 1.00
Efficient Instance Motion-Aware Point Cloud Scene Prediction 1.00
SOLO-SMap: Semantic-Aided Online LiDAR Odometry and 3D Static Mapping for Dynamic Scenes 1.00
Dynamicity Adaptation for Multi-object Tracking and Segmentation: Toward Improved Association Correction 1.00
LGDD: Local-Global Synergistic Dual-Branch 3D Object Detection Using 4D Radar 1.00
CA2Point: Learning Keypoint Detection and Description with Context Aggregation and Cross Augmentation 1.00
MambaNUT: Nighttime UAV Tracking via Mamba-based Adaptive Curriculum Learning 1.00
DPSN: Dual Prior Knowledge Induced Tactile paving and Obstacle Joint Segmentation Network 0.86
Embracing Dynamics: Dynamics-aware 4D Gaussian Splatting SLAM 1.00
Dynamic Action Localization and Recognition for Intelligent Perception of Surgical Robots 1.00
RoadsideSplat: Robust 3D Gaussian Reconstruction from Monocular Roadside Surveillance 1.00
SynthDrive: Scalable Real2Sim2Real Sensor Simulation Pipeline for High-Fidelity Asset Generation and Driving Data Synthesis 1.00
TIETracker: A CLIP-based RGB-T Tracking via Feature Interaction and Semantic Enhancement 1.00
SA-MVSNet: Spatial-aware Multi-view Stereo Network with Attention Cost Volume 1.00
Multi-target Association and Localization with Distributed Drone Following: A Factor Graph Approach 1.00
Anyview: General Indoor 3D Object Detection with Variable Frames 0.86
ThermalLoc: A Vision Transformer-Based Approach for Robust Thermal Camera Relocalization in Large-Scale Environments 1.00
Graph2Scene: Versatile 3D Indoor Scene Generation with Interaction-aware Scene Graph 1.00
BookBot: A Robotic Manipulation Benchmark for Voice-Driven Book Recognition and Grasping in Cluttered Environments 1.00
Reducing Redundancy in VSLAM: VLMs-driven Keyframe Selection using Multi-dimensional Semantic Information 1.00
Generalizable and Actionable Part Detection and Manipulation with SAM-rectified Segmentation and Iterative Pose Refinement 1.00
Cross-modal State Space Modeling for Real-time RGB-thermal Wild Scene Semantic Segmentation 0.92
ACP-MVS: Efficient Multi-View Stereo with Attention-based Context Perception 1.00
Zero-Shot Temporal Interaction Localization for Egocentric Videos 1.00
NeuroLoc: Encoding Navigation Cells for 6-DOF Camera Localization 1.00
Towards Physically Realizable Adversarial Attacks in Embodied Vision Navigation 1.00
Motion-Feat: Motion Blur-Aware Local Feature Description for Image Matching 0.71
IoU-Aware Clustering for Anchor Configuration Determination in Efficient Defect Detection 1.00
Wireless Collaborative Inference Acceleration Based on Distillation for Weed Detection and Instance Segmentation 1.00
MCTrack: A Unified 3D Multi-Object Tracking Framework for Autonomous Driving 0.92
LIM: A Low-Complexity Local Feature Image Matching Network for Real-Time Embedded Applications 1.00
UAV-DETR: Efficient End-to-End Object Detection for Unmanned Aerial Vehicle Imagery 1.00
Correspondence-Free Pose Estimation with Patterns: A Unified Approach for Multi-Dimensional Vision 1.00
ManipGPT: Is Affordance Segmentation by Large Vision Models Enough for Articulated Object Manipulation? 1.00
Reusing Attention for One-stage Lane Topology Understanding 0.96
HFDNet: High-Frequency Divergence Attention Network for Underwater Segmentation 1.00
Multi-View Normal and Distance Guidance Gaussian Splatting for Surface Reconstruction 1.00
MovSAM: A Single-image Moving Object Segmentation Framework Based on Deep Thinking 0.80
Efficient Multimodal 3D Object Detector via Instance-Level Contrastive Distillation 1.00
SimWorld: A Unified Benchmark for Simulator-Conditioned Scene Generation via World Model 1.00
ORA-NET: Enhancing Image Feature Matching through Oriented Overlapping Region Alignment 1.00
STG-Avatar: Animatable Human Avatars via Spacetime Gaussian 0.86
Unidirectional Point-Voxel Fusion for Enhanced 3D Single Object Tracking 1.00
Vision-Driven 2D Supervised Fine-Tuning Framework for Bird’s Eye View Perception 1.00
GSO-SLAM: Robust Monocular SLAM with Global Structure Optimization 1.00
Self-Supervised Enhancement for Depth from a Lightweight ToF Sensor with Monocular Images 1.00
MRMT-PR: A Multi-Scale Reverse-View Mamba-Transformer for LiDAR Place Recognition 0.83
AlignCAPE: Support and Query Feature Aligning for Category-Agnostic Pose Estimation 1.00
Recognizing Skeleton-Based Actions As Points 1.00
Overlap-Aware Feature Learning for Robust Unsupervised Domain Adaptation for 3D Semantic Segmentation 0.75
Unveiling the Potential of Segment Anything Model 2 for RGB-Thermal Semantic Segmentation with Language Guidance 1.00
HeightAware-BEV: Height-Aware Feature Mapping for Efficient Bird’s-Eye-View Perception 1.00
LiDAR-Inertial Odometry in Dynamic Driving Scenarios using Label Consistency Detection 1.00
BEVPointNet3D: Fusing Bird’s Eye View and Point Cloud Features for Robust 3D Lane Detection 1.00
GTAD: Global Temporal Aggregation Denoising Learning for 3D Semantic Occupancy Prediction 1.00
SAGENet: Binaural Echo-Based 3D Depth Estimation with Sparse Angular Queries and Refined Geometric Cues 1.00
Depth Matters: Exploring Deep Interactions of RGB-D for Semantic Segmentation in Traffic Scenes 1.00
KDMOS:Knowledge Distillation for Motion Segmentation 1.00
VCADNet: Vision-based Circular Accessible Depth Prediction for UGV Perception 1.00
One-Shot Affordance Grounding of Deformable Objects in Egocentric Organizing Scenes 1.00
EDSOD: An Encoder-Decoder, Diffusion-model, and Swin-Transformer-based Small Object Detector 1.00
DPGP: A Hybrid 2D-3D Dual Path Potential Ghost Probe Zone Prediction Framework for Safe Autonomous Driving 0.93
Simpler Is Better: Revisiting Doppler Velocity for Enhanced Moving Object Tracking with FMCW LiDAR 1.00
DISCOVERSE: Efficient Robot Simulation in Complex High-Fidelity Environments 1.00
Tracking Any Point with Frame-Event Fusion Network at High Frame Rate 1.00
SplatPose: Geometry-Aware 6-DoF Pose Estimation from Single RGB Image via 3D Gaussian Splatting 1.00
MaskSem: Semantic-Guided Masking for Learning 3D Hybrid High-Order Motion Representation 1.00
Reducing Scene Graph Generation Parameters Towards UAV Understanding of Structured Environments 1.00
Fusion Scene Context: Robust and Efficient LiDAR Place Recognition Across Season 1.00
Neural-Link: Non-Overlapping MPC Fusion and Passive Inertial Sensing on Soft Platforms 1.00
CSVO: Complementary-Pathway Spatial-Enhanced Visual Odometry for Extreme Environments with Brain-Inspired Vision Sensors 1.00
SR3D: Unleashing Single-view 3D Reconstruction for Transparent and Specular Object Grasping 1.00
PartGrasp: Generalizable Part-level Grasping via Semantic-Geometric Alignment 1.00
Neural Collision Detection for Constrained Grasp Pose Optimization in Cluttered Environments 1.00
Region-Centric 6-Dof Grasp Detection: A Data-Efficient Solution for Cluttered Scenes 1.00
MISCGrasp: Leveraging Multiple Integrated Scales and Contrastive Learning for Enhanced Volumetric Grasping 1.00
A Novel Large-Scale Collaborative Mapping Framework with Heterogeneous Point Clouds for Aerial-Ground Robots 1.00
sEMG-Based Continues Motion Prediction of Shoulder exoskeleton Control Using the VGANet Model* 1.00
Self-Distilled Stereo Matching: Real-Time Domain Generalization for Robotic Depth Perception 1.00
360Recon: An Accurate Reconstruction Method based on Depth Fusion from 360 Images 1.00
PB-MOT: Pose-aware Association Boosted Online 3D Multi-Object Tracking 1.00
RoboEngine: Plug-and-Play Robot Data Augmentation with Semantic Robot Segmentation and Background Generation 1.00
RCGNet: RGB-based Category-Level 6D Object Pose Estimation with Geometric Guidance 1.00
ROA-BEV: 2D Region-Oriented Attention for BEV-based 3D Object Detection 1.00
Adjacent-view Transformers for Supervised Surround-view Depth Estimation 0.88
Normalized Triangulation for Calibrated Dual-View 3D Human Pose Estimation 1.00
Pathfinder for Low-altitude Aircraft with Binary Neural Network 1.00
Region-Aware 6D Grasping for Industrial Bin-Picking: A Sim2Real Label Self-Generation and Hybrid Evaluation Framework 0.60
DynamicPose: Real-time and Robust 6D Object Pose Tracking for Fast-Moving Cameras and Objects 1.00
IMM-MOT: A Novel 3D Multi-object Tracking Framework with Interacting Multiple Model Filter 1.00
DynamicGSG: Dynamic 3D Gaussian Scene Graphs for Environment Adaptation 1.00
A New Unsupervised Infrared and Visible Image Fusion Method Based on Salient Object Segmentation under Poor Illumination 1.00
Advancing Depth Anything Model for Unsupervised Monocular Depth Estimation in Endoscopy 1.00
PAVLM: Advancing Point Cloud based Affordance Understanding Via Vision-Language Model 0.13
YO-CSA-T: A Real-time Badminton Tracking System Utilizing YOLO Based on Contextual and Spatial Attention 1.00
Learning Upright and Forward-Facing Object Poses using Category-level Canonical Representations 1.00
1 1.00
SynSeg: A synthetic data-driven approach for robust subcellular structure segmentation 1.00
5 4.46
Physical mechanisms governing generalization and hallucination in deep learning for imaging through scattering media 1.00
Pushing the limits of fluorescence imaging with a restoration neural network aggregating large-view statistics 1.00
Bridging the latency gap with a continuous stream evaluation framework in event-driven perception 0.77
Ultra-efficient physical field computing by complex-valued network quantization 0.69
A foundation model for multi-task cross-distribution restoration of fluorescence microscopy images 1.00
1 0.95
CELLECT: contrastive embedding learning for large-scale efficient cell tracking 0.95
1 0.88
Deeper detection limits in astronomical imaging using self-supervised spatiotemporal denoising 0.88
1 1.00
Motif field combined with two-stream feature fusion network and double detection head for identification and prediction of microalgae in seawater 1.00

Numerical information only (in the tables above) is licensed under a Creative Commons Attribution-NonCommercial-ShareAlike 4.0 International.