New papers on Imaging & sensing
51 new papers on imaging & sensing in the last 7 days, within Robotics & engineering. These are the 50 Pipette rates most worth reading, with the main result in the authors' own words.
The best of the week
Dynamic flexible optical sensing based on liquid integrated circuits
Here, we propose a DFOS system based on liquid integrated circuits (LICs), integrating photosensitive ionic liquid (PIL) and stretchable liquid metal electrodes to achieve active curvature modulation for geometry-adaptive optical sensing.
Peer-reviewed journalReal-world useORION-CMR: On-scanner Reporting with Integrated Foundation Model for End-to-End Cardiac MRI Analysis and Interpretation
We present ORION-CMR (On-scanner Reporting with Integrated fOunda-tioN Model), the first clinically evaluated scanner-native end-to-end CMR foundation model.
PreprintClaims a big stepReal-world useIntegrating Local Detail and Global Context: A Dual-Input Multi-Task Learning Framework for Bone Tumor Diagnosis
To address the limitations of existing single-view models, we present a dual-input, multi-task learning framework that, to our knowledge, is the first to apply bidirectional cross-modal attention between a lesion crop and the full radiograph for joint segmentation and subtype classification.
PreprintBold claims, read criticallyReal-world useAutoRASOR: Autonomous Rapid Scanning Electron Microscope Operator
We introduce AutoRASOR, a task-agnostic autonomous SEM pipeline that captures the multi-scale morphology of an unknown sample without domain-specific pre-training, fine-tuning, or human prompting.
PreprintBold claims, read criticallyReal-world useRevolutionizing Diffusion MRI Microstructure Mapping via Global Inversion
We instead cast MM as a single global inverse problem, reconstructing the entire parameter field jointly from all measurements of a subject.
PreprintReal-world useSignal recovery in the polychromatic computed tomography model: Injectivity and complexity
We construct a measurement ensemble that ensures perfect signal recovery almost surely provided , thereby isolating the injectivity threshold up to an additive factor .
PreprintBioimpedance meets biomechanics: Wearable electrical impedance myography encodes fascicle and activation dynamics
These findings validate EIM as a robust, real-time biomechanical tool for sensing muscle kinematics and neuromuscular activity during functional movement, expanding its use in assistive robotic control and injury prevention.
Peer-reviewed journalBold claims, read criticallyReal-world useA Unified Frequency-Domain Model for Cascaded Filter-Interpolation Modulation in Tomographic Reconstruction
Here, we introduce a unified frequency-domain model that conceptualizes the combined effect of filtering and interpolation in the filtered backprojection (FBP) algorithm as a cascaded modulation process.
PreprintBold claims, read criticallyReal-world useJEVQA - Video Quality from Metadata, Bitstream, and Pixel Features with a General-Purpose Decision Model
Our results show that trained models on the same features remain clearly ahead in both studies, but that zero-shot classifiers are promising.
PreprintAdaptive Color Grading
Results show that K-nearest neighbors is an effective prediction strategy, outperforming state-of-the-art end-to-end methods for image enhancement.
PreprintReal-world useCode availableThe EventCV Library for Event-Based Robotic Vision
Here, we present EventCV, an open-source and extensible Rust library with OpenCV-style Python bindings that lowers the entry barrier to working with event cameras.
PreprintReal-world useM3GD: Multi-Modal Multi-View Geometric Diffusion for Camera--LiDAR Novel View Synthesis
We present M3GD, a Camera--LiDAR multimodal representation for generative NVS that composes independently pretrained 2D image and 3D point-cloud foundation models without separately pretraining a cross-modal translator.
PreprintReal-world useHierarchical Filter Band Selection for Multispectral Object Classification
The proposed method achieves a 28.9% relative reduction in classification error on the SMM50 dataset compared to the state-of-the-art.
Preprint with a published versionReal-world useAWR-Net: Decoupling Anatomy and Appearance for 3D Fetal Brain Ultrasound Synthesis
To address this challenge, we separate anatomical correspondence learning from ultrasound appearance adaptation at both the data and model levels.
PreprintReal-world useMatcherCompass: A Deployment-Aware Benchmark to Guide Image Matcher Selection in the Wild
We present MatcherCompass, a deployment-aware benchmark for choosing local feature matchers in field robotics.
PreprintReal-world useDisparity Estimation of Planar Reflective Surfaces Using Specular Reflections From a Single Light Source
To address this issue, the novel Specular Reflection Disparity Estimation SRDE algorithm is introduced, which is specifically designed for the constrained scenario of planar, textureless objects and single-source illumination.
Preprint with a published versionReal-world useTransparentize A Shallow Cryosphere: High-Resolution Subsurface Imaging using UAV-Borne GPR A review and prospective
This contribution reviews recent progress in UAV-borne GPR for cryosphere investigations, encompassing system architectures, operational strategies, and representative deployment practices.
Preprint with a published versionReal-world useBeyond the Flat Seafloor: A Closed-Form Two-View Constraint to Aid Sidescan Sonar Reconstruction
Rather than make similar approximations, this paper focuses on a multi-view geometry based approach and formalizes a two-view geometric constraint and proves that a shared feature is constrained to a locus within the intersection of a sphere and a plane.
PreprintReal-world useVideo Based Assessment of Surgical Skills Using Frozen Pretrained Video Foundation Models
We introduce VBA-Net+, a video-only framework for Fundamentals of Laparoscopic Surgery (FLS) score regression and pass-fail classification using pretrained video foundation models as frozen feature extractors.
Preprint with a published versionReal-world useRate-distortion optimization for full-reference image quality metrics via stochastic Hessian estimates
Across five metrics for Kodak and CLIC in VVC, IDQD-RDO achieves 14.2-36.7 % BD-rate savings under the target metric with no decoder changes and incurs 10-30 % encoding complexity overhead.
PreprintReal-world useForecasting Intrathecal Tracer Enhancement from Pre-Contrast Brain MRI: Direct Regression versus Flow Matching
Much of tracer enhancement is thus predictable from anatomy and timing, and direct regression is the more accurate and cheaper choice at this data scale.
PreprintReal-world useRecursive Uncertainty-Gated Image Registration for Learning-based Algorithms
We propose Recursive Uncertainty-Gated Image Registration (RUGI), an algorithm for iteratively refining deformation fields predicted by learning-based registration models.
PreprintReal-world usePhysics-Guided Multi-Objective Deep Learning for Ultrasound RF Data Interpolation in Resource-Constrained Imaging
We present a physics-guided, data-driven framework for sparse-to-dense radio-frequency (RF) reconstruction that aligns training with downstream image formation.
PreprintReal-world useWhen is a closed-form RGB->S/P ratio adequate? A hyperspectral characterization on natural scenes for mesopic display
On a daylight radiance time-series, a six-scalar closed form (three photopic and three scotopic channel weights) reproduces spectral S/P with a median error of ~0.07 that is time-invariant once the RGB input is chromatically adapted to D65; evaluated in un-adapted sRGB the error instead carries a color-temperature tilt across illuminants (~0.19), so adaptation is the enabling step for this use case.
PreprintReal-world useUncertainty-driven training for three-dimensional calibrated lung nodule classification
In this work, we present an uncertainty-driven training framework for three-dimensional computed tomography (CT) lung nodule classification, where validation-based uncertainty estimates guide loss reweighting to enhance predictive performance and probability calibration.
Preprint with a published versionReal-world useWS-NeRF: A Mamba-Driven World-State-Aware Adaptive Deblurring Neural Radiance Field
In this paper, we propose a novel Mamba-driven world-state-aware adaptive deblurring neural radiance field, termed WS-NeRF, to address image degradation and 3D inconsistency.
PreprintSomaNet: Weakly Supervised Learning for Instance Soma Segmentation in 3D Electron Microscopy with Partial Annotations
To address this challenge, we propose SomaNet, a weakly supervised framework for 3D EM soma instance segmentation under partial annotation constraints.
PreprintCode availableHigh-Temporal-Resolution Motion Correction in Magnetic Resonance Fingerprinting Using a Quantitative Scout and Compact Spiral Navigators
We propose a navigation framework that integrates compact k-space navigators throughout the MRF acquisition, enabling sub-second motion estimation at minimal sequence overhead.
PreprintReal-world useAnatomically Faithful Artifact Suppression in SENSE Accelerated Brain MRI
Conclusion: ART-Net improved agreement between SENSE4 and fully sampled T1-weighted images in a single-center, held-out test cohort while maintaining segmentation-derived anatomical measurements.
PreprintReal-world useScalable photoacoustic tomography implementations accounting for the spatial impulse response of transducers
Two implementations are proposed, differing only in how this quantity is evaluated: a quadrature over points of the surface, as in existing works, or a closed-form area, which never discretizes the surface.
PreprintReal-world useMulticentre Bi-atrial Segmentation from LGE-MRI for Atrial Fibrillation with a 2D and 3D Framework
This study presents a two-stage segmentation framework and benchmarking platform for evaluating how ROI localisation, encoder design, 2D/3D dimensionality, and ensemble fusion affect bi-atrial wall and cavity segmentation across multicentre LGE-MRI datasets.
PreprintReal-world useCharacterizing Experiments with Synthetic MSI/HSI Data: A Structured Taxonomy for Agri-Food Research
We present a literature-informed taxonomy for experiments using synthetic MSI/HSI data through spectral reconstruction and/or data augmentation, with a deliberate focus on agri-food research.
PreprintDoes DCGAN-Based Synthetic Augmentation Improve Brain Tumor MRI Classification? An Empirical Study
The two models achieved the same overall accuracy of 96%, while macro F1 remained effectively unchanged and ROC-AUC decreased slightly from 0.987 to 0.982 after augmentation.
PreprintReal-world useOpportunistic Conditional Entropy Coding with Frozen Analysis and Synthesis Transforms
We introduce a single entropy model that conditions on a previously decoded latent when available and falls back to a standard hyperprior otherwise.
PreprintReal-world useScalable SSIM Estimation from PSNR for Per-Title and Context-Adaptive Encoding Workflows
We propose ApproxSSIMate, a low-complexity method for estimating SSIM from PSNR combined with reference-sequence statistics computed once per sequence and reused across every candidate encode.
PreprintReal-world useClassification-oriented adaptive sensing via posterior sampling
We introduce a classification-driven extension motivated by the closed-form posterior covariance of a class-conditional Gaussian mixture model, which decomposes into within-class and between-class uncertainty.
PreprintQuality Assessment of 3D Gaussian Splatting: Distortions, Benchmarks, and Open Challenges
This survey reviews recent 3DGS quality assessment studies from four perspectives, covering distortion characteristics, subjective benchmarks, objective metric reliability, and emerging directions for native 3DGS evaluation.
Preprint with a published versionRobust, Estimator-Agnostic Dynamic 3DGS Compression
We concatenate groups of frames into one Gaussian set, augment each Gaussian with a frame index, and pass it to a static (i.e., non-temporal) 3DGS codec, converting temporal redundancy into spatial redundancy.
PreprintResolution-Flexible Decoding for Hybrid Neural Video Representations
In this paper, we propose a resolution-flexible decoder framework for hybrid NVRs.
PreprintLesion-Gated Hybrid Synthesis for Virtual Contrast-Enhanced Breast MRI: A MAMA-SYNTH Challenge Solution
Conclusion: A test-compatible lesion probability map derived from the pre-contrast image can coordinate complementary tumour-focused regression and background-focused perceptual synthesis, enabling balanced virtual contrast enhancement under a multi-metric challenge setting.
PreprintReal-world useBRiDCT: Fast Two-Dimensional DCTs Using SIMD: SIMD Organization, Register Blocking, and Numerical Verification
On one Apple M3 Max, the final native library has lower execution times than the tested general-purpose routes---Apple's Accelerate framework (vDSP), FFTW and Ooura---across their available cases from to .
PreprintReal-world useModel-Based Iterative Reconstruction with View-Dependent Detector Displacements for Cone-Beam CT
We extend the separable distance-driven MBIR model to incorporate recorded horizontal and vertical detector displacements into the distance-driven overlap computation while leaving measured projections unchanged.
PreprintReal-world useComplementary-Aperture Pulse Sequencing for Fundamental-Band Nonlinearity Parameter Imaging
These results demonstrate that CAPS preserves depletion-based B/A estimation without using the pressure scaling factor as a reconstruction input, supporting experimental and in vivo implementation without pressure-ratio calibration.
PreprintReal-world useEntropy-map SSIM analysis of Salt and Pepper Noise Removal via Recursive Median Filterring
We show that SSIM-Map is more sensitive to residual impulse artifacts, blur, and local intensity transitions, and therefore complements the conventional SSIM-Img metric.
PreprintLC3EM: Long-Range Context Extrapolation Enhanced Entropy Model for Coordinate-based Overfitting Image Codecs
Inspired by the prediction mechanism in traditional codecs, we propose a new entropy-modeling strategy that introduces complementary prediction modes with region-adaptive soft mode selection, rather than relying on a single learned predictor to model diverse types of redundancy.
PreprintDA-Lion: Efficient Neural Video Representation via Direction-Aware Optimization
To address this, we propose Direction-Aware Lion (DA-Lion), a task-driven optimizer tailored for NVR.
PreprintCode availableVGG16-MCA UNet: Whole-Tumor Segmentation in 2D FLAIR MRI with Decoder-Side Channel Attention
We present VGG16-MCA UNet, a hybrid architecture pairing an ImageNet-pretrained VGG16 encoder with a decoder in which a Multi-Channel Attention (MCA) module recalibrates features after each skip-connection fusion, trained with the Focal Tversky loss to counter severe foreground-background imbalance.
PreprintReal-world useCode availableA Reference-Based Protocol for Assessing Image Displacement and Scale Stability
The contribution is a workflow that reports completerun variability, typical deviations and estimator coverage together; the case study does not establish support efficacy, equivalence or transient damping.
PreprintGraph-Based Semi-Supervised Hyperspectral Image Classification with Distance-Aware Spatial Measure
This work proposes a composite kernel approach that includes a third kernel dealing exclusively with the relative spatial position of the pixels.
PreprintAdaptive Tiling for Least-Squares Phase Unwrapping: Runtime and Accuracy
In single-threaded experiments on a heterogeneous image dataset, optimized adaptive partitions use fewer tiles but remain slower than the optimized grid, and some reconstructions lose substantial accuracy.
Preprint