Papers nuevos sobre Imágenes y sensores
51 papers nuevos sobre imágenes y sensores en los últimos 7 días, dentro de Robótica e ingeniería. Acá están los 50 que Pipette considera más valiosos, con el resultado principal en palabras de sus autores.
Lo mejor de la semana
Dynamic flexible optical sensing based on liquid integrated circuits
Here, we propose a DFOS system based on liquid integrated circuits (LICs), integrating photosensitive ionic liquid (PIL) and stretchable liquid metal electrodes to achieve active curvature modulation for geometry-adaptive optical sensing.
Revista con revisión por paresUso en el mundo realORION-CMR: On-scanner Reporting with Integrated Foundation Model for End-to-End Cardiac MRI Analysis and Interpretation
We present ORION-CMR (On-scanner Reporting with Integrated fOunda-tioN Model), the first clinically evaluated scanner-native end-to-end CMR foundation model.
PreprintDice ser un gran avanceUso en el mundo realIntegrating Local Detail and Global Context: A Dual-Input Multi-Task Learning Framework for Bone Tumor Diagnosis
To address the limitations of existing single-view models, we present a dual-input, multi-task learning framework that, to our knowledge, is the first to apply bidirectional cross-modal attention between a lesion crop and the full radiograph for joint segmentation and subtype classification.
PreprintAfirmaciones fuertes, leer con cuidadoUso en el mundo realAutoRASOR: Autonomous Rapid Scanning Electron Microscope Operator
We introduce AutoRASOR, a task-agnostic autonomous SEM pipeline that captures the multi-scale morphology of an unknown sample without domain-specific pre-training, fine-tuning, or human prompting.
PreprintAfirmaciones fuertes, leer con cuidadoUso en el mundo realRevolutionizing Diffusion MRI Microstructure Mapping via Global Inversion
We instead cast MM as a single global inverse problem, reconstructing the entire parameter field jointly from all measurements of a subject.
PreprintUso en el mundo realSignal recovery in the polychromatic computed tomography model: Injectivity and complexity
We construct a measurement ensemble that ensures perfect signal recovery almost surely provided , thereby isolating the injectivity threshold up to an additive factor .
PreprintBioimpedance meets biomechanics: Wearable electrical impedance myography encodes fascicle and activation dynamics
These findings validate EIM as a robust, real-time biomechanical tool for sensing muscle kinematics and neuromuscular activity during functional movement, expanding its use in assistive robotic control and injury prevention.
Revista con revisión por paresAfirmaciones fuertes, leer con cuidadoUso en el mundo realA Unified Frequency-Domain Model for Cascaded Filter-Interpolation Modulation in Tomographic Reconstruction
Here, we introduce a unified frequency-domain model that conceptualizes the combined effect of filtering and interpolation in the filtered backprojection (FBP) algorithm as a cascaded modulation process.
PreprintAfirmaciones fuertes, leer con cuidadoUso en el mundo realJEVQA - Video Quality from Metadata, Bitstream, and Pixel Features with a General-Purpose Decision Model
Our results show that trained models on the same features remain clearly ahead in both studies, but that zero-shot classifiers are promising.
PreprintAdaptive Color Grading
Results show that K-nearest neighbors is an effective prediction strategy, outperforming state-of-the-art end-to-end methods for image enhancement.
PreprintUso en el mundo realCódigo disponibleThe EventCV Library for Event-Based Robotic Vision
Here, we present EventCV, an open-source and extensible Rust library with OpenCV-style Python bindings that lowers the entry barrier to working with event cameras.
PreprintUso en el mundo realM3GD: Multi-Modal Multi-View Geometric Diffusion for Camera--LiDAR Novel View Synthesis
We present M3GD, a Camera--LiDAR multimodal representation for generative NVS that composes independently pretrained 2D image and 3D point-cloud foundation models without separately pretraining a cross-modal translator.
PreprintUso en el mundo realHierarchical Filter Band Selection for Multispectral Object Classification
The proposed method achieves a 28.9% relative reduction in classification error on the SMM50 dataset compared to the state-of-the-art.
Preprint con versión publicadaUso en el mundo realAWR-Net: Decoupling Anatomy and Appearance for 3D Fetal Brain Ultrasound Synthesis
To address this challenge, we separate anatomical correspondence learning from ultrasound appearance adaptation at both the data and model levels.
PreprintUso en el mundo realMatcherCompass: A Deployment-Aware Benchmark to Guide Image Matcher Selection in the Wild
We present MatcherCompass, a deployment-aware benchmark for choosing local feature matchers in field robotics.
PreprintUso en el mundo realDisparity Estimation of Planar Reflective Surfaces Using Specular Reflections From a Single Light Source
To address this issue, the novel Specular Reflection Disparity Estimation SRDE algorithm is introduced, which is specifically designed for the constrained scenario of planar, textureless objects and single-source illumination.
Preprint con versión publicadaUso en el mundo realTransparentize A Shallow Cryosphere: High-Resolution Subsurface Imaging using UAV-Borne GPR A review and prospective
This contribution reviews recent progress in UAV-borne GPR for cryosphere investigations, encompassing system architectures, operational strategies, and representative deployment practices.
Preprint con versión publicadaUso en el mundo realBeyond the Flat Seafloor: A Closed-Form Two-View Constraint to Aid Sidescan Sonar Reconstruction
Rather than make similar approximations, this paper focuses on a multi-view geometry based approach and formalizes a two-view geometric constraint and proves that a shared feature is constrained to a locus within the intersection of a sphere and a plane.
PreprintUso en el mundo realVideo Based Assessment of Surgical Skills Using Frozen Pretrained Video Foundation Models
We introduce VBA-Net+, a video-only framework for Fundamentals of Laparoscopic Surgery (FLS) score regression and pass-fail classification using pretrained video foundation models as frozen feature extractors.
Preprint con versión publicadaUso en el mundo realRate-distortion optimization for full-reference image quality metrics via stochastic Hessian estimates
Across five metrics for Kodak and CLIC in VVC, IDQD-RDO achieves 14.2-36.7 % BD-rate savings under the target metric with no decoder changes and incurs 10-30 % encoding complexity overhead.
PreprintUso en el mundo realForecasting Intrathecal Tracer Enhancement from Pre-Contrast Brain MRI: Direct Regression versus Flow Matching
Much of tracer enhancement is thus predictable from anatomy and timing, and direct regression is the more accurate and cheaper choice at this data scale.
PreprintUso en el mundo realRecursive Uncertainty-Gated Image Registration for Learning-based Algorithms
We propose Recursive Uncertainty-Gated Image Registration (RUGI), an algorithm for iteratively refining deformation fields predicted by learning-based registration models.
PreprintUso en el mundo realPhysics-Guided Multi-Objective Deep Learning for Ultrasound RF Data Interpolation in Resource-Constrained Imaging
We present a physics-guided, data-driven framework for sparse-to-dense radio-frequency (RF) reconstruction that aligns training with downstream image formation.
PreprintUso en el mundo realWhen is a closed-form RGB->S/P ratio adequate? A hyperspectral characterization on natural scenes for mesopic display
On a daylight radiance time-series, a six-scalar closed form (three photopic and three scotopic channel weights) reproduces spectral S/P with a median error of ~0.07 that is time-invariant once the RGB input is chromatically adapted to D65; evaluated in un-adapted sRGB the error instead carries a color-temperature tilt across illuminants (~0.19), so adaptation is the enabling step for this use case.
PreprintUso en el mundo realUncertainty-driven training for three-dimensional calibrated lung nodule classification
In this work, we present an uncertainty-driven training framework for three-dimensional computed tomography (CT) lung nodule classification, where validation-based uncertainty estimates guide loss reweighting to enhance predictive performance and probability calibration.
Preprint con versión publicadaUso en el mundo realWS-NeRF: A Mamba-Driven World-State-Aware Adaptive Deblurring Neural Radiance Field
In this paper, we propose a novel Mamba-driven world-state-aware adaptive deblurring neural radiance field, termed WS-NeRF, to address image degradation and 3D inconsistency.
PreprintSomaNet: Weakly Supervised Learning for Instance Soma Segmentation in 3D Electron Microscopy with Partial Annotations
To address this challenge, we propose SomaNet, a weakly supervised framework for 3D EM soma instance segmentation under partial annotation constraints.
PreprintCódigo disponibleHigh-Temporal-Resolution Motion Correction in Magnetic Resonance Fingerprinting Using a Quantitative Scout and Compact Spiral Navigators
We propose a navigation framework that integrates compact k-space navigators throughout the MRF acquisition, enabling sub-second motion estimation at minimal sequence overhead.
PreprintUso en el mundo realAnatomically Faithful Artifact Suppression in SENSE Accelerated Brain MRI
Conclusion: ART-Net improved agreement between SENSE4 and fully sampled T1-weighted images in a single-center, held-out test cohort while maintaining segmentation-derived anatomical measurements.
PreprintUso en el mundo realScalable photoacoustic tomography implementations accounting for the spatial impulse response of transducers
Two implementations are proposed, differing only in how this quantity is evaluated: a quadrature over points of the surface, as in existing works, or a closed-form area, which never discretizes the surface.
PreprintUso en el mundo realMulticentre Bi-atrial Segmentation from LGE-MRI for Atrial Fibrillation with a 2D and 3D Framework
This study presents a two-stage segmentation framework and benchmarking platform for evaluating how ROI localisation, encoder design, 2D/3D dimensionality, and ensemble fusion affect bi-atrial wall and cavity segmentation across multicentre LGE-MRI datasets.
PreprintUso en el mundo realCharacterizing Experiments with Synthetic MSI/HSI Data: A Structured Taxonomy for Agri-Food Research
We present a literature-informed taxonomy for experiments using synthetic MSI/HSI data through spectral reconstruction and/or data augmentation, with a deliberate focus on agri-food research.
PreprintDoes DCGAN-Based Synthetic Augmentation Improve Brain Tumor MRI Classification? An Empirical Study
The two models achieved the same overall accuracy of 96%, while macro F1 remained effectively unchanged and ROC-AUC decreased slightly from 0.987 to 0.982 after augmentation.
PreprintUso en el mundo realOpportunistic Conditional Entropy Coding with Frozen Analysis and Synthesis Transforms
We introduce a single entropy model that conditions on a previously decoded latent when available and falls back to a standard hyperprior otherwise.
PreprintUso en el mundo realScalable SSIM Estimation from PSNR for Per-Title and Context-Adaptive Encoding Workflows
We propose ApproxSSIMate, a low-complexity method for estimating SSIM from PSNR combined with reference-sequence statistics computed once per sequence and reused across every candidate encode.
PreprintUso en el mundo realClassification-oriented adaptive sensing via posterior sampling
We introduce a classification-driven extension motivated by the closed-form posterior covariance of a class-conditional Gaussian mixture model, which decomposes into within-class and between-class uncertainty.
PreprintQuality Assessment of 3D Gaussian Splatting: Distortions, Benchmarks, and Open Challenges
This survey reviews recent 3DGS quality assessment studies from four perspectives, covering distortion characteristics, subjective benchmarks, objective metric reliability, and emerging directions for native 3DGS evaluation.
Preprint con versión publicadaRobust, Estimator-Agnostic Dynamic 3DGS Compression
We concatenate groups of frames into one Gaussian set, augment each Gaussian with a frame index, and pass it to a static (i.e., non-temporal) 3DGS codec, converting temporal redundancy into spatial redundancy.
PreprintResolution-Flexible Decoding for Hybrid Neural Video Representations
In this paper, we propose a resolution-flexible decoder framework for hybrid NVRs.
PreprintLesion-Gated Hybrid Synthesis for Virtual Contrast-Enhanced Breast MRI: A MAMA-SYNTH Challenge Solution
Conclusion: A test-compatible lesion probability map derived from the pre-contrast image can coordinate complementary tumour-focused regression and background-focused perceptual synthesis, enabling balanced virtual contrast enhancement under a multi-metric challenge setting.
PreprintUso en el mundo realBRiDCT: Fast Two-Dimensional DCTs Using SIMD: SIMD Organization, Register Blocking, and Numerical Verification
On one Apple M3 Max, the final native library has lower execution times than the tested general-purpose routes---Apple's Accelerate framework (vDSP), FFTW and Ooura---across their available cases from to .
PreprintUso en el mundo realModel-Based Iterative Reconstruction with View-Dependent Detector Displacements for Cone-Beam CT
We extend the separable distance-driven MBIR model to incorporate recorded horizontal and vertical detector displacements into the distance-driven overlap computation while leaving measured projections unchanged.
PreprintUso en el mundo realComplementary-Aperture Pulse Sequencing for Fundamental-Band Nonlinearity Parameter Imaging
These results demonstrate that CAPS preserves depletion-based B/A estimation without using the pressure scaling factor as a reconstruction input, supporting experimental and in vivo implementation without pressure-ratio calibration.
PreprintUso en el mundo realEntropy-map SSIM analysis of Salt and Pepper Noise Removal via Recursive Median Filterring
We show that SSIM-Map is more sensitive to residual impulse artifacts, blur, and local intensity transitions, and therefore complements the conventional SSIM-Img metric.
PreprintLC3EM: Long-Range Context Extrapolation Enhanced Entropy Model for Coordinate-based Overfitting Image Codecs
Inspired by the prediction mechanism in traditional codecs, we propose a new entropy-modeling strategy that introduces complementary prediction modes with region-adaptive soft mode selection, rather than relying on a single learned predictor to model diverse types of redundancy.
PreprintDA-Lion: Efficient Neural Video Representation via Direction-Aware Optimization
To address this, we propose Direction-Aware Lion (DA-Lion), a task-driven optimizer tailored for NVR.
PreprintCódigo disponibleVGG16-MCA UNet: Whole-Tumor Segmentation in 2D FLAIR MRI with Decoder-Side Channel Attention
We present VGG16-MCA UNet, a hybrid architecture pairing an ImageNet-pretrained VGG16 encoder with a decoder in which a Multi-Channel Attention (MCA) module recalibrates features after each skip-connection fusion, trained with the Focal Tversky loss to counter severe foreground-background imbalance.
PreprintUso en el mundo realCódigo disponibleA Reference-Based Protocol for Assessing Image Displacement and Scale Stability
The contribution is a workflow that reports completerun variability, typical deviations and estimator coverage together; the case study does not establish support efficacy, equivalence or transient damping.
PreprintGraph-Based Semi-Supervised Hyperspectral Image Classification with Distance-Aware Spatial Measure
This work proposes a composite kernel approach that includes a third kernel dealing exclusively with the relative spatial position of the pixels.
PreprintAdaptive Tiling for Least-Squares Phase Unwrapping: Runtime and Accuracy
In single-threaded experiments on a heterogeneous image dataset, optimized adaptive partitions use fewer tiles but remain slower than the optimized grid, and some reconstructions lose substantial accuracy.
Preprint