Searcharxiv⌕ Search

arXiv · 2609.39202

3D Reconstruction from Arthroscopic Images using NeRF: a preliminary in-silico study

Abstract

In knee arthroscopy surgery, accurate registration between preoperative and intraoperative anatomy is a critical step for patient-specific navigation. Achieving an accurate registration requires a reliable 3D reconstruction of the joint during surgery. Preoperative 3D models can be obtained from patient imaging through segmentation and reconstruction, but generating an intraoperative 3D representation remains particularly challenging. Arthroscopic imaging suffers from a limited field of view, low surface texture, and strong specular reflections, which make conventional feature-based 3D reconstruction methods unreliable. In this work, we investigate the application of MIS-NeRF (Minimally-Invasive Surgery Neural Radiance Fields) for reconstructing intraoperative knee 3D models from monocular arthroscopic images. The approach is evaluated on six simulated arthroscopic acquisitions representing six patient-specific knee 3D models. Both qualitative and quantitative results are presented to assess the reconstruction quality. The reconstructed knee 3D models were evaluated through their rendered images, achieving PSNR (Peak Signal-to-Noise Ratio) of 31.88 $\pm$ 2.82, SSIM (Structural Similarity Index) of 0.98 $\pm$ 0.004 and LPIPS (Learned Perceptual Image Patch Similarity) of 0.017 $\pm$ 0.006. These preliminary results suggest the feasibility of NeRF-based reconstruction in the challenging context of arthroscopy and may represent a promising step toward accurate in-silico preoperative-to-intraoperative 3D registration for computer-assisted orthopedic surgery. Further, validation on real arthroscopic data will be necessary to assess clinical applicability.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Hermine Kitio Tsamo, Agathe Yvinou, Daniel Pizarro, Samad Barri Khojasteh, Alexandre Tronchot, Antoine Ferreira, Eric Stindel, Aziliz Guezou-Philippe, Guillaume Dardenne. 2026-09-30. 3D Reconstruction from Arthroscopic Images using NeRF: a preliminary in-silico study. https://arxiv.org/abs/2609.39202

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

SAIPAS: Simulating aspect-angles-invariant physical adversarial attacks on SAR target recognition models

Synthetic aperture radar (SAR) enables versatile, all-time, all-weather remote sensing. Coupled with automatic target recognition (ATR) leveraging machine learning (ML), SAR is empowering a wide range of Earth observation and surveillance applications. However, the surge of attacks based on adversarial perturbations against the ML algorithms underpinning SAR ATR is prompting the need for systematic research into adversarial perturbation mechanisms. Research in this area began in the digital (image) domain and evolved into the (simulated) physical domain, resulting in physical adversarial attacks (PAAs) that strategically exploit corner reflectors as attack vectors to evade ML-based ATR. Existing PAAs assume that the attacker knows the SAR platform's aspect angles, restricting their applicability to idealized scenarios. We propose the Simulated Aspect-angle-Invariant Physical Adversarial SAR attack (SAIPAS), a framework that determines adversarially effective positions and orientations of any given set of reflectors, regardless of their number or size, even when the attacker lacks knowledge of the SAR platform's aspect angles. This is enabled by rigorous physics-based modeling of the reflected signal and the SAR imaging process. To facilitate mapping between image and scene coordinates, we additionally propose a method for generating bounding boxes in densely sampled azimuthal SAR images, allowing the target object to serve as a spatial reference. The resulting adversarial configurations offer a clear physical interpretation while maintaining high fooling rates across continuous aspect trajectories under more realistic operational assumptions (69.1% for AConvNet for a four-reflector white-box attack). This paper has supplementary material available, which demonstrates the SAIPAS.

eess.IV↗

A Unified Deep Learning Framework for Motion Correction in Medical Imaging

Deep learning has shown significant value in medical image registration for motion correction; however, current techniques are either limited by the type and range of motion they can handle or require iterative inference and/or retraining for new imaging data. To address these limitations, we introduce UniMo, a Unified Motion Correction framework that uses deep neural networks to correct various types of motion in medical imaging. UniMo uses an alternating optimization scheme with a unified loss function to train an integrated model of 1) an equivariant neural network for global rigid motion correction and 2) an encoder-decoder network for local deformations. It features a geometric deformation augmenter that 1) enhances the robustness of global motion correction by addressing local deformations, whether caused by non-rigid motion or geometric distortions, and 2) generates augmented data to improve training. As a hybrid model that uses both image intensities and shapes, UniMo is robust to appearance variations and generalizes to various imaging modalities without retraining. We trained and tested UniMo for motion tracking in fetal magnetic resonance imaging, which is challenging due to 1) both large rigid and non-rigid motion and 2) large variations in image appearance. We then tested the trained model, without retraining, on three public datasets: MedMNIST, lung CT, and BraTS. UniMo surpassed existing motion correction methods in accuracy and, notably, enabled one-time training on a single modality while maintaining high stability and adaptability across multiple unseen imaging datasets. By offering a unified solution to motion correction, UniMo marks a significant advance in challenging applications with a mixture of bulk motion and local deformations. Code is available at https://github.com/IntelligentImaging/UNIMO

eess.IV↗

MedForj: An open, large-scale foundational generative prior for high-resolution 3D brain MRI

This work introduces MedForj, a suite of 3D foundational generative priors based on diffusion models. The MedForj models were trained on $72{,}659$ 1~mm isotropic 3D $T_1$-weighted MRI human brain image volumes from $38{,}174$ subjects, drawn from a curated corpus of $80{,}675$ volumes from $42{,}506$ subjects spanning $38$ publicly available datasets. These training images were manually inspected to exclude those with poor quality and excessive pathology, and otherwise were minimally processed. The models include six different diffusion training strategies: rectified flow, latent diffusion rectified flow, flow matching, velocity prediction, clean prediction, and noise prediction. Image samples produced by each of these models were compared to each other and against real, ground truth data under downstream segmentation distributions, FID, five inverse problems, and blind human inspection in an observer study. Flow matching was the strongest strategy overall, achieving the best inverse problem solving results at $28.80$~dB PSNR and $0.874$ SSIM averaged over the five forward problems, the highest rate of reconstructions judged real by blind human raters at $72.6\%$, and the closest per-structure match to real segmented anatomy in a permutation test. It was not best everywhere: rectified flow produced the most convincing unconditional samples in the observer study and the best FID, and the latent rectified-flow model achieved the smallest joint distributional distance to real anatomy. No other strategy, however, performed consistently well across all four evaluations. We therefore recommend flow matching as the default MedForj prior, while releasing every strategy so that the choice can be revisited per application. All model weights and corresponding code are publicly available at https://github.com/piksl-research/medforj.

eess.IV↗