SearcharxivSearch

arXiv subjects

Wei Ouyang

Publications and source records attributed to Wei Ouyang.

At least 19 recordsLinked to original sources

The Setting of IMU Parameters in Kalman Filtering-based Information Fusion

The setting or tuning of specifications for the inertial measurement unit (IMU) is tricky in sensor fusion. The underneath conundrum is caused by the fact that the working condition of IMU is more complex than the stationary calibration scenario. Since the noises and biases instabilities calibrated under static condition cannot accommodate other cases, the effective tuning of IMU parameters largely hinges on the experience or profound understanding of the system. In the current work, the setting method of IMU parameters based on Allan variance calibration is delved into within the Kalman filtering framework. Specifically, the relationship between the power sepctral density and Allan variance is leveraged in formulating the process uncertainty in continuous-time filtering. Three typical IMU-based sensor fusion systems, including INS/GNSS integration, LiDAR-inertial odometry, and visual-inertial odometry are considered to show the feasibility and effectiveness of this parameter setting process.

cs.RO

Predictive single cell foundation model for gene regulation and aging with privacy-preserving tabular learning

Pre-trained foundation models (FMs) have begun transforming single-cell genomics, but scaling them raises privacy concerns. Moreover, unlike text data, single-cell data is unordered and exhibits a unique tabular structure that current single-cell FMs overlook. We introduce Tabula, a privacy-preserving FM designed with federated learning (FL) that explicitly models the tabular structure of single-cell data. To deploy Tabula, we further developed Chiron, a decentralized AI agent-enabled platform for collaborative training across institutions without sharing raw data. Beyond strong performance across downstream benchmarks, Tabula reveals combinatorial regulatory logic across diverse biological systems, including hematopoiesis, pancreatic endogenesis, neurogenesis, and cardiogenesis. Using a new scRNA-seq dataset of paired young and aged human fibroblasts, Tabula nominates rejuvenation factors through age- and identity score-guided in silico prioritization, outperforming conventional approaches. Thus, Tabula represents an important advance in single-cell foundation modeling by integrating tabular learning with FL, paving the way toward privacy-preserving virtual cells for human health.

cs.LG

Meta-learning-enhanced implicit full waveform inversion

Implicit full waveform inversion (IFWI) introduces implicit neural representations to parameterize the subsurface velocity model as a continuous function of spatial coordinates, which alleviates the dependence on the initial model and improves inversion flexibility. However, IFWI still requires a large number of iterative updates for each new exploration area, leading to slow convergence, high computational cost, and a lack of mechanisms to share prior knowledge across different geological settings, thereby limiting its efficiency and generalization capability. To further accelerate convergence and enhance cross-area generalization, we propose a meta-learning-based implicit full waveform inversion method, referred to as Meta-learning-enhanced implicit full waveform inversion (Meta-IFWI). In this framework, the subsurface velocity model is represented using an implicit neural network with periodic activation functions (SIREN), while a meta-learning strategy is employed to pretrain a single network on multiple velocity inversion tasks. Through this process, the network learns shared inversion priors and rapid adaptation strategies across different geological scenarios. For a new inversion task, the pretrained Meta-IFWI model can be efficiently adapted to the observed seismic data with only a few gradient updates, significantly reducing the number of iterations required for inversion. Numerical experiments conducted on in-distribution models, including layered synthetic models and the Overthrust model, as well as out-of-distribution complex models such as Marmousi 2, demonstrate that, compared with conventional IFWI, the proposed Meta-IFWI achieves improved inversion accuracy while substantially accelerating convergence and reducing computational cost. Moreover, Meta-IFWI exhibits enhanced robustness and stronger cross-area generalization capability.

physics.geo-ph

Adaptive Self-Supervised Surface-Related Multiple Suppression

Effective suppression of surface-related multiples is essential to prevent imaging artifacts and erroneous structural interpretations. While conventional approaches rely on accurate priors or subsurface model knowledge, and supervised learning methods require labeled data that are impractical to obtain for real seismic data. To overcome these limitations, a recently proposed self-supervised learning (SSL) framework integrates multi-dimensional convolution (MDC) for multiple generation with a two-stage training strategy, eliminating the need for both prior knowledge and labeled data. However, their approach requires manual selection of a scaling factor to match the amplitudes between the MDC-generated multiples and the true multiples, thus introducing subjectivity and limiting its practical applicability. In this study, we propose an adaptive SSL method that treats the scaling factor as a learnable parameter, jointly optimized with the network weights in a unified single-stage training pipeline. This dynamic scaling implicitly introduces amplitude diversity into the training data, acting as an implicit regularizer that improves the network's robustness to amplitude variations of surface-related multiples. We further design a composite loss function with homoscedastic uncertainty-based adaptive weighting, which automatically balances the contributions of multiple loss terms without manual tuning. Synthetic and field data examples demonstrate that our method robustly and effectively suppresses surface-related multiples while preserving primary reflections, with migration results confirming improved subsurface imaging quality.

physics.geo-ph

Restless Multi-Process Multi-Armed Bandits with Applications to Self-Driving Microscopies

High-content screening microscopy generates large amounts of live-cell imaging data, yet its potential remains constrained by the inability to determine when and where to image most effectively. Optimally balancing acquisition time, computational capacity, and photobleaching budgets across thousands of dynamically evolving regions of interest remains an open challenge, further complicated by limited field-of-view adjustments and sensor sensitivity. Existing approaches either rely on static sampling or heuristics that neglect the dynamic evolution of biological processes, leading to inefficiencies and missed events. Here, we introduce the restless multi-process multi-armed bandit (RMPMAB), a new decision-theoretic framework in which each experimental region is modeled not as a single process but as an ensemble of Markov chains, thereby capturing the inherent heterogeneity of biological systems such as asynchronous cell cycles and heterogeneous drug responses. Building upon this foundation, we derive closed-form expressions for transient and asymptotic behaviors of aggregated processes, and design scalable Whittle index policies with sub-linear complexity in the number of imaging regions. Through both simulations and a real biological live-cell imaging dataset, we show that our approach achieves substantial improvements in throughput under resource constraints. Notably, our algorithm outperforms Thomson Sampling, Bayesian UCB, epsilon-Greedy, and Round Robin by reducing cumulative regret by more than 37% in simulations and capturing 93% more biologically relevant events in live imaging experiments, underscoring its potential for transformative smart microscopy. Beyond improving experimental efficiency, the RMPMAB framework unifies stochastic decision theory with optimal autonomous microscopy control, offering a principled approach to accelerate discovery across multidisciplinary sciences.

stat.AP

CT-ESKF: A General Framework of Covariance Transformation-Based Error-State Kalman Filter

Invariant extended Kalman filter (InEKF) possesses excellent trajectory-independent property and better consistency compared to conventional extended Kalman filter (EKF). However, when applied to scenarios involving both global-frame and body-frame observations, InEKF may fail to preserve its trajectory-independent property. This work introduces the concept of equivalence between error states and covariance matrices among different error-state Kalman filters, and shows that although InEKF exhibits trajectory independence, its covariance propagation is actually equivalent to EKF. A covariance transformation-based error-state Kalman filter (CT-ESKF) framework is proposed that unifies various error-state Kalman filtering algorithms. The framework gives birth to novel filtering algorithms that demonstrate improved performance in integrated navigation systems that incorporate both global and body-frame observations. Experimental results show that the EKF with covariance transformation outperforms both InEKF and original EKF in a representative INS/GNSS/Odometer integrated navigation system.

eess.SY

A Multi-view Landmark Representation Approach with Application to GNSS-Visual-Inertial Odometry

Invariant Extended Kalman Filter (IEKF) has been a significant technique in vision-aided sensor fusion. However, it usually suffers from high computational burden when jointly optimizing camera poses and the landmarks. To improve its efficiency and applicability for multi-sensor fusion, we present a multi-view pose-only estimation approach with its application to GNSS-Visual-Inertial Odometry (GVIO) in this paper. Our main contribution is deriving a visual measurement model which directly associates landmark representation with multiple camera poses and observations. Such a pose-only measurement is proven to be tightly-coupled between landmarks and poses, and maintain a perfect null space that is independent of estimated poses. Finally, we apply the proposed approach to a filter based GVIO with a novel feature management strategy. Both simulation tests and real-world experiments are conducted to demonstrate the superiority of the proposed method in terms of efficiency and accuracy.

cs.RO

On second-order weak sharp minima of general nonconvex set-constrained optimization problems

This paper explores local second-order weak sharp minima for a broad class of nonconvex optimization problems. We propose novel second-order optimality conditions formulated through the use of classical and lower generalized support functions. These results are based on asymptotic second-order tangent cones and outer second-order tangent sets. Specifically, our findings eliminate the necessity of assuming convexity in the constraint set and/or the outer second-order tangent set, or the nonemptiness of the outer second-order tangent set. Furthermore, unlike traditional approaches, our sufficient conditions do not rely on strong assumptions such as the uniform second-order regularity of the constraint set and the property of uniform approximation of the critical cones.

math.OC

Inhomogeneous plane waves in attenuative anisotropic porous media

We investigate the propagation of inhomogeneous plane waves in poro-viscoelastic media, explicitly incorporating both velocity and attenuation anisotropy. Starting from classical Biot theory, we present a fractional differential equation describing wave propagation in attenuative anisotropic porous media that accommodates arbitrary anisotropy in both velocity and attenuation. Then, instead of relying on the traditional complex wave vector approach, we derive new Christoffel and energy balance equations for general inhomogeneous waves by employing an alternative formulation based on the complex slowness vector. The phase velocities and complex slownesses of inhomogeneous fast and slow quasi-compressional (qP1 and qP2) and quasi-shear (qS1 and qS2) waves are determined by solving an eighth-degree algebraic equation. By invoking the derived energy balance equation along with the computed complex slowness, we present explicit and concise expressions for energy velocities. Additionally, we analyze dissipation factors defined by two alternative measures: the ratio of average dissipated energy density to either average strain energy density or average stored energy density. We clarify and discuss the implications of these definitional differences in the context of general poro-viscoelastic anisotropic media. Finally, our expressions are degenerated to give their counterparts of the homogeneous waves as a special case, and the reduced forms are identical to those presented by the existing poro-viscoelastic theory. Several examples are provided to illustrate the propagation characteristics of inhomogeneous plane waves in unbounded attenuative vertical transversely isotropic porous media.

physics.geo-ph

Adapting Video Diffusion Models for Time-Lapse Microscopy

We present a domain adaptation of video diffusion models to generate highly realistic time-lapse microscopy videos of cell division in HeLa cells. Although state-of-the-art generative video models have advanced significantly for natural videos, they remain underexplored in microscopy domains. To address this gap, we fine-tune a pretrained video diffusion model on microscopy-specific sequences, exploring three conditioning strategies: (1) text prompts derived from numeric phenotypic measurements (e.g., proliferation rates, migration speeds, cell-death frequencies), (2) direct numeric embeddings of phenotype scores, and (3) image-conditioned generation, where an initial microscopy frame is extended into a complete video sequence. Evaluation using biologically meaningful morphological, proliferation, and migration metrics demonstrates that fine-tuning substantially improves realism and accurately captures critical cellular behaviors such as mitosis and migration. Notably, the fine-tuned model also generalizes beyond the training horizon, generating coherent cell dynamics even in extended sequences. However, precisely controlling specific phenotypic characteristics remains challenging, highlighting opportunities for future work to enhance conditioning methods. Our results demonstrate the potential for domain-specific fine-tuning of generative video models to produce biologically plausible synthetic microscopy data, supporting applications such as in-silico hypothesis testing and data augmentation.

cs.CV

New second-order optimality conditions for directional optimality of a general set-constrained optimization problem

In this paper we derive new second-order optimality conditions for a very general set-constrained optimization problem where the underlying set may be nononvex. We consider local optimality in specific directions (i.e., optimal in a directional neighborhood) in pursuit of developing these new optimality conditions. First-order necessary conditions for local optimality in given directions are provided by virtue of the corresponding directional normal cones. Utilizing the classical and/or the lower generalized support function, we obtain new second-order necessary and sufficient conditions for local optimality of general nonconvex constrained optimization problem in given directions via both the corresponding asymptotic second-order tangent cone and outer second-order tangent set. Our results do not require convexity and/or nonemptyness of the outer second-order tangent set. This is an important improvement to other results in the literature since the outer second-order tangent set can be nonconvex and empty even when the set is convex.

math.OC

An Immediate Update Strategy of Multi-State Constraint Kalman Filter

The lightweight Multi-state Constraint Kalman Filter (MSCKF) has been well-known for its high efficiency, in which the delayed update has been usually adopted since its proposal. This work investigates the immediate update strategy of MSCKF based on timely reconstructed 3D feature points and measurement constraints. The differences between the delayed update and the immediate update are theoretically analyzed in detail. It is found that the immediate update helps construct more observation constraints and employ more filtering updates than the delayed update, which improves the linearization point of the measurement model and therefore enhances the estimation accuracy. Numerical simulations and experiments show that the immediate update strategy significantly enhances MSCKF even with a small amount of feature observations.

cs.RO

Generative Diffusion Model for Seismic Imaging Improvement of Sparsely Acquired Data and Uncertainty Quantification

Seismic imaging from sparsely acquired data faces challenges such as low image quality, discontinuities, and migration swing artifacts. Existing convolutional neural network (CNN)-based methods struggle with complex feature distributions and cannot effectively assess uncertainty, making it hard to evaluate the reliability of their processed results. To address these issues, we propose a new method using a generative diffusion model (GDM). Here, in the training phase, we use the imaging results from sparse data as conditional input, combined with noisy versions of dense data imaging results, for the network to predict the added noise. After training, the network can predict the imaging results for test images from sparse data acquisition, using the generative process with conditional control. This GDM not only improves image quality and removes artifacts caused by sparse data, but also naturally evaluates uncertainty by leveraging the probabilistic nature of the GDM. To overcome the decline in generation quality and the memory burden of large-scale images, we develop a patch fusion strategy that effectively addresses these issues. Synthetic and field data examples demonstrate that our method significantly enhances imaging quality and provides effective uncertainty quantification.

physics.geo-ph

BioImage.IO Chatbot: A Community-Driven AI Assistant for Integrative Computational Bioimaging

We present the BioImage$.$IO Chatbot, an AI assistant powered by Large Language Models and supported by a community-driven knowledge base and toolset. This chatbot is designed to cater to a wide range of user needs through a flexible extension mechanism that spans from information retrieval to AI-enhanced analysis and microscopy control. Embracing open-source principles, the chatbot is designed to evolve through community contributions. By simplifying navigation through the intricate bioimaging landscape, the BioImage$.$IO Chatbot empowers life sciences to progress by leveraging the collective expertise and innovation of its users.

cs.AI

Clifford Algebra-Based Iterated Extended Kalman Filter with Application to Low-Cost INS/GNSS Navigation

The traditional GNSS-aided inertial navigation system (INS) usually exploits the extended Kalman filter (EKF) for state estimation, and the initial attitude accuracy is key to the filtering performance. To spare the reliance on the initial attitude, this work generalizes the previously proposed trident quaternion within the framework of Clifford algebra to represent the extended pose, IMU biases and lever arms on the Lie group. Consequently, a quasi-group-affine system is established for the low-cost INS/GNSS integrated navigation system, and the right-error Clifford algebra-based EKF (Clifford-RQEKF) is accordingly developed. The iterated filtering approach is further applied to significantly improve the performances of the Clifford-RQEKF and the previously proposed trident quaternion-based EKFs. Numerical simulations and experiments show that all iterated filtering approaches fulfill the fast and global convergence without the prior attitude information, whereas the iterated Clifford-RQEKF performs much better than the others under especially large IMU biases.

eess.SY

JDLL: A library to run Deep Learning models on Java bioimage informatics platforms

We present JDLL, an agile Java library that offers a comprehensive toolset/API to unify the development of high-end applications of DL for bioimage analysis and to streamline their installation and maintenance. JDLL provides all the functions required to consume DL models seamlessly, without being burdened by the configuration of the Python-based DL frameworks, within Java bioimage informatics platforms. Moreover, it allows the deployment of pre-trained models in the Bioimage Model Zoo (BMZ) by shipping the logic to connect to the BMZ website, download and run a selected model inference.

eess.IV

Roadmap on Deep Learning for Microscopy

Through digital imaging, microscopy has evolved from primarily being a means for visual observation of life at the micro- and nano-scale, to a quantitative tool with ever-increasing resolution and throughput. Artificial intelligence, deep neural networks, and machine learning are all niche terms describing computational methods that have gained a pivotal role in microscopy-based research over the past decade. This Roadmap is written collectively by prominent researchers and encompasses selected aspects of how machine learning is applied to microscopy image data, with the aim of gaining scientific knowledge by improved image quality, automated detection, segmentation, classification and tracking of objects, and efficient merging of information from multiple imaging modalities. We aim to give the reader an overview of the key developments and an understanding of possibilities and limitations of machine learning for microscopy. It will be of interest to a wide cross-disciplinary audience in the physical sciences and life sciences.

physics.optics

A Trident Quaternion Framework for Inertial-based Navigation Part II: Error Models and Application to Initial Alignment

This work deals with error models for trident quaternion framework proposed in the companion paper (Part I) and further uses them to investigate the odometer-aided static/in-motion inertial navigation attitude alignment for land vehicles. By linearizing the trident quaternion kinematic equation, the left and right trident quaternion error models are obtained, which are found to be equivalent to those derived from profound group affine. The two error models are used to design their corresponding extended Kalman filters (EKF), namely, the left-quaternion EKF (LQEKF) and the right-quaternion EKF (RQEKF). Simulations and field tests are conducted to evaluate their actual performances. Owing to the high estimation consistency, the L/RQEKF converge much faster in the static alignment than the traditional error model-based EKF, even under arbitrary large heading initialization. For the in-motion alignment, the L/RQEKF possess much larger convergence region than the traditional EKF does, although they still require the aid of attitude initialization so as to avoid large initial attitude errors.

cs.RO