Searcharxiv⌕ Search

arXiv subjects

Yan Chang

Publications and source records attributed to Yan Chang.

31 records · Page 2Linked to original sources

ReMEmbR: Building and Reasoning Over Long-Horizon Spatio-Temporal Memory for Robot Navigation

Navigating and understanding complex environments over extended periods of time is a significant challenge for robots. People interacting with the robot may want to ask questions like where something happened, when it occurred, or how long ago it took place, which would require the robot to reason over a long history of their deployment. To address this problem, we introduce a Retrieval-augmented Memory for Embodied Robots, or ReMEmbR, a system designed for long-horizon video question answering for robot navigation. To evaluate ReMEmbR, we introduce the NaVQA dataset where we annotate spatial, temporal, and descriptive questions to long-horizon robot navigation videos. ReMEmbR employs a structured approach involving a memory building and a querying phase, leveraging temporal information, spatial information, and images to efficiently handle continuously growing robot histories. Our experiments demonstrate that ReMEmbR outperforms LLM and VLM baselines, allowing ReMEmbR to achieve effective long-horizon reasoning with low latency. Additionally, we deploy ReMEmbR on a robot and show that our approach can handle diverse queries. The dataset, code, videos, and other material can be found at the following link: https://nvidia-ai-iot.github.io/remembr

cs.RO↗

DualFluidNet: an Attention-based Dual-pipeline Network for FLuid Simulation

Fluid motion can be considered as a point cloud transformation when using the SPH method. Compared to traditional numerical analysis methods, using machine learning techniques to learn physics simulations can achieve near-accurate results, while significantly increasing efficiency. In this paper, we propose an innovative approach for 3D fluid simulations utilizing an Attention-based Dual-pipeline Network, which employs a dual-pipeline architecture, seamlessly integrated with an Attention-based Feature Fusion Module. Unlike previous methods, which often make difficult trade-offs between global fluid control and physical law constraints, we find a way to achieve a better balance between these two crucial aspects with a well-designed dual-pipeline approach. Additionally, we design a Type-aware Input Module to adaptively recognize particles of different types and perform feature fusion afterward, such that fluid-solid coupling issues can be better dealt with. Furthermore, we propose a new dataset, Tank3D, to further explore the network's ability to handle more complicated scenes. The experiments demonstrate that our approach not only attains a quantitative enhancement in various metrics, surpassing the state-of-the-art methods but also signifies a qualitative leap in neural network-based simulation by faithfully adhering to the physical laws. Code and video demonstrations are available at https://github.com/chenyu-xjtu/DualFluidNet.

cs.CV↗

Direct and Inverse scattering in a three-dimensional planar waveguide

In this paper, we study the direct and inverse scattering of the Schrödinger equation in a three-dimensional planar waveguide. For the direct problem, we derive a resonance-free region and resolvent estimates for the resolvent of the Schrödinger operator in such a geometry. Based on the analysis of the resolvent, several inverse problems are investigated. First, given the potential function, we prove the uniqueness of the inverse source problem with multi-frequency data. We also develop a Fourier-based method to reconstruct the source function. The capability of this method is numerically illustrated by examples. Second, the uniqueness and increased stability of an inverse potential problem from data generated by incident waves are achieved in the absence of the source function. To derive the stability estimate, we use an argument of quantitative analytic continuation in complex theory. Third, we prove the uniqueness of simultaneously determining the source and potential by active boundary data generated by incident waves. In these inverse problems, we only use the limited lateral Dirichlet boundary data at multiple wavenumbers within a finite interval.

math.AP↗

Inverse source problem of the biharmonic equation from multi-frequency phaseless data

This work deals with an inverse source problem for the biharmonic wave equation. A two-stage numerical method is proposed to identify the unknown source from the multi-frequency phaseless data. In the first stage, we introduce some artificially auxiliary point sources to the inverse source system and establish a phase retrieval formula. Theoretically, we point out that the phase can be uniquely determined and estimate the stability of this phase retrieval approach. Once the phase information is retrieved, the Fourier method is adopted to reconstruct the source function from the phased multi-frequency data. The proposed method is easy-to-implement and there is no forward solver involved in the reconstruction. Numerical experiments are conducted to verify the performance of the proposed method.

math.NA↗

A novel Newton method for inverse elastic scattering problems

This work is concerned with an inverse elastic scattering problem of identifying the unknown rigid obstacle embedded in an open space filled with a homogeneous and isotropic elastic medium. A Newton-type iteration method relying on the boundary condition is designed to identify the boundary curve of the obstacle. Based on the Helmholtz decomposition and the Fourier-Bessel expansion, we explicitly derive the approximate scattered field and its derivative on each iterative curve. Rigorous mathematical justifications for the proposed method are provided. Numerical examples are presented to verify the effectiveness of the proposed method.

math.NA↗

Mathematical and numerical study of an inverse source problem for the biharmonic wave equation

In this paper, we study the inverse source problem for the biharmonic wave equation. Mathematically, we characterize the radiating sources and non-radiating sources at a fixed wavenumber. We show that a general source can be decomposed into a radiating source and a non-radiating source. The radiating source can be uniquely determined by Dirichlet boundary measurements at a fixed wavenumber. Moreover, we derive a Lipschitz stability estimate for determining the radiating source. On the other hand, the non-radiating source does not produce any scattered fields outside the support of the source function. Numerically, we propose a novel source reconstruction method based on Fourier series expansion by multi-wavenumber boundary measurements. Numerical experiments are presented to verify the accuracy and efficiency of the proposed method.

math.NA↗

Jointly determining the point sources and obstacle from Cauchy data

A numerical method is developed for recovering both the source locations and the obstacle from the scattered Cauchy data of the time-harmonic acoustic field. First of all, the incident and scattered components are decomposed from the coupled Cauchy data by the representation of the single-layer potentials and the solution to the resulting linear integral system. As a consequence of this decomposition, the original problem of joint inversion is reformulated into two decoupled subproblems: an inverse source problem and an inverse obstacle scattering problem. Then, two sampling-type schemes are proposed to recover the shape of the obstacle and the source locations, respectively. The sampling methods rely on the specific indicator functions defined on target-oriented probing domains of circular shape. The error estimates of the decoupling procedure are established and the asymptotic behaviors of the indicator functions are analyzed. Extensive numerical experiments are also conducted to verify the performance of the sampling schemes.

math.NA↗

Recovering source location, polarization, and shape of obstacle from elastic scattering data

We consider an inverse elastic scattering problem of simultaneously reconstructing a rigid obstacle and the excitation sources using near-field measurements. A two-phase numerical method is proposed to achieve the co-inversion of multiple targets. In the first phase, we develop several indicator functionals to determine the source locations and the polarizations from the total field data, and then we manage to obtain the approximate scattered field. In this phase, only the inner products of the total field with the fundamental solutions are involved in the computation, and thus it is direct and computationally efficient. In the second phase, we propose an iteration method of Newton's type to reconstruct the shape of the obstacle from the approximate scattered field. Using the layer potential representations on an auxiliary curve inside the obstacle, the scattered field together with its derivative on each iteration surface can be easily derived. Theoretically, we establish the uniqueness of the co-inversion problem and analyze the indicating behavior of the sampling-type scheme. An explicit derivative is provided for the Newton-type method. Numerical results are presented to corroborate the effectiveness and efficiency of the proposed method.

math.NA↗

Co-inversion of a scattering cavity and its internal sources: uniqueness, decoupling and imaging

This paper concerns the simultaneous reconstruction of a sound-soft cavity and its excitation sources from the total-field data. Using the single-layer potential representations on two measurement curves, this co-inversion problem can be decoupled into two inverse problems: an inverse cavity scattering problem and an inverse source problem. This novel decoupling technique is fast and easy to implement since it is based on a linear system of integral equations. Then the uncoupled subproblems are respectively solved by the modified optimization and sampling method. We also establish the uniqueness of this co-inversion problem and analyze the stability of our method. Several numerical examples are presented to demonstrate the feasibility and effectiveness of the proposed method.

math.NA↗

Simultaneous recovery of an obstacle and its excitation sources from near-field scattering data

This paper is concerned with the inverse problem of determining an obstacle and the corresponding incident point sources in the Helmholtz equation from near-field scattering data. An optimization method is proposed to simultaneously recover both the obstacle and source locations. Moreover, a two-step sampling scheme with novel indicator functions is proposed to produce a good initial guess for solving the optimization problem. Theoretically, we analyze the convergence properties of the optimization method and the indicating behaviors of the indicator functions. Several numerical examples are presented to show the effectiveness of the proposed method.

math.AP↗

SafetyNet: Safe planning for real-world self-driving vehicles using machine-learned policies

In this paper we present the first safe system for full control of self-driving vehicles trained from human demonstrations and deployed in challenging, real-world, urban environments. Current industry-standard solutions use rule-based systems for planning. Although they perform reasonably well in common scenarios, the engineering complexity renders this approach incompatible with human-level performance. On the other hand, the performance of machine-learned (ML) planning solutions can be improved by simply adding more exemplar data. However, ML methods cannot offer safety guarantees and sometimes behave unpredictably. To combat this, our approach uses a simple yet effective rule-based fallback layer that performs sanity checks on an ML planner's decisions (e.g. avoiding collision, assuring physical feasibility). This allows us to leverage ML to handle complex situations while still assuring the safety, reducing ML planner-only collisions by 95%. We train our ML planner on 300 hours of expert driving demonstrations using imitation learning and deploy it along with the fallback layer in downtown San Francisco, where it takes complete control of a real vehicle and navigates a wide variety of challenging urban driving scenarios.

cs.RO↗

Energy Efficiency and Emission Testing for Connected and Automated Vehicles Using Real-World Driving Data

By using the onboard sensing and external connectivity technology, connected and automated vehicles (CAV) could lead to improved energy efficiency, better routing, and lower traffic congestion. With the rapid development of the technology and adaptation of CAV, it is more critical to develop the universal evaluation method and the testing standard which could evaluate the impacts on energy consumption and environmental pollution of CAV fairly, especially under the various traffic conditions. In this paper, we proposed a new method and framework to evaluate the energy efficiency and emission of the vehicle based on the unsupervised learning methods. Both the real-world driving data of the evaluated vehicle and the large naturalistic driving dataset are used to perform the driving primitive analysis and coupling. Then the linear weighted estimation method could be used to calculate the testing result of the evaluated vehicle. The results show that this method can successfully identify the typical driving primitives. The couples of the driving primitives from the evaluated vehicle and the typical driving primitives from the large real-world driving dataset coincide with each other very well. This new method could enhance the standard development of the energy efficiency and emission testing of CAV and other off-cycle credits.

cs.OH↗

An Optimal LiDAR Configuration Approach for Self-Driving Cars

LiDARs plays an important role in self-driving cars and its configuration such as the location placement for each LiDAR can influence object detection performance. This paper aims to investigate an optimal configuration that maximizes the utility of on-hand LiDARs. First, a perception model of LiDAR is built based on its physical attributes. Then a generalized optimization model is developed to find the optimal configuration, including the pitch angle, roll angle, and position of LiDARs. In order to fix the optimization issue with off-the-shelf solvers, we proposed a lattice-based approach by segmenting the LiDAR's range of interest into finite subspaces, thus turning the optimal configuration into a nonlinear optimization problem. A cylinder-based method is also proposed to approximate the objective function, thereby making the nonlinear optimization problem solvable. A series of simulations are conducted to validate our proposed method. This proposed approach to optimal LiDAR configuration can provide a guideline to researchers to maximize the utility of LiDARs.

cs.RO↗