SearcharxivSearch

arXiv subjects

Matthias Schultheis

Publications and source records attributed to Matthias Schultheis.

10 recordsLinked to original sources

ALMA Central Molecular Zone Exploration Survey (ACES) I: Overview

The mass flows and energy cycles within the inner regions of galaxies exert a powerful influence on the evolution of the galaxy population. The centre of the Milky Way is the only galactic nucleus for which it is possible to resolve the physical mechanisms that drive these cycles, namely star formation and feedback, while also tracing global (>100 pc) processes which determine where and when star formation and feedback occur. We present an overview of ACES, the 'Atacama Large Millimeter/submillimeter Array (ALMA) CMZ Exploration Survey', a ~1.5" angular resolution, 0.2-3 km/s spectral resolution ALMA Band 3 (85-102 GHz), survey of the 'Central Molecular Zone' (CMZ) -- the inner-100 pc of the Galaxy (l = 359.4 deg to 0.8 deg). ACES spectral setup is tuned to observe optimal tracers of the physical, chemical, and kinematic conditions in over 70 spectral features (e.g. HCO+, HNCO, SiO, H40alpha, complex molecules) of the gas in the CMZ, to derive the properties of all potentially star-forming Galactic Centre gas, from global scales (100 pc) to dense ~0.05 pc structures that are expected to host individual star-forming cores, down to sub-sonic (<0.4 km/s) velocity resolution. In this overview paper, we provide the scientific justification for the ACES survey, explain the choice of observational setup, and describe the data legacy products. Finally, we show some of the initial ACES data which highlight the power of ACES' combination of high angular resolution, unprecedented spatial dynamic range, sensitivity, spectral resolution and spectral bandwidth as an illustration of how ACES aims to understand how global processes set the location, intensity, and timescales for star formation and feedback in the CMZ.

astro-ph.GA

Probabilistic inverse optimal control for non-linear partially observable systems disentangles perceptual uncertainty and behavioral costs

Inverse optimal control can be used to characterize behavior in sequential decision-making tasks. Most existing work, however, is limited to fully observable or linear systems, or requires the action signals to be known. Here, we introduce a probabilistic approach to inverse optimal control for partially observable stochastic non-linear systems with unobserved action signals, which unifies previous approaches to inverse optimal control with maximum causal entropy formulations. Using an explicit model of the noise characteristics of the sensory and motor systems of the agent in conjunction with local linearization techniques, we derive an approximate likelihood function for the model parameters, which can be computed within a single forward pass. We present quantitative evaluations on stochastic and partially observable versions of two classic control tasks and two human behavioral tasks. Importantly, we show that our method can disentangle perceptual factors and behavioral costs despite the fact that epistemic and pragmatic actions are intertwined in sequential decision-making under uncertainty, such as in active sensing and active learning. The proposed method has broad applicability, ranging from imitation learning to sensorimotor neuroscience.

cs.LG

Reinforcement Learning with Non-Exponential Discounting

Commonly in reinforcement learning (RL), rewards are discounted over time using an exponential function to model time preference, thereby bounding the expected long-term reward. In contrast, in economics and psychology, it has been shown that humans often adopt a hyperbolic discounting scheme, which is optimal when a specific task termination time distribution is assumed. In this work, we propose a theory for continuous-time model-based reinforcement learning generalized to arbitrary discount functions. This formulation covers the case in which there is a non-exponential random termination time. We derive a Hamilton-Jacobi-Bellman (HJB) equation characterizing the optimal policy and describe how it can be solved using a collocation method, which uses deep learning for function approximation. Further, we show how the inverse RL problem can be approached, in which one tries to recover properties of the discount function given decision data. We validate the applicability of our proposed approach on two simulated problems. Our approach opens the way for the analysis of human discounting in sequential decision-making tasks.

cs.LG

Inverse Optimal Control Adapted to the Noise Characteristics of the Human Sensorimotor System

Computational level explanations based on optimal feedback control with signal-dependent noise have been able to account for a vast array of phenomena in human sensorimotor behavior. However, commonly a cost function needs to be assumed for a task and the optimality of human behavior is evaluated by comparing observed and predicted trajectories. Here, we introduce inverse optimal control with signal-dependent noise, which allows inferring the cost function from observed behavior. To do so, we formalize the problem as a partially observable Markov decision process and distinguish between the agent's and the experimenter's inference problems. Specifically, we derive a probabilistic formulation of the evolution of states and belief states and an approximation to the propagation equation in the linear-quadratic Gaussian problem with signal-dependent noise. We extend the model to the case of partial observability of state variables from the point of view of the experimenter. We show the feasibility of the approach through validation on synthetic data and application to experimental data. Our approach enables recovering the costs and benefits implicit in human sequential sensorimotor behavior, thereby reconciling normative and descriptive approaches in a computational framework.

cs.LG

POMDPs in Continuous Time and Discrete Spaces

Many processes, such as discrete event systems in engineering or population dynamics in biology, evolve in discrete space and continuous time. We consider the problem of optimal decision making in such discrete state and action space systems under partial observability. This places our work at the intersection of optimal filtering and optimal control. At the current state of research, a mathematical description for simultaneous decision making and filtering in continuous time with finite state and action spaces is still missing. In this paper, we give a mathematical description of a continuous-time partial observable Markov decision process (POMDP). By leveraging optimal filtering theory we derive a Hamilton-Jacobi-Bellman (HJB) type equation that characterizes the optimal solution. Using techniques from deep learning we approximately solve the resulting partial integro-differential equation. We present (i) an approach solving the decision problem offline by learning an approximation of the value function and (ii) an online algorithm which provides a solution in belief space using deep reinforcement learning. We show the applicability on a set of toy examples which pave the way for future methods providing solutions for high dimensional problems.

cs.LG

Probabilistic Trajectory Segmentation by Means of Hierarchical Dirichlet Process Switching Linear Dynamical Systems

Using movement primitive libraries is an effective means to enable robots to solve more complex tasks. In order to build these movement libraries, current algorithms require a prior segmentation of the demonstration trajectories. A promising approach is to model the trajectory as being generated by a set of Switching Linear Dynamical Systems and inferring a meaningful segmentation by inspecting the transition points characterized by the switching dynamics. With respect to the learning, a nonparametric Bayesian approach is employed utilizing a Gibbs sampler.

stat.ML

Receding Horizon Curiosity

Sample-efficient exploration is crucial not only for discovering rewarding experiences but also for adapting to environment changes in a task-agnostic fashion. A principled treatment of the problem of optimal input synthesis for system identification is provided within the framework of sequential Bayesian experimental design. In this paper, we present an effective trajectory-optimization-based approximate solution of this otherwise intractable problem that models optimal exploration in an unknown Markov decision process (MDP). By interleaving episodic exploration with Bayesian nonlinear system identification, our algorithm takes advantage of the inductive bias to explore in a directed manner, without assuming prior knowledge of the MDP. Empirical evaluations indicate a clear advantage of the proposed algorithm in terms of the rate of convergence and the final model fidelity when compared to intrinsic-motivation-based algorithms employing exploration bonuses such as prediction error and information gain. Moreover, our method maintains a computational advantage over a recent model-based active exploration (MAX) algorithm, by focusing on the information gain along trajectories instead of seeking a global exploration policy. A reference implementation of our algorithm and the conducted experiments is publicly available.

cs.LG

APOGEE Chemical Abundances of Globular Cluster Giants in the Inner Galaxy

We report chemical abundances obtained by SDSS-III/APOGEE for giant stars in five globular clusters located within 2.2 kpc of the Galactic centre. We detect the presence of multiple stellar populations in four of those clusters (NGC 6553, NGC 6528, Terzan 5, and Palomar 6) and find strong evidence for their presence in NGC 6522. All clusters present a significant spread in the abundances of N, C, Na, and Al, with the usual correlations and anti-correlations between various abundances seen in other globular clusters. Our results provide important quantitative constraints on theoretical models for self-enrichment of globular clusters, by testing their predictions for the dependence of yields of elements such as Na, N, C, and Al on metallicity. They also confirm that, under the assumption that field N-rich stars originate from globular cluster destruction, they can be used as tracers of their parental systems in the high- metallicity regime.

astro-ph.GA

Chemical tagging with APOGEE: Discovery of a large population of N-rich stars in the inner Galaxy

Formation of globular clusters (GCs), the Galactic bulge, or galaxy bulges in general, are important unsolved problems in Galactic astronomy. Homogeneous infrared observations of large samples of stars belonging to GCs and the Galactic bulge field are one of the best ways to study these problems. We report the discovery by APOGEE of a population of field stars in the inner Galaxy with abundances of N, C, and Al that are typically found in GC stars. The newly discovered stars have high [N/Fe], which is correlated with [Al/Fe] and anti-correlated with [C/Fe]. They are homogeneously distributed across, and kinematically indistinguishable from, other field stars in the same volume. Their metallicity distribution is seemingly unimodal, peaking at [Fe/H]~-1, thus being in disagreement with that of the Galactic GC system. Our results can be understood in terms of different scenarios. N-rich stars could be former members of dissolved GCs, in which case the mass in destroyed GCs exceeds that of the surviving GC system by a factor of ~8. In that scenario, the total mass contained in so-called "first-generation" stars cannot be larger than that in "second-generation" stars by more than a factor of ~9 and was certainly smaller. Conversely, our results may imply the absence of a mandatory genetic link between "second generation" stars and GCs. Last, but not least, N-rich stars could be the oldest stars in the Galaxy, the by-products of chemical enrichment by the first stellar generations formed in the heart of the Galaxy.

astro-ph.GA

The Apache Point Observatory Galactic Evolution Experiment (APOGEE)

The Apache Point Observatory Galactic Evolution Experiment (APOGEE), one of the programs in the Sloan Digital Sky Survey III (SDSS-III), has now completed its systematic, homogeneous spectroscopic survey sampling all major populations of the Milky Way. After a three year observing campaign on the Sloan 2.5-m Telescope, APOGEE has collected a half million high resolution (R~22,500), high S/N (>100), infrared (1.51-1.70 microns) spectra for 146,000 stars, with time series information via repeat visits to most of these stars. This paper describes the motivations for the survey and its overall design---hardware, field placement, target selection, operations---and gives an overview of these aspects as well as the data reduction, analysis and products. An index is also given to the complement of technical papers that describe various critical survey components in detail. Finally, we discuss the achieved survey performance and illustrate the variety of potential uses of the data products by way of a number of science demonstrations, which span from time series analysis of stellar spectral variations and radial velocity variations from stellar companions, to spatial maps of kinematics, metallicity and abundance patterns across the Galaxy and as a function of age, to new views of the interstellar medium, the chemistry of star clusters, and the discovery of rare stellar species. As part of SDSS-III Data Release 12, all of the APOGEE data products are now publicly available.

astro-ph.IM