SearcharxivSearch

arXiv subjects

Pengyu Zhou

Publications and source records attributed to Pengyu Zhou.

9 recordsLinked to original sources

CXR-LT 2026 Challenge: Multi-Center Long-Tailed and Zero Shot Chest X-ray Classification

Chest X-ray (CXR) interpretation is hindered by the long-tailed distribution of pathologies and the open-world nature of clinical environments. Existing benchmarks often rely on closed-set classes from a single institution, failing to capture the prevalence of rare diseases or the appearance of novel findings. To address this, we present the CXR-LT challenge. The first event, CXR-LT 2023, established a large-scale benchmark for long-tailed multi-label CXR classification and identified key challenges in rare disease recognition. CXR-LT 2024 further expanded the label space and introduced a zero-shot task to study generalization to unseen findings. Building on the success of CXR-LT 2023 and 2024, this third iteration of the benchmark introduces a multi-center dataset comprising over 145,000 images from PadChest and NIH Chest X-ray datasets. Additionally, all development and test sets in CXR-LT 2026 are annotated by radiologists, providing a more reliable and clinically grounded evaluation than report-derived labels. The challenge defines two core tasks this year: (1) Robust Multi-Label Classification on 30 known classes and (2) Open-World Generalization to 6 unseen (out-of-distribution) rare disease classes. This paper summarizes the overview of the CXR-LT 2026 challenge. We describe the data collection and annotation procedures, analyze solution strategies adopted by participating teams, and evaluate head-versus-tail performance, calibration, and cross-center generalization gaps. Our results show that vision-language foundation models improve both in-distribution and zero-shot performance, but detecting rare findings under multi-center shift remains challenging. Our study provides a foundation for developing and evaluating AI systems in realistic long-tailed and open-world clinical conditions.

cs.CV

Electrostatic quadrupole lens for focusing a cryogenic $^{205}$TlF molecular beam

Precision measurements with molecules require molecular beams with high flux and narrow divergence. Here we present the design, implementation, and characterization of an electrostatic quadrupole lens (EQL) for focusing a cryogenic beam of $^{205}$TlF molecules. The EQL consists of four electrodes operated at spatially alternating potentials up to $\pm 30$ kV. This configuration generates a transverse restoring force that focuses molecules prepared in selected manifold of hyperfine states. At the optimal applied voltage, the EQL increases the detected molecular signal by a factor of $14.4(4)$ at the current detection position. The measured lens gain, transverse Doppler spectra, and transverse spatial profiles are all reproduced by trajectory simulations using independently measured beam parameters as input. Based on this agreement, we project a gain of $16.5(8)$ for the full $7.9~\mathrm{m}$ CeNTREX beamline. These results establish the EQL as a key component for increasing sensitivity in precision experiments using cryogenic beams of polar molecules.

physics.atom-ph

A Fully GPU-Accelerated Framework for High-Performance Configuration Interaction Selection with Neural Network Quantum States

AI-driven methods have demonstrated considerable success in tackling the central challenge of accurately solving the Schrödinger equation for complex many-body systems. Among neural network quantum state (NNQS) approaches, the NNQS-SCI (Selected Configuration Interaction) method stands out as a state-of-the-art technique, recognized for its high accuracy and scalability. However, its application to larger systems is severely constrained by a hybrid CPU-GPU architecture. Specifically, centralized CPU-based global de-duplication creates a severe scalability barrier due to communication bottlenecks, while host-resident coupled-configuration generation induces prohibitive computational overheads. We introduce QiankunNet-cuSCI, a fully GPU-accelerated SCI framework designed to overcome these bottlenecks. It first integrates a distributed, load-balanced global de-duplication algorithm to minimize redundancy and communication overhead at scale. To address compute limitations, it employs specialized, fine-grained CUDA kernels for exact coupled configuration generation. Finally, to break the single-GPU memory barrier exposed by this full acceleration, it incorporates a GPU memory-centric runtime featuring GPU-side pooling, streaming mini-batches, and overlapped offloading. This design enables much larger configuration spaces and shifts the bottleneck from host-side limitations back to on-device inference. Our evaluation demonstrates that our work fundamentally expands the scale of solvable problems. On an NVIDIA A100 cluster with 64 GPUs, our work achieves up to 2.32X end-to-end speedup over the highly-optimized NNQS-SCI baseline while preserving the same chemical accuracy. Furthermore, it demonstrates excellent distributed performance, maintaining over 90% parallel efficiency in strong scaling tests.

cs.DC

Overview of the CXR-LT 2026 Challenge: Multi-Center Long-Tailed and Zero Shot Chest X-ray Classification

Chest X-ray (CXR) interpretation is hindered by the long-tailed distribution of pathologies and the open-world nature of clinical environments. Existing benchmarks often rely on closed-set classes from single institutions, failing to capture the prevalence of rare diseases or the appearance of novel findings. To address this, we present the CXR-LT 2026 challenge. This third iteration of the benchmark introduces a multi-center dataset comprising over 145,000 images from PadChest and NIH Chest X-ray datasets. The challenge defines two core tasks: (1) Robust Multi-Label Classification on 30 known classes and (2) Open-World Generalization to 6 unseen (out-of-distribution) rare disease classes. We report the results of the top-performing teams, evaluating them via mean Average Precision (mAP), AUROC, and F1-score. The winning solutions achieved an mAP of 0.5854 on Task 1 and 0.4315 on Task 2, demonstrating that large-scale vision-language pre-training significantly mitigates the performance drop typically associated with zero-shot diagnosis.

cs.CV

Global Dynamics Of Quadratic And Cubic Planar Quasi-homogeneous Differential Systems

In this paper we obtain the global dynamics and phase portraits of quadratic and cubic quasi-homogeneous but non-homogeneous systems. We first prove that all planar quadratic and cubic quasi-homogeneous but non-homogeneous polynomial systems can be reduced to three homogeneous ones. Then for the homogeneous systems, we employ blow-up method, normal sector method, Poincaré compactification and other techniques to discuss their dynamics. Finally we characterize the global phase portraits of quadratic and cubic quasi-homogeneous but non-homogeneous polynomial systems.

math.DS

FaaSKeeper: Learning from Building Serverless Services with ZooKeeper as an Example

FaaS (Function-as-a-Service) revolutionized cloud computing by replacing persistent virtual machines with dynamically allocated resources. This shift trades locality and statefulness for a pay-as-you-go model more suited to variable and infrequent workloads. However, the main challenge is to adapt services to the serverless paradigm while meeting functional, performance, and consistency requirements. In this work, we push the boundaries of FaaS computing by designing a serverless variant of ZooKeeper, a centralized coordination service with a safe and wait-free consensus mechanism. We define synchronization primitives to extend the capabilities of scalable cloud storage and outline a set of requirements for efficient computing with serverless. In FaaSKeeper, the first coordination service built on serverless functions and cloud-native services, we explore the limitations of serverless offerings and propose improvements essential for complex and latency-sensitive applications. We share serverless design lessons based on our experiences of implementing a ZooKeeper model deployable to clouds today. FaaSKeeper maintains the same consistency guarantees and interface as ZooKeeper, with a serverless price model that lowers costs up to 110-719x on infrequent workloads.

cs.DC

NNQS-Transformer: an Efficient and Scalable Neural Network Quantum States Approach for Ab initio Quantum Chemistry

Neural network quantum state (NNQS) has emerged as a promising candidate for quantum many-body problems, but its practical applications are often hindered by the high cost of sampling and local energy calculation. We develop a high-performance NNQS method for \textit{ab initio} electronic structure calculations. The major innovations include: (1) A transformer based architecture as the quantum wave function ansatz; (2) A data-centric parallelization scheme for the variational Monte Carlo (VMC) algorithm which preserves data locality and well adapts for different computing architectures; (3) A parallel batch sampling strategy which reduces the sampling cost and achieves good load balance; (4) A parallel local energy evaluation scheme which is both memory and computationally efficient; (5) Study of real chemical systems demonstrates both the superior accuracy of our method compared to state-of-the-art and the strong and weak scalability for large molecular systems with up to $120$ spin orbitals.

quant-ph

Fast shimming algorithm based on Bayesian optimization for magnetic resonance based dark matter search

The sensitivity and accessible mass range of magnetic resonance searches for axionlike dark matter depends on the homogeneity of applied magnetic fields. Optimizing homogeneity through shimming requires exploring a large parameter space which can be prohibitively time consuming. We have automated the process of tuning the shim-coil currents by employing an algorithm based on Bayesian optimization. This method is especially suited for applications where the duration of a single optimization step prohibits exploring the parameter space extensively or when there is no prior information on the optimal operation point. Using the Cosmic Axion Spin Precession Experiment (CASPEr)-gradient low-field apparatus, we show that for our setup this method converges after approximately 30 iterations to a sub-10 parts-per-million field homogeneity which is desirable for our dark matter search.

astro-ph.CO

Charm and beauty isolation from heavy flavor decay electrons in p+p and Pb+Pb collisions at $\sqrt{s_{\mathrm{NN}}}$ = 5.02 TeV at LHC

We present an analysis on the heavy flavor hadron decay electrons with charm and beauty contributions decomposed via a data driven method in p+p and Pb+Pb collisions at $\sqrt{s_{\mathrm{NN}}}$ = 5.02 TeV at LHC. The transverse momentum $p_{\mathrm{T}}$ spectra, nuclear modification factor $R_{\mathrm{AA}}$ and azimuthal anisotropic flow $v_2$ distributions of electrons from charm and beauty decays are obtained. We find that the electron $R_{\mathrm{AA}}$ from charm ($R_{\mathrm{AA}}^{\mathrm{c\rightarrow e}}$) and beauty ($R_{\mathrm{AA}}^{\mathrm{b\rightarrow e}}$) decays are suppressed at $p_{\mathrm{T}}$ $>$ 2.0 and $p_{\mathrm{T}}$ $>$ 3.0 GeV/$c$ in Pb+Pb collisions, respectively, which indicates that charm and beauty interact with and lose their energy in the hot-dense medium. A less suppression of electron $R_{\mathrm{AA}}$ from beauty decays than that from charm decays at 2.0 $<$ $p_{\mathrm{T}}$ $<$ 8.0 GeV/$c$ is observed, which is consistent with the mass-dependent partonic energy loss scenario. A non-zero electron $v_2$ from beauty decays ($v_{2}^{\mathrm{b\rightarrow e}}$) is observed and in good agreement with ALICE measurement. At low $p_{\mathrm{T}}$ region from 1.0 to 3.0 GeV/$c$, a discrepancy between RHIC and LHC results is observed with 68\% confidence level, which suggests different degree of thermalization of beauty quark under different temperatures of the medium. At 3.0 GeV/$c$ $<$ $p_{\mathrm{T}}$ $<$ 7.0 GeV/$c$, $v_{2}^{\mathrm{b\rightarrow e}}$ deviates from a number-of-constituent-quark (NCQ) scaling hypothesis, which favors that beauty quark is unlikely thermalized in heavy-ion collisions at LHC energy.

nucl-ex