SearcharxivSearch

arXiv subjects

Minh Dinh

Publications and source records attributed to Minh Dinh.

4 recordsLinked to original sources

AutoVerifier: An Agentic Automated Verification Framework Using Large Language Models

Scientific and Technical Intelligence (S&TI) analysis requires verifying complex technical claims across rapidly growing literature, where existing approaches fail to bridge the verification gap between surface-level accuracy and deeper methodological validity. We present AutoVerifier, an LLM-based agentic framework that automates end-to-end verification of technical claims without requiring domain expertise. AutoVerifier decomposes every technical assertion into structured claim triples of the form (Subject, Predicate, Object), constructing knowledge graphs that enable structured reasoning across six progressively enriching layers: corpus construction and ingestion, entity and claim extraction, intra-document verification, cross-source verification, external signal corroboration, and final hypothesis matrix generation. We demonstrate AutoVerifier on a contested quantum computing claim, where the framework, operated by analysts with no quantum expertise, automatically identified overclaims and metric inconsistencies within the target paper, traced cross-source contradictions, uncovered undisclosed commercial conflicts of interest, and produced a final assessment. These results show that structured LLM verification can reliably evaluate the validity and maturity of emerging technologies, turning raw technical documents into traceable, evidence-backed intelligence assessments.

cs.AI

Latent Equivariant Operators for Robust Object Recognition: Promises and Challenges

Despite the successes of deep learning in computer vision, difficulties persist in recognizing objects that have undergone group-symmetric transformations rarely seen during training$\unicode{x2013}$for example objects seen in unusual poses, scales, positions, or combinations thereof. Equivariant neural networks are a solution to the problem of generalizing across symmetric transformations, but require knowledge of transformations a priori. An alternative family of architectures proposes to learn equivariant operators in a latent space, from examples of symmetric transformations. Here, using simple datasets of rotated and translated noisy MNIST, we illustrate how such architectures can successfully be harnessed for out-of-distribution classification, thus overcoming the limitations of both traditional and equivariant networks. While conceptually enticing, we discuss challenges ahead on the path of scaling these architectures to more complex datasets. Our code is available at https://github.com/BRAIN-Aalto/equivariant_operator.

cs.CV

Adaptive Optical Multi-Spectral Matrix Approach for Label-free High-resolution Imaging through Complex Scattering Media

Imaging through complex scattering media is severely limited by aberrations and scattering which obscure images and reduce resolution. Confocal and temporal gatings partly filter out multiple scattering but are severely degraded by wavefront distortions. Adaptive optics restore resolution by correcting low-order aberrations and matrix-based imaging enables more complex wavefront corrections. However, they struggle to undo high-order aberrations under strong scattering, preventing imaging at greater depths. To address these challenges, we present Scattering Matrix Tomography (SMT), an approach that makes full use of the wavefront engineering capability of scattering matrix and extreme adaptive optics. SMT reformulates imaging through complex media as a numerical optimization and employs Zernike-mode wavefront regularization and coarse-to-fine nonconvex optimization strategy to reverse severe aberrations, enabling noninvasive high-resolution volumetric imaging in multiple scattering regime. Based on the spectrally-resolved matrix measurement, SMT achieves a depth-over-resolution ratio above 900 beneath $ex~vivo$ mouse brain tissue and volumetric imaging at over three transport mean free paths inside an opaque colloid, where conventional methods fail to correct strong aberrations under these challenging conditions. SMT is noninvasive, label-free, and works both inside and outside the scattering media, making it suitable for various applications, including medical imaging, biological science, device inspection, and colloidal physics.

physics.optics

CoughTrigger: Earbuds IMU Based Cough Detection Activator Using An Energy-efficient Sensitivity-prioritized Time Series Classifier

Persistent coughs are a major symptom of respiratory-related diseases. Increasing research attention has been paid to detecting coughs using wearables, especially during the COVID-19 pandemic. Among all types of sensors utilized, microphone is most widely used to detect coughs. However, the intense power consumption needed to process audio signals hinders continuous audio-based cough detection on battery-limited commercial wearable products, such as earbuds. We present CoughTrigger, which utilizes a lower-power sensor, an inertial measurement unit (IMU), in earbuds as a cough detection activator to trigger a higher-power sensor for audio processing and classification. It is able to run all-the-time as a standby service with minimal battery consumption and trigger the audio-based cough detection when a candidate cough is detected from IMU. Besides, the use of IMU brings the benefit of improved specificity of cough detection. Experiments are conducted on 45 subjects and our IMU-based model achieved 0.77 AUC score under leave one subject out evaluation. We also validated its effectiveness on free-living data and through on-device implementation.

cs.LG