SearcharxivSearch

arXiv subjects

Xianglin Ji

Publications and source records attributed to Xianglin Ji.

3 recordsLinked to original sources

ASI-Bench: At the Dawn of Artificial Superintelligence

Artificial superintelligence (ASI) requires AI to move beyond mastering existing knowledge toward exploring the unknown, creating new knowledge, and turning new ideas into verifiable results. However, the capabilities of today's AI systems are still largely built on learning, compressing, and applying existing human knowledge. Accordingly, existing benchmarks primarily test whether AI can produce correct answers based on learned knowledge, or whether it can complete tasks under extensive human guidance. We therefore introduce ASI-Bench, the first benchmark to jointly evaluate AI systems' capabilities of innovative exploration and autonomous scientific execution across general research domains, and the first to progressively withdraw human methodological guidance within the same research project to test how far AI can proceed on its own. Built by over 40 experts with the cost of 31,000+ human hours, ASI-Bench contains 60 project-level research tasks across 11 scientific domains and progressively reduces methodological guidance to test whether AI can independently select methods, conduct research, and produce verifiable results. All tasks undergo expert review, AI-assisted auditing, sandbox execution, and scorer validation. Across 18 state-of-the-art agent--model configurations, the average score drops from 50.91 with full methodological guidance to 29.10 with only the method specified and 26.62 when agents must determine the method themselves. This sharp decline shows that current systems remain heavily dependent on human guidance and are still far from autonomously conducting end-to-end, project-level scientific research. ASI-Bench is open to the world. We invite researchers and builders everywhere to contribute new tasks, challenge the limits of today's AI, and help accelerate humanity's collective path toward artificial superintelligence at https://asibench.apexin.ai/submit.

cs.AI

PHITSBench: an execution-scored benchmark for AI-assisted PHITS radiation-transport input generation using natural language

We introduce PHITSBench, an execution-scored benchmark for the Monte Carlo Particle and Heavy Ion Transport code System (PHITS). PHITSBench comprises 282 transport-scorable tasks spanning three common workflow categories: parameter editing (Edit), syntax repair (Repair ), and complete simulation generation from natural-language descriptions (Reproduce). Each task is evaluated using a Composite Metric Score that combines execution success with agreement between generated and reference transport observables. Using PHITSBench, we evaluate five GPT-5.4-based configurations ranging from zero-shot prompting to knowledge-augmented and agentic workflows. Without domain-specific knowledge, the model performs well on editing and repair tasks (95% and 70% success, respectively) but fails to generate correct simulations from scratch (0% success on the Reproduce track). A structured, machine-readable PHITS knowledge catalog, supplied alongside the user manual, raises single-shot Reproduce-task success to 57%. Agentic execution provides a further improvement to 66-73%, but at increased computational cost. Failure analysis shows that the remaining errors are dominated by incorrect selection and configuration of physical observables rather than syntax generation. These results suggest that future progress in AI-assisted radiation-transport modeling will depend as much on machine-readable knowledge bases, curated domain-training datasets, and execution-grounded evaluation environments as on advances in foundation models themselves.

cs.AI

Large Transverse Thermoelectric Effect in Weyl Semimetal TaIrTe$_4$ Engineered for Photodetection

Anomalous local photocurrent generation via second-order nonlinear and thermoelectric responses is a signature of many topological semimetals. The emergence of these photocurrents is inherently linked to symmetry breaking and anisotropy of their crystal lattices. Studies of type-II Weyl semimetals of group C$_{2v}$ (WTe$_2$, MoTe$_2$, TaIrTe$_4$) have reported anomalous, nonlocal photocurrents localized to crystals edges or far from electrodes, which are highly dependent on the geometry of the material sample. While originally attributed to a nonlinear charge current response, it was recently shown that these currents could instead be attributed to the anisotropic Seebeck coefficients of the materials. Here, we confirm that anomalous photocurrents observed in TaIrTe$_4$ under either visible or far-infrared far-field illumination originate from the large transverse thermoelectric effect. We engineer the mutual orientation of crystal edges and electrodes as well as the thermal environment of TaIrTe$_4$ to control and amplify its spatial photocurrent response. We show that substrate engineering can locally enhance photocurrent. This framework of thermal device engineering can enable broadband photo detection schemes by leveraging spectral and spatial dependence of photocurrents for applications like wavefront sensing, beam positioning, and edge detection.

cond-mat.mtrl-sci