SearcharxivSearch

arXiv subjects

Ricky Ho

Publications and source records attributed to Ricky Ho.

3 recordsLinked to original sources

Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Large language models (LLMs) have shown impressive capabilities, but still struggle with complex reasoning tasks requiring multiple steps. While prompt-based methods like Chain-of-Thought (CoT) can improve LLM reasoning at inference time, optimizing reasoning capabilities during training remains challenging. We introduce LaTent Reasoning Optimization (LaTRO), a principled framework that formulates reasoning as sampling from a latent distribution and optimizes it via variational approaches. LaTRO enables LLMs to concurrently improve both their reasoning process and ability to evaluate reasoning quality, without requiring external feedback or reward models. We validate LaTRO through experiments on GSM8K and ARC-Challenge datasets using multiple model architectures. On GSM8K, LaTRO improves zero-shot accuracy by an average of 12.5% over base models and 9.6% over supervised fine-tuning across Phi-3.5-mini, Mistral-7B, and Llama-3.1-8B. Our findings suggest that pre-trained LLMs possess latent reasoning capabilities that can be unlocked and enhanced through our proposed optimization approach in a self-improvement manner. The code of LaTRO is available at \url{https://github.com/SalesforceAIResearch/LaTRO}.

cs.AI

Scaling Knowledge Graph Construction through Synthetic Data Generation and Distillation

Document-level knowledge graph (KG) construction faces a fundamental scaling challenge: existing methods either rely on expensive large language models (LLMs), making them economically nonviable for large-scale corpora, or employ smaller models that produce incomplete and inconsistent graphs. We find that this limitation stems not from model capabilities but from insufficient training on high-quality document-level KG data. To address this gap, we introduce SynthKG, a multi-step data synthesis pipeline that generates high-quality document-KG pairs through systematic chunking, decontextualization, and structured extraction using LLMs. By fine-tuning a smaller LLM on synthesized document-KG pairs, we streamline the multi-step process into a single-step KG generation approach called Distill-SynthKG. Furthermore, we repurpose existing question-answering datasets to construct KG evaluation datasets and introduce new evaluation metrics. Using KGs produced by Distill-SynthKG, we also design a novel graph-based retrieval framework for RAG. Experimental results demonstrate that Distill-SynthKG not only surpasses all baseline models in KG quality (including models up to eight times larger) but also consistently improves in retrieval and question-answering tasks. Additionally, our proposed graph retrieval framework outperforms all KG-retrieval methods across multiple benchmark datasets.

cs.CL

High Efficiency, Low Cost, RF Sources for Accelerators and Colliders

Several high efficiency, low cost, RF sources are in development or recently completed. All are designed to provide operating efficiencies exceeding 80% and provide more than 100 kW of output power with a focus on high average power or CW operation. The sources include (1) a magnetron system with amplitude and phase control, a multiple beam, power grid-tube based source, a multiple beam inductive output tube, and a klystron using the core oscillation method. The estimated cost for the magnetron system and multiple beam power grid-tube source are one dollar per Watt and 75 cents per Watt, respectively. Operating frequencies span the range from 300 MHz (power grid tube) to 1.3 GHz (magnetron and klystron). This paper describes the basic operation of the sources, indicates the status and schedule, and provides available experimental results.

physics.acc-ph