A fully GPU-based workflow for building physics emulators of hypersonic flows
arXiv:2606.13742v1 Announce Type: new
Abstract: The ability to resolve complex physical phenomena with high fidelity and at low computational cost is central to addressing key challenges in modern engineering. A prime example lies in hypersonic flows, where the precise prediction of the full flowfi...
Efficient On-Device Diffusion LLM Inference with Mobile NPU
arXiv:2606.13740v1 Announce Type: new
Abstract: Diffusion large language models (dLLMs) accelerate generation by denoising multiple tokens in parallel, making them attractive for latency-sensitive mobile inference. However, repeated denoising introduces substantial computation on smartphones. Mobil...
FedSPC: Shared Parameter Correction for Personalized Federated Learning
arXiv:2606.13748v1 Announce Type: new
Abstract: Personalized federated learning (PFL) is one of the important approaches in federated learning for addressing statistical heterogeneity while enabling client-specific adaptation. Many PFL methods split the model into shared and personalized parameters...
arXiv:2606.13741v1 Announce Type: new
Abstract: This paper presents the design, development, and implementation of a specialized forecast-then-optimize algorithmic pricing tool for sales campaigns in fashion e-commerce. Sales events present unique challenges for pricing including volatile demand pa...
Can Editing 1 Neuron Fix Repetition Loops in LLMs?
arXiv:2606.13705v1 Announce Type: new
Abstract: Yes. Can it cure doom loops? Probably not.
The Gemma 4 instruction-tuned models share a reproducible failure: on long factual enumeration prompts, such as listing every episode of a TV series, the 88 IAU constellations, or the 151 original Pokemon, ...
arXiv:2606.13707v1 Announce Type: new
Abstract: The recent success of agent swarms has shifted the paradigm of large language model (LLM)-based agents from single-agent workflows to multi-agent systems, highlighting the importance of agent orchestration for task decomposition and collaboration. How...
A Deep Reinforcement Learning (DRL)-Based Transformer Method for Solving the Open Shop Scheduling Problem
arXiv:2606.13682v1 Announce Type: new
Abstract: The open shop scheduling problem (OSSP) arises in many industrial and service settings but remains computationally challenging as the number of jobs and machines increases. While exact methods quickly become intractable, classical dispatching rules an...
UP-NRPA: User Portrait based Nested Rollout Policy Adaptation for Planning with Large Language Models in Goal-oriented Dialogue Systems
arXiv:2606.13683v1 Announce Type: new
Abstract: To address the challenge that current dialogue policy planning methods struggle to dynamically adapt to diverse user characteristics, this paper proposes a User Portrait based Nested Rollout Policy Adaptation (UP-NRPA) online framework with Large Lang...
Hybrid Open-Ended Tri-Evolution Makes Better Deep Researcher
arXiv:2606.13710v1 Announce Type: new
Abstract: Deep research and agent evolution serve as de-facto tasks for AI agents in real-world applications toward artificial general intelligence. The former enables autonomous retrieval and integration of information in open-ended environments to tackle open...
Pythagoras-Prover: Advancing Efficient Formal Proving via Augmented Lean Formalisation
arXiv:2606.12594v1 Announce Type: new
Abstract: Modern Lean theorem provers achieve strong performance only with substantial training and inference compute, driven in part by scarce verified proof data and the long reasoning traces of formal proof search, making both supervised fine-tuning (SFT) an...
Arbor: Tree Search as a Cognition Layer for Autonomous Agents
arXiv:2606.12563v1 Announce Type: new
Abstract: Arbor is a multi-agent framework that introduces structured tree search as a cognition layer for autonomous agents operating in large, stateful action spaces. Prior autonomous optimization systems operate on isolated targets with stateless evaluation....
PersonaDrive: Human-Style Retrieval-Augmented VLA Agents for Closed-Loop Driving Simulation
arXiv:2606.12616v1 Announce Type: new
Abstract: Closed-loop driving simulators typically populate their environments with non-ego traffic agents that behave largely the same way, produced either by rule-based traffic managers or by learned models trained toward a single behavioral mode. Recent work...
arXiv:2606.12587v1 Announce Type: new
Abstract: Traditionally, decision support studies how humans use machine learning models to make better decisions. In modern agentic systems, this division of roles is increasingly reversed: AI agents act on behalf of users, while humans and tools becomes suppo...
ToolSense: A Diagnostic Framework for Auditing Parametric Tool Knowledge in LLMs
arXiv:2606.12451v1 Announce Type: new
Abstract: Large language models deployed as agents over large tool catalogs face a critical tool-retrieval bottleneck. As embedding-based retrieval approaches rely on compact encoders that may under-capture specialized tool semantics, parametric tool retrieval ...
ProHiFlo: Hierarchical Flow Matching with Functional Guidance for De Novo Protein Generation
arXiv:2606.11243v1 Announce Type: new
Abstract: De novo protein generation has transformative potential in therapeutic design, enzyme engineering, and synthetic biology. While diffusion-based and flow matching approaches have achieved progress, they typically operate at single resolution and lack m...
Restless bandits with imperfect binary feedback: PCL-indexability analysis and computation
arXiv:2606.11192v1 Announce Type: new
Abstract: We study restless bandits with binary latent states and imperfect binary feedback, motivated by opportunistic spectrum access with sensing errors. For the associated belief-state model, we develop a partial conservation laws (PCL)-based analytical and...
To Intervene or Not: Guiding Inference-time Alignment with Probabilistic Model Blending
arXiv:2606.11201v1 Announce Type: new
Abstract: The wide deployment of LLMs has made model alignment necessary to make newly trained models safely and effectively respond to user instructions. Among different methods, inference-time alignment is often cheaper as it intervenes (i.e., offers guidance...
Dual-Stance Evaluation of Sycophancy: The Structure of Agreement and the Limits of Intervention
arXiv:2606.11205v1 Announce Type: new
Abstract: Activation steering can shift LLM behaviour, but standard evaluations do not typically test whether a sycophancy-reduction direction also suppresses agreement with factually correct statements. We introduce dual-stance evaluation, which tests both sta...
Position: Hippocampal Explicit Memory Is the Cornerstone for AGI
arXiv:2606.11245v1 Announce Type: new
Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities across various tasks, raising expectations for Artificial General Intelligence (AGI). This position paper argues that integrating explicit memory is the cornerstone for advancing L...
From Explicit Elements to Implicit Intent: A Predefined Library for Auditable Behavioral Inference
arXiv:2606.11207v1 Announce Type: new
Abstract: We present SemantiClean, a modular framework for extracting structured semantic signals from e-commerce session data and driving pluggable inference targets including purchase intent, customer segmentation, and product affinity through a shared elemen...
Automated Mediator for Human Negotiation: Pre-Mediation via a Structured LLM Pipeline
arXiv:2606.11379v1 Announce Type: new
Abstract: Pre-mediation, the preparatory phase preceding direct human negotiation, plays a critical role in achieving mutually beneficial agreements, yet is often omitted due to cost, time, and limited access to trained mediators. We introduce an automated medi...
Knowing When to Ask: Self-Gated Clarification for Hierarchical Language Agents
arXiv:2606.11349v1 Announce Type: new
Abstract: In hierarchical reasoning, failures often originate at intermediate decision points where the agent commits to a wrong branch without recognizing that it lacks critical information. Rather than treating clarification as an external uncertainty trigger...
arXiv:2606.11337v1 Announce Type: new
Abstract: Scientific AI agents increasingly retrieve evidence, reason across sources, and synthesize conclusions used in consequential decisions. Yet, their ability to do so in high-stakes domains such as health remains unclear. We introduce SciConBench, a larg...
Mechanistic Analysis of Alignment Algorithms in Language Models
arXiv:2606.09850v1 Announce Type: new
Abstract: Post-training alignment algorithms are predominantly evaluated as black boxes, obscuring how they reshape language models' internal computations. We present a systematic mechanistic analysis of six preference-optimization methods: PPO, DPO, SimPO, ORP...