Understanding LoRA Rank Trade-offs in Diffusion Model Fine-Tuning
arXiv:2609.10656v1 Announce Type: new
Abstract: Selecting LoRA rank for diffusion fine-tuning requires balancing quality and compute cost. We present a controlled study on CIFAR-10 using a DDPM U-Net with ranks {2,4,8,16,32}, fixed optimization settings, and a reproducible local-folder pytorch-fid ...
Probabilistic Focal Search: Accelerating Bounded-Suboptimal Search via Lower-Bound Advancement
arXiv:2609.10584v1 Announce Type: new
Abstract: Bounded-suboptimal search seeks a solution within a factor $w$ of optimal while reducing search effort. Focal Search (FS) uses heuristic guidance within FOCAL, the frontier nodes eligible under the threshold $w f_{\min}$, but its deterministic policy ...
A Multi-Stage Rule-Chaining Framework for Compositional and Interpretable Cognitive Reasoning
arXiv:2609.10654v1 Announce Type: new
Abstract: The Abstraction and Reasoning Corpus (ARC) benchmarks cognitive generalization, the ability to infer and apply abstract rules from limited examples. This paper presents a multi-stage rule-chaining framework that performs compositional reasoning across...
M3-Former: Multimodal Transformer with Mixture-of-Experts for Long-Term Vessel Trajectory Prediction
arXiv:2609.10559v1 Announce Type: new
Abstract: To address the challenges of behavioral multimodality, limited semantic utilization, and long-term error accumulation in vessel trajectory prediction, this paper proposes M3-Former, a multimodal trajectory prediction framework enhanced by large langua...
An Autonomous GeoAI Agent for Arctic Eco-Navigation
arXiv:2609.09374v1 Announce Type: new
Abstract: Arctic maritime navigation is becoming increasingly important as changing sea-ice conditions expand seasonal accessibility while simultaneously introducing substantial operational, environmental, and community risks. Arctic route planning is inherentl...
Adaptive Entangled Game Modules in Artificial General Intelligence
arXiv:2609.09226v1 Announce Type: new
Abstract: We introduce a probability-wave framework for modeling the collective behavior of interacting adaptive agents, deriving testable eigenmodes through a generalized behavioral intelligence (GBI) nonlocal probability-wave equation. This framework captures...
HB-PVI: A Hierarchical Bayesian Personalization and Value-of-Information Framework for Complex Activity Recognition
arXiv:2609.05582v1 Announce Type: new
Abstract: Personalization can improve activity-recognition performance, but participant-specific gains are heterogeneous, and every additional calibration label has an acquisition cost. This study presents HB-PVI, a hierarchical Bayesian personalization and val...
Capsule Lens: Locating and Tracking Concept Geometry in Model Representations
arXiv:2609.05575v1 Announce Type: new
Abstract: Understanding how concepts are encoded in the internal representations of machine learning models is a central problem in mechanistic interpretability, essential both for the science of deep learning and for the trustworthy deployment of increasingly ...
AhaBench: Do Agents Learn from Prior Experience? A Benchmark for Long-Horizon Continual Learning
arXiv:2609.05435v1 Announce Type: new
Abstract: Modern language agents are expected to operate over long horizons: they ask follow-up questions, reuse worked examples, handle tool feedback, and adapt to delayed consequences. Most evaluations still reset the agent after a prompt or score only the fi...
Beyond Right and Wrong: Evaluating Second-order Social Reasoning in Large Language Models
arXiv:2609.05437v1 Announce Type: new
Abstract: Previous AI alignment efforts have focused primarily on first-order social norms -- teaching models what is socially acceptable or unacceptable (e.g., `do not steal'). However, social intelligence depends not only on norm recognition, but also on anti...
When Does Memory Help? A Cost-Aware Evaluation of Long-Term Memory in Tool-Using LLM Agents
arXiv:2609.05441v1 Announce Type: new
Abstract: Long-term memory for LLM agents is evaluated today by conversational recall benchmarks (LoCoMo, LongMemEval), which measure question answering over dialogue history, not whether remembered facts change what a tool-using agent does. We present MERIT (M...
Damage-Aware Bandit Pruning for Vision and Language Transformers
arXiv:2609.05448v1 Announce Type: new
Abstract: Structured post-training pruning of transformers requires selecting complete functional units whose suppression causes limited degradation. We formulate structured-unit selection for language and vision transformers as a damage-aware multi-armed bandi...
CriticGen: Generation-Aware Evaluation as Actionable Feedback
arXiv:2609.05439v1 Announce Type: new
Abstract: Current evaluation methods for large language models are coarse-grained and decoupled from generation, producing generic explanations that fail to provide actionable feedback for model improvement. We propose CriticGen, a fine-grained, generation-awar...
arXiv:2609.04304v1 Announce Type: new
Abstract: We present Iris-mini and Iris-pro, two search agents trained at the 35B-A3B and 397B-A17B scales, together with the data pipeline and training recipe behind them. Tasks are reverse-constructed from the hyperlink structure of a web corpus: we author mu...
Evaluating Large Language Models for Forced Outage Risk Prediction: Benefits and Comparison to Machine Learning
arXiv:2609.04272v1 Announce Type: new
Abstract: This study examines the ability of large language models (LLMs) to predict the risk of weather-related forced outages in the distribution grid in a zero-shot framework, without labeled training data. The problem is formulated as a binary severity clas...
Quantum-Assisted Memory-Efficient Training for Parameter-Intensive Wi-Fi-Based Human Activity Recognition
arXiv:2609.04271v1 Announce Type: new
Abstract: Wi-Fi-based human activity recognition (HAR) has become an important part of integrated sensing and communications, paving the way for a range of context-aware services. However, most existing Wi-Fi-based HAR systems rely on deep learning (DL) models ...
Spectral-Target Physical Latent Structuring for JEPA-Style World Models
arXiv:2609.04264v1 Announce Type: new
Abstract: Latent world models have become increasingly popular as a method to predict and plan in latent space rather than pixel space. Recent architectures, such as LeWorldModel (LeWM), jointly train the encoder and predictor using regularization techniques li...
ProToMEx: Rapid, Interpretable Explanations via Structured Representations
arXiv:2609.04265v1 Announce Type: new
Abstract: Existing post-hoc explainers for machine learning classifiers primarily focus on feature attribution, assigning importance scores to individual features. While valuable, this approach struggles to articulate the complex, combinatorial patterns that of...
A Data Fusion Framework for Grounding Aerospace Surrogate Model via Experimental Wind-Tunnel Observations
arXiv:2609.04267v1 Announce Type: new
Abstract: Aerodynamic surrogate models trained on high-fidelity CFD data reproduce numerical predictions of both scalar outputs and entire fields accurately, yet their predictive fidelity is limited by systematic discrepancies between CFD and experimental obser...
Harbor Adapters and Harbor-Index: Infrastructure and a Curated Meta-Dataset for Large-Scale Agentic Evaluation
arXiv:2609.04298v1 Announce Type: new
Abstract: Evaluating agents on the growing number of agentic benchmarks is challenging because they often require complex environments and agent integrations. We introduce Harbor Adapters, a unified evaluation infrastructure for agentic benchmarks. Our work mak...
arXiv:2609.04239v1 Announce Type: new
Abstract: This technical report presents EXAONE Forecast for Finance (EXAONE Finance), a financial time series (TS) foundation model (TSFM) tailored to financial forecasting. Recent TSFMs achieve strong zero-shot performance through large-scale pretraining. How...
Equation Recast for Canonical Operator Learning Across Parametric PDEs
arXiv:2609.02982v1 Announce Type: new
Abstract: Learning solution operators across broad parameter ranges can require substantial coverage of both input functions and physical parameters, particularly for purely data-driven parametric models. In addition, the resulting models may fail silently outs...
From Euclidean to Graph-Structured Data: A Survey of Collaborative Learning
arXiv:2609.02984v1 Announce Type: new
Abstract: The conventional approach to machine learning, that is, collecting data, training models, and performing inference in a single location, faces fundamental limitations, including scalability and privacy, that restrict its applicability. To address thes...
The Geometry of Ignorance: LLMs Know When to Temper Bayesian Priors
arXiv:2609.02959v1 Announce Type: new
Abstract: What does a language model predict when it has few clues? The answer lurks in its unembedding geometry: a single direction of the unembedding matrix encodes the unigram distribution of the training corpus, which serves as the Bayesian prior the model ...