arXiv:2609.03209v1 Announce Type: new
Abstract: We study a governed approach to enterprise analytics: a language model interprets the question, while deterministic policy selects and runs a pre-approved analytical program that returns both results and evidence. We show that this restriction can rem...
From Euclidean to Graph-Structured Data: A Survey of Collaborative Learning
arXiv:2609.02984v1 Announce Type: new
Abstract: The conventional approach to machine learning, that is, collecting data, training models, and performing inference in a single location, faces fundamental limitations, including scalability and privacy, that restrict its applicability. To address thes...
The Geometry of Ignorance: LLMs Know When to Temper Bayesian Priors
arXiv:2609.02959v1 Announce Type: new
Abstract: What does a language model predict when it has few clues? The answer lurks in its unembedding geometry: a single direction of the unembedding matrix encodes the unigram distribution of the training corpus, which serves as the Bayesian prior the model ...
Equation Recast for Canonical Operator Learning Across Parametric PDEs
arXiv:2609.02982v1 Announce Type: new
Abstract: Learning solution operators across broad parameter ranges can require substantial coverage of both input functions and physical parameters, particularly for purely data-driven parametric models. In addition, the resulting models may fail silently outs...
WMLLM: Self-Evolving Optimization Agents via Predict-Then-Act World Modeling
arXiv:2609.01608v1 Announce Type: new
Abstract: Black-box optimization problems remain challenging because of large, weakly structured, and high-dimensional search spaces. Existing methods often suffer from poor sample efficiency because they rely on direct candidate generation or trial-and-error r...
Prompt-Space Meta-Learning Does Not Transfer Across Users: A Frozen-LLM Negative Result
arXiv:2609.01615v1 Announce Type: new
Abstract: Personalizing a frozen large language model (LLM) to individual users is often framed as a meta-learning problem in prompt space: each user is a task, and one seeks a shared natural-language adaptation policy that, given a handful of the user's labele...
CliffRank: A Dual-Branch Framework for Activity-Cliff Ranking Prediction
arXiv:2609.01673v1 Announce Type: new
Abstract: Activity-cliff ranking remains difficult because local structural changes can cause large activity differences, while high-quality data that resolve the underlying mechanisms remain limited. To use available activity labels more effectively, we combin...
When Does Information Sharing Improve Decentralized Discovery? Aggregation, Independent Rescue, and Equilibrium Selection
arXiv:2609.01814v1 Announce Type: new
Abstract: Information sharing can improve a pooled estimate while eliminating independent rescue actions. This paper separates those effects in exact finite discovery models. A centralized action-budget profile shows that equal one-person accuracy can coexist w...
Induction and Inquiry via Probabilistic Reasoning over Language and Code
arXiv:2609.01815v1 Announce Type: new
Abstract: How humans grow and maintain abstract knowledge from the sparse, streaming noisy data of experience is a longstanding challenge in cognitive science. Any computational account must satisfy at least three desiderata: It must be (1) data-efficient and c...
EvalDetectBench: A Benchmark for Measuring Evaluation Awareness in Frontier Language Models
arXiv:2609.01611v1 Announce Type: new
Abstract: Frontier large language models can often recognize when they are being evaluated, a capability known as evaluation awareness. If models behave differently in evaluations than in deployment, this undermines the validity of evaluation results, which are...
DISTAL: Distillation and Self-Supervised Pretraining for Structure-Agnostic Materials Property Prediction
arXiv:2609.00059v1 Announce Type: new
Abstract: Materials property prediction remains difficult in low-data settings, where many target properties are supported by only a limited number of labeled samples. Models with the strongest predictive accuracy often depend on crystal structures, which restr...
Task-Specific Prompt with Global Context for Multi-Task Graph Pre-Training
arXiv:2609.00047v1 Announce Type: new
Abstract: Graph prompt learning is an effective paradigm to adapt pre-trained graph models to downstream tasks in low-resource scenarios. However, existing multi-task graph pre-training frameworks generally use randomly initialized prompts, leading to poor alig...
Convergence issues in Relational Concept Analysis based on AOC-posets
arXiv:2609.00054v1 Announce Type: new
Abstract: Formal Concept Analysis (FCA) is an approach for conceptual classification building and rule discovery from a binary table describing a set of objects by a set of attributes. Extensions have been proposed to deal with non-binary and more complex data,...
ReNFT: Repairing Mode Collapse in Reward Post-Training via Internal Probability-Mass Recalibration
arXiv:2609.00061v1 Announce Type: new
Abstract: Reward post-training of diffusion generators inevitably concentrates probability mass on a few reward-favored modes, a mode collapse that erases within-prompt diversity. Existing methods for mitigating collapse rely on external signals or interfaces, ...
REAL-Q: E2E LLM Quantization via Dynamic Gradient Descent
arXiv:2609.00049v1 Announce Type: new
Abstract: Post-training quantization (PTQ) is essential for deploying large language models (LLMs) under strict resource constraints. State-of-the-art PTQ methods quantize each layer with a single closed-form second-order solver: to remain analytically tractabl...
I-CARE: Analysis of interference-related phenomena in a controllable, diverse and representative unlearning setting for text-to-image models
arXiv:2609.00003v1 Announce Type: new
Abstract: Machine unlearning studies the removal of knowledge from an AI model, making the system forget a concept it previously learned. Despite rapid progress in generative machine unlearning, the unintended degradation of semantically related concepts that s...
Incremental Risk Assessment of Progressive Elder Financial Scams via Instruction-Tuned Small Language Models
arXiv:2609.00005v1 Announce Type: new
Abstract: Financial scams targeting older adults increasingly occur through text and voice channels such as email, SMS, and phone calls, unfolding over multiple conversational turns that begin with impersonation or casual contact, escalate through trust buildin...
Discrete-Time MDP Modeling for Multi-Item Capacitated Lot Sizing with Stochastic Demand Timing
arXiv:2609.00004v1 Announce Type: new
Abstract: This paper studies a finite-horizon multi-item capacitated lot-sizing problem in which demand quantities are deterministic, while demand-arrival periods are stochastic. Each demand occurs once within a known time window and must be satisfied no later ...
Long-Horizon State Tracking in LLMs: Executing MD5 through a Deep Sequence of Dependent Tool Calls
arXiv:2609.00012v1 Announce Type: new
Abstract: Long-horizon tasks remain uncommon in large language model (LLM) evaluation, and for a reason: when each step depends on the last, per-step accuracy that looks excellent in isolation decays catastrophically, as errors cascade and the end-to-end failur...
HyperWorld: Hypergraph-Structured State Serialization Improves Learned Textual World Models
arXiv:2609.00002v1 Announce Type: new
Abstract: World models enable language-model agents to predict environment dynamics and plan before acting. In text environments, the model must learn symbolic action effects from serialized state descriptions, but the role of serialization structure remains un...
Equivariant Sheaf Neural Networks: Learning Geometric Transport on Graphs
arXiv:2608.28853v1 Announce Type: new
Abstract: Equivariant graph neural networks provide a principled way to model geometric systems, but efficient first-order architectures remain limited in how vector information can be transformed as it moves across a graph. We introduce \textsc{ESNN}, an Equiv...
Curvature Cryptanalysis of Smooth Transformer Feed-Forward Networks
arXiv:2608.28843v1 Announce Type: new
Abstract: We show that smooth two-layer feed-forward networks (FFNs) expose an additional structural model extraction channel under a chosen-input raw-output oracle at the FFN branch; consider transformer FFN branches with GELU or SiLU activations under chosen-...
The Halt Vector: Internalizing a Causal Steering Intervention for Efficient Reasoning
arXiv:2608.28859v1 Announce Type: new
Abstract: Reasoning models do not stop when they know the answer. On DeepSeek-R1-Distill-Qwen-7B the chain of thought runs about twice as long as the model's own answer probability takes to settle, and how much of that excess is removable varies from problem to...