ReCBM: Uncertainty-Gated Relational Reasoning for Concept Bottleneck Models
arXiv:2608.10004v1 Announce Type: new
Abstract: Concept Bottleneck Models (CBMs) provide an interpretable framework by grounding predictions in human-understandable concepts, enabling semantic inspection and test-time intervention. Recent variants have improved CBMs through richer concept represent...
arXiv:2608.09949v1 Announce Type: new
Abstract: This study evaluates the application of Large Language Models (LLMs) in complex biological systems, evolving from data analysis to autonomous, AI-guided experimentation. The framework is driven by data from a 49-channel phytosensor network, encompassi...
SkillConsist: Detecting Inconsistencies in Agent Skills via Bidirectional Graph Alignment
arXiv:2608.07639v1 Announce Type: new
Abstract: Agent Skills provide reusable capabilities to LLM agents. Agent Skill inconsistencies can expose undisclosed dangerous behavior or cause wrong Skill selection. Recent Agent Skill research has increasingly examined Agent Skill consistency detection. Ex...
Data-Driven Fire-Zone Segmentation for Improved Short-Term Wildfire Prediction
arXiv:2608.07472v1 Announce Type: new
Abstract: Wildfire prediction models typically discretize study areas into uniform grids, ignoring the heterogeneous spatial distribution of ignitions. We challenge this paradigm by showing that how data is discretized matters more than which model is used. We ...
Application of Artificial Intelligence for Fraudulent Banking Operations Recognition
arXiv:2608.07471v1 Announce Type: new
Abstract: This study considers the task of applying artificial intelligence to recognize bank fraud. In recent years, due to the COVID19 pandemic, bank fraud has become even more common due to the massive transition of many operations to online platforms and th...
Evolving Safety Landscape of Multi-modal Large Language Models: A Survey of Emerging Threats and Safeguards
arXiv:2608.07535v1 Announce Type: new
Abstract: Multi-modal large language models (MLLMs) integrate heterogeneous modalities through modality alignment and fusion, enabling stronger understanding and reasoning. However, this architectural shift reshapes the safety landscape of machine learning. Inc...
Emotion in an active inference model of human driving
arXiv:2608.07480v1 Announce Type: new
Abstract: Active inference has emerged as a principled framework for modeling adaptive behavior by balancing goal-directed action with uncertainty reduction. It has been successfully applied across biological and artificial systems, including recent work on hum...
Determinization in Structure Theories: A Unified Framework via Closure, Comparability, and Joint Admissibility
arXiv:2608.07476v1 Announce Type: new
Abstract: We develop a formal framework for constructing canonical interpretations from plural structure theories. A structure theory is a triple T = ({\Sigma}, A, I) consisting of a signature, axioms, and an inference policy, whose admissible interpretation fa...
Training Variable Long Sequences with Data-Centric Parallel
arXiv:2608.07524v1 Announce Type: new
Abstract: Training deep learning models on variable long sequences poses significant computational challenges. Existing methods force a difficult trade-off between efficiency and ease-of-use. Simple approaches use static configurations that cause workload imbal...
Fixed and Adaptive Topological DeepONets: Functional Measurements on Hausdorff Locally Convex Spaces
arXiv:2608.06428v1 Announce Type: new
Abstract: Deep Operator Networks (DeepONets; arXiv:1910.03193) typically encode an input function through point values on a fixed discretization. Building on the Topological DeepONet framework of Ismailov (arXiv:2603.11972), we replace point samples by continuo...
Sharding Prevents LLM Oversight Failures and Adversarial Exploitation
arXiv:2608.06422v1 Announce Type: new
Abstract: Giving an LLM judge more compute does not necessarily make it check more requirements. When one call must return many verdicts, some decisions become weakly grounded in the evidence, even when that call receives the same token or tool budget as a pane...
Latent Fact-Checking: Detecting Misinformation through Activation Engineering
arXiv:2608.06417v1 Announce Type: new
Abstract: The proliferation of misinformation online has driven demand for scalable detection systems. While most existing approaches rely on surface-level linguistic features or external knowledge retrieval, we examine truthfulness as a geometric property of a...
arXiv:2608.06427v1 Announce Type: new
Abstract: Generative models can reproduce an observational distribution while encoding an incorrect causal structure. We study a sequential game in which a structural causal generator proposes observational and interventional distributions, while an adversarial...
Risk-Aware Decision Policies for Agents Under Noisy Perception
arXiv:2608.06420v1 Announce Type: new
Abstract: Perception in biological systems is inherently noisy, requiring organisms to make decisions under uncertainty where misclassification can be costly or fatal. We present an Artificial Life predator-prey model of foraging under noisy perception, and com...
Beyond Routing Weights: Faithful Response-Level Interpretation of Mixture-of-Experts Reward Models via Contribution Contrast
arXiv:2608.06400v1 Announce Type: new
Abstract: Reward models are central to learning from human preferences, yet identifying what drives their predictions remains challenging. Recent sparse Mixture-of-Experts (MoE) reward models seek to improve interpretability by routing prompts to specialized ex...
ADIAS: Automated Design of Interactive Agentic Systems
arXiv:2608.06410v1 Announce Type: new
Abstract: Automated agent design improves agent harnesses through iterative revision, evaluation, and feedback summarization. Existing methods are largely candidate-centric: cross-round experience is organized around candidate agents, which leaves the repair pr...
Interpretable Unsupervised Community Detection with LLM-Symbolized Structured Processes
arXiv:2608.06402v1 Announce Type: new
Abstract: Community detection is a fundamental task in graph analytics that aims to identify cohesive groups of entities with similar behaviors or interests. Classic objective-driven methods struggle with complex graph structures, while deep-learning approaches...
EntropyMoE: Entropy-Aware Sparse Expert Routing for Tokenizer-Free LLMs
arXiv:2608.06398v1 Announce Type: new
Abstract: Recent byte-level large language models (LLMs) have made tokenizer-free modeling increasingly competitive by grouping bytes into dynamically sized patches. However, existing byte-patch architectures still apply the same dense feed-forward computation ...
Towards Multi-Label Graph Foundation Models: from Single-Vector Representation Learning to Multi-Semantic Basis Learning
arXiv:2608.06394v1 Announce Type: new
Abstract: Multi-label node classification is an important yet challenging task in graph learning, where nodes exhibit multiple semantics simultaneously. Existing methods for multi-label node classification can effectively model multiple labels, while only consi...
arXiv:2608.05234v1 Announce Type: new
Abstract: Building reliable applications that leverage large language models (LLMs) remains a significant challenge. While LLMs offer impressive capabilities across diverse tasks, their outputs often lack accuracy and provide no clear measure of confidence. Thi...
MS-MLB: An Open Machine Learning Benchmark for Blood-Based MS Classification
arXiv:2608.05196v1 Announce Type: new
Abstract: Multiple sclerosis (MS) is diagnosed through clinical assessment, magnetic resonance imaging, laboratory evidence when appropriate, and exclusion of better explanations. Blood RNA expression data may contain disease associated immune signal, but a blo...
arXiv:2608.05242v1 Announce Type: new
Abstract: In this work, we explore an alternative paradigm for spatial reasoning by explicitly disentangling 3D perception from reasoning, rather than jointly acquiring implicit 3D perception and reasoning through large-scale training. Our key observation is th...
Decoupling Perception from Description: Computation-Grounded Representation Alignment between Multivariate Time Series and Language
arXiv:2608.05238v1 Announce Type: new
Abstract: Training multimodal models to align time series with language runs into a self-supervision trap. The usual recipe asks an LLM to read a series and write a description, so label quality is capped by the perceptual skill the model is supposed to learn. ...
When Do Corrective Features Help? An Agent for Corrective Feature Discovery on Black-Box Forecasters
arXiv:2608.05207v1 Announce Type: new
Abstract: Frozen pretrained forecasters often fail in structured, recurring ways that are costly to repair through fine-tuning. We study corrective feature discovery: mining interpretable features of a frozen forecaster's residual to drive a lightweight post-ho...