AI papers — 2026-08-11

Gromov-Wasserstein Quantization and Clustering: Structure, Rates, and Algorithms. This paper introduces and analyzes the Gromov–Wasserstein (GW) quantization... A Recommendation System Approach for Interference-Robust Sensor Subset Selection. This paper develops a method for sensor-subset selection for tracking. A Single Atom in Front of a Mirror is a Universal Reservoir Computer. Core Claim The paper demonstrates that "the minimal architecture can carry the... DashArena: Benchmarking LLMs on Interactive Analytic Dashboard Generation. DashArena: Benchmarking LLMs on Interactive Analytic Dashboard Generation... Dual-Domain Cross-Modal Decoding for Clinical Text-Guided Medical Image Segmentation. Dual-Domain Cross-Modal Decoding (DD-CMD) is proposed for clinical text-guided... Putting Registers to Work: Task Registers for Token Pruning in Vision Transformers. This paper investigates whether token-pruning policies transfer across... Most biomedical publications show signs of LLM-assisted writing. Based on the paper "Most Biomedical Publications Show Signs of LLM-Assisted... Hypothesis Frontier: Verifier Guided LLM and Symbolic Search for First-Order Induction. Hypothesis Frontier: Verifier-Guided LLM–Symbolic Search for First-Order... Detecting an Effect Is Not Learning to Act on It: A Reward-SNR Floor for LLM Acquisition Agents. Core Thesis The paper makes a distinction that is easy to miss: detecting that... Ex-Omni-2D: Expressive Omni-Modal Dialogue Models with Native Visual Presence. Ex-Omni-2D is an omni-modal dialogue framework that generates a coordinated... Synthesizing Probabilistic Saturating Counters with Differentially Private Formal Guarantees. This paper presents a formal security analysis of Probabilistic Saturating... PAC-Bayes Beyond Parameter Space: Behavioral Equivalence, Z-Information, and Exact Complexity Decomposition. PAC-Bayes theory provides generalization guarantees by controlling the... Order Matters: LVLMs as Judges for Temporal Reasoning in Image Sequences. Order Matters: LVLMs as Judges for Temporal Reasoning in Image Sequences Core... A Joint-Distribution Route to Fair Representations with Continuous Sensitive Attributes. This paper addresses fair representation learning when the sensitive attribute... Telemetry and Concealment in Self-Adapting Generative AI: Logging Architecture, Adversarial Model Hiding, and the Limits of Detection. This paper addresses the governance problem posed by continually self-adapting... MARCO: Click-Intent Decomposition for Calibrated Ads Conversion Prediction. MARCO: Click-Intent Decomposition for Calibrated Ads Conversion Prediction... Temporally Grounded Compositional Camera Motion Understanding via Geometric Knowledge Distillation. This paper introduces a new formulation for camera-motion understanding called... ProTAGAD: A Foundation Model for TAG Anomaly Detection with Decoupled Topological and Textual Prototypes. ProTAGAD: A Foundation Model for TAG Anomaly Detection with Decoupled... Compute-Optimal Is Not Cluster-Optimal: Systems-Aware Scaling for Sparse Mixture-of-Experts. This paper introduces MOSAIC (Model Optimization via Systems-Aware TraIning... Probing and steering biology across Boltz-1s trunk-diffusion boundary. This paper investigates how biological information is represented and... ImpactHO: Importance-Aware KV Cache Transfer for Multi-User Edge LLM Handover. ImpactHO: Importance-Aware KV Cache Transfer for Multi-User Edge LLM Handover... Pair-Centric Graph Rewiring for Over-Squashing via Optimal Transport-Guided Communication Alignment. This paper proposes PairAlign (PAR), a pair-centric graph rewiring framework... Diffract: Spectral View of LLM Domain Adaptation. The paper "Diffract: Spectral View of LLM Domain Adaptation" studies continual... ChemWorld: Programmable Chemical Worlds for Controlled and Replayable Agent Experimentation. ChemWorld is a programmable chemical environment in which reusable process and... Agent Safety Should Be a Runtime Contract. Position: The paper argues that agent safety should not be treated as a... Defending against Model Extraction for GNNs with Model Reprogramming. Graph Neural Networks (GNNs) serve as the backbone for high-stakes applications... DEFT: Data-Efficient Frequency-domain Top-k Sampling via Inverse Discrete Fourier Transform for Spatiotemporal Dynamical Systems Modeling. DEFT: Data-Efficient Frequency-Domain Top-K Sampling via Inverse Discrete... Imaginative Generative AI: Crossing the Entropy Wall into Worlds Beyond Imitation. Generative AI models are primarily designed to imitate the data distribution,... Posterior contraction rates in Sobolev norms and Bayesian derivative estimation for infinite-dimensional exponential families. The paper studies posterior contraction in positive-order Sobolev norms and... Deployment Decision Reliability: A Generalizability-Theory Framework for Sizing Long-Horizon Agent Evaluations. Deployment Decision Reliability: A Generalizability-Theory Framework for Sizing... Towards Efficient Reasoning in LLM-Based Recommender Systems via Model Merging. This paper introduces REAM (Reasoning-HEad-Aware Merging), the first model... Three Tokens Force Exponential Feature Rank in Nonnegative Kernel Attention. This paper studies the expressive power of attention mechanisms by isolating... Causality Sum Rules in Conventional Scattering Matrices. This paper, "Causality Sum Rules in Conventional Scattering Matrices" by Ning... SkillLens: Visual Skill Cards for Retrieval-Augmented GUI Action Prediction and On-Policy Distillation. SkillLens converts heterogeneous GUI experience into Visual Skill Cards (VSCs)... Quantum Coordination Advantages in AI State-Tracking Tasks: Semantic Compilation and Latent Memory. This paper proves inference-time quantum coordination advantages for specified... ProbGuard: Calibrated Safety Risk Estimation from LLM Output Distributions. ProbGuard is the first completely probabilistic, architecture-agnostic... Recovering Wasted Compute in Autoresearch Agents. This paper studies the modeling pipeline at the core of autoresearch systems... TangPoetryBench: A Multi-Dimensional Benchmark and Rubric-Conditioned Evaluator for Poetry-to-Image Generation. TangPoetryBench: A Multi-Dimensional Benchmark and Rubric-Conditioned Evaluator... How to Verify Consistency of Probabilistic Claims. Problem Statement This paper addresses the question of whether a probabilistic... EVIL-Detect for NLPCC 2026 Shared Task 6: LLM-Generated Text Detection. This paper presents EVIL-Detect, a multi-signal ensemble framework with... Self-evolving network verifiers. Self-evolving network verifiers Authors: Ioannis Protogeros, Tibor Schneider,... GARLIC: Graph Attention-based Relational Learning of Multivariate Time Series in Intensive Care. GARLIC: Graph Attention-based Relational Learning of Multivariate Time Series... HyperFix: Combinatorial Nonlinear Correction for Task Vector Merging. HyperFix: Combinatorial Nonlinear Correction for Task Vector Merging introduces... Apodex Discovery: Reality Benchmarks and Environments for Evaluating and Building Discoverative Artificial Intelligence. Here is a detailed summary of the paper "Apodex Discovery: Reality Benchmarks... ConVAWG: A Retrieval-Grounded Framework for Controlled Synthetic Dialogue Generation in Violence Against Women and Girls. ConVAWG: A Retrieval-Grounded Framework for Controlled Synthetic Dialogue... Flow Straight to Reality: Perceptually Consistent Flow Matching for Efficient Image Restoration. PCFlow: Perceptually Consistent Flow Matching for Efficient Image Restoration... Every Token Counts: Exact Likert-Scale Distributions for Measuring LLM Attitudes and Biases. This paper introduces an analytically exact framework for the controlled... Optimistic Rates for Multiclass PAC Learning. This paper resolves the open problem of optimistic rates for multiclass PAC... SpeedTuning: Speeding Up Policy Execution with Lightweight Reinforcement Learning. SPEED TUNING: Speeding Up Policy Execution with Lightweight Reinforcement... Cross-View Feature Matching: Survey, Benchmarking, and Foundation-Model Perspectives. This survey presents a unified review of cross-view feature matching, a... Workflow Cards: Structured Summaries of Workflow Executions Using Provenance Data. Workflow Cards: Structured Summaries of Workflow Executions Using Provenance... MAP-Graph: Provenance-Aware Shared Memory for Multi-Agent Workflows. MAP-Graph is a provenance-aware memory layer for multi-agent workflows that... TimeRoute: Time-Aware Modality Routing and Diffusion for Multi-Modal Recommendation. TimeRoute: Time-Aware Modality Routing and Diffusion for Multi-Modal... The Illusion of Cross-Lingual Safety in Low-Resource Languages. This paper investigates whether safety alignment in large language models... From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop. The Workshop on Trustworthy Natural Language Processing (TrustNLP), co-located... CosMAP: Contrastive Manifold Approximation and Projection for Dimensionality Reduction of Omics and Genealogical Data. CosMAP: Contrastive Manifold Approximation and Projection for Dimensionality... Bayesian Symbolic Regression with Entropic Reinforcement Learning. Bayesian Symbolic Regression with Entropic Reinforcement Learning Oussama... Hidden in Plain Sight: Diffusion-Based Unrestricted Robotic Attacks on Vision-Language-Action Models. DURA is a diffusion-based unrestricted robotic attack that generates visually... Operationalising Relative Causal Knowledge: Backbone Identifiability from Private Reports on a Shared Outcome. Operationalising Relative Causal Knowledge: Backbone Identifiability from... DistilVDR: A Compact End-to-End Visual Document Retriever via Dual-Student Distillation. DistilVDR: A Compact End-to-End Visual Document Retriever via Dual-Student... From Faulty Memories to Corrected Actions: Dependency-Guided Rollback Repair for Memory-Augmented Agents. Problem and Motivation Persistent memory in language-model agents enables... Whisper-Aware LLM: Self-Supervised Uncertainty Learning for Robust Whispered Speech Recognition. The paper introduces the Whisper-Aware LLM, a framework designed to address the... MedUP: Awakening Unified Understanding and Perception in Medical Vision-Language Models. MedUP: Awakening Unified Understanding and Perception in Medical... Hierarchical Compositionality for An Assistive AI Agent. This paper presents an architecture for personalized command disambiguation in... RadFusion: Towards Threshold-Controllable Radiology Report Generation. RadFusion: Towards Threshold-Controllable Radiology Report Generation Summary... Mapping and Measuring the Behavioral Evolution of Large Language Models. The paper "Mapping and Measuring the Behavioral Evolution of Large Language... GitSkills: A Dataset of Agent Skills on GitHub. GitSkills is a dataset of 3,797,117 SKILL.md files collected from 282,200... Agentic Instruction Data Selection: Let DataMaster Interpret Your Intent. The paper introduces DataMaster, an agentic instruction data selection system... A lower bound for stepsize-based acceleration of gradient descent. This paper establishes a new lower bound on the convergence rate of plain... Improving TensorSketch Using Complex Random Variables. Improving TensorSketch Using Complex Random Variables Summary This paper... Weightless Fine-Tuning: Personalizing LLMs via Logit-Space Transport. Weightless Fine-Tuning (WFT) is a training-free, decoding-time method that... Chemically Meaningful Textualization Enables Explainable Validation of Metal-Organic Frameworks by Large Language Models. This paper demonstrates that large language models (LLMs) can serve as... SQuaT: Self-Supervised Knowledge Distillation via Student-Aware Quantized Teacher Features. SQuaT (Student-Aware Quantized Teacher Features) is a label-free... Towards Unified Dynamic Face Landmark Detection. This paper introduces Unified Dynamic Face Landmark Detection, a novel... Auditing Chinese Web-scale Corpora via Sampled BPE Token Statistics. This paper presents SAMPLED-BPE, a lightweight token-level auditing pipeline... Uncertainty-Aware Deep Learning for Genomics Applications: Insights from an Empirical Study. This paper presents an empirical analysis of uncertainty quantification (UQ) in... Uncertainty-Aware Compositional Localization and Placement Assessment of Catheters and Tubes in Chest X-Rays. Uncertainty-Aware Compositional Localization and Placement Assessment of... Multi-Granular Rationale-Guided Molecular LLM for Property Prediction. MR-MoL is a multi-granular rationale-guided molecular LLM for property... Tensor-normal maximum likelihood estimation at the operator-norm sample threshold. Let be independent Gaussian tensors in whose covariance is a Kronecker product... Better, Faster, Stronger: Programmatic Skill Learning Best Reduces Agent Cost. SpeedRunner is a method for programmatic skill learning that reduces agent cost. Generative Learning for Quantum Measurement Design. Generative Learning for Quantum Measurement Design Authors: Jun Dai, Olivier... Dual-Loop Self-Evolution via Verifiable Emotion Feedback for Multi-Turn Empathetic Dialogue. The paper introduces a dual-loop self-evolution framework for multi-turn... Seeds Before Objectives: Rethinking Evaluation for Low-Resource Garhwali ASR. This paper presents a rigorous, multi-seed evaluation of Automatic Speech... Calibrating Post-Training Feature Shifts for LLM Data Contamination Detection. The paper proposes CALIBDCD, a calibration framework for feature-based data... Physics-informed Diffusion Generative Model for Time-Series Data Synthesis in Dynamic Systems. PhysDGM is a stepwise physics-embedded diffusion generative model for... Measure, Don't Optimize: Forecasting Recovery in LLM Unlearning. Core Contribution The paper introduces J-Access, an inference-time auditing... Mitigating Context Interference for Reliable and Efficient Search Agents. This paper investigates the issue of context interference in multi-turn search... MoE Proxy Models for Low-Cost Failure Reproduction and Diagnosis in LLM RL Post-Training. This paper proposes a multi-view, frequency-aware expert pruning method for... When Is a General Factor Distinguishable? Non-Proportionality, Stable Structure, and the Bifactor Decision. When Is a General Factor Distinguishable? Non-Proportionality, Stable... Simplex Relaxation for Discrete Diffusion. Core Contribution The paper introduces Simplax, "an exact... MD-ProTector: Positioning Multiple Data-Driven Prototypes for LLM-Generated Text Detection. MD-ProTector is an input-only encoder detector for LLM-generated text detection... What We Know about Responsible AI Practices in Industry: A Half Decade of Empirical Research. This paper synthesizes current knowledge about Responsible AI (RAI) practices... The Next Challenge for Agentic Cybersecurity: A Realistic, Contamination-Free Reverse Engineering Benchmark. SRE-Bench is the first realistic, contamination-free reverse engineering (RE)... Partially Observable Learning for Multi-Platform Dispatch Optimization. This paper proposes POLO, a Partially Observable multi-agent reinforcement... Terminal Symmetry as a Decision Resource: Statewise Refinement for Anytime Verified Construction. Core Contribution This paper develops a decision-resource view of terminal... Backdoor Decontamination Dynamics in LLM Agents. Backdoor Decontamination Dynamics in LLM Agents Authors: Gabriel Huang, Abhay... Why Does CLAUDE.md Keep Growing? Catastrophic Remembering in Agentic Coding. Agentic coding READMEs like CLAUDE.md grow without bound in real repositories,... Evidence-Grounded Trustworthy Multimodal Reasoning and Evaluation Benchmark in Complex Urban Scenes. The paper introduces AD2-Bench, a large-scale benchmark for evaluating... Once Poisoned, Arbitrarily Controlled: A Programmable Backdoor in VLMs. This paper introduces the first any-to-any backdoor attack on Vision-Language... BREAD: Baseline-Referenced Explanations for Anomaly Diagnosis. BREAD: Baseline-Referenced Explanations for Anomaly Diagnosis proposes a... Adaptation of Generalist Robot Policies with Minimal Data. Core Problem and Setting The paper introduces minimal-data adaptation (MDA), a... Lifecycle-Optimal Tokenization: Vocabulary Size as a Deployment-Regime-Dependent Infrastructure Parameter. This paper argues that tokenizer vocabulary size in large language models... Market-Information-Aware Gated-LoRA of Foundation Models for Transferable Day-Ahead Electricity Price Forecasting. This paper proposes a market-information-aware adaptation framework that... Cross-Corpus Evaluation of Generalizable Vulnerability Detection in IoT Firmware. This paper introduces IoTVulBench, a human-verified benchmark for cross-corpus... Entropy-based Code Adversarial Translation for Real-world Repository Migration. This paper introduces Entropy-based Code Adversarial Translation (ECAT), a... Share First, Route What Remains: A Unified Framework for Token-Adaptive MoE Computation. Mixture-of-experts (MoE) models have recently moved beyond routing a fixed... TRACES: A Benchmark for Epistemic Reliability in Scientific Reasoning by LLMs. TRACES: A Benchmark for Epistemic Reliability in Scientific Reasoning by LLMs... Unlocking the Power of Medical Tabular Data via Semantic-Aware Multimodal Pre-training. The paper "Unlocking the Power of Medical Tabular Data via Semantic-Aware... Rethinking Text-Based Image Retrieval in Specific Domain. The paper introduces SecMM-TBIR, a multi-match benchmark for surveillance... Conversational Orchestration for Organic 6G. The paper proposes a lightweight, decentralized conversational orchestration... Association-based Privacy Attacks in Wireless Protocols: Formal Modeling and Mitigation. This paper formally investigates the root causes of pairing-based privacy... Principal Trait Analysis: Towards Deriving "Skills" in Human-AI Collaboration. Principal Trait Analysis (PTA) is a novel, data-driven algorithm inspired by... The GenAI Catch-22: Use of Generative Artificial Intelligence in Norwegian Newsrooms During the 2025 Parliamentary Election. The GenAI Catch-22: Use of Generative Artificial Intelligence in Norwegian... Conditional Independence Tests for Constraint-Based Causal Discovery: A Survey. This survey reviews conditional independence (CI) testing with emphasis on... Conflict and Congruency Effects in Large Language Models: In-Weight and In-Context Competition in a Verbal Conflict Task. This paper introduces a novel verbal-only conflict task, the "crayon task," to... Rationale-Guided Learning for Multimodal Emotion Recognition. RATIONALE-GUIDED LEARNING FOR MULTIMODAL EMOTION RECOGNITION The paper proposes... Dueling Deep Q-Learning for Intrusion Detection. This study proposes a novel approach to intrusion detection systems (IDS) by... Scheduling Mixed RL Rollouts Beyond Prefix Locality. MISA-T is a routing-layer admission policy for mixed rollout serving in... Leveraging Human Reading Behavior for Keyphrase Extraction: A Webcam-based Eye-tracking Corpus. This paper investigates whether lightweight webcam-based eye-tracking features... On Understanding, Identifying, and Mitigating Vulnerabilities in Agentic Large Language Models. This systematic literature review, conducted under PRISMA 2020 guidelines,... myMediWhisper: Construction of Burmese Medical Speech Corpus and Whisper Fine-Tuning for Clinical Dialogue ASR. This paper presents myMediWhisper, a Burmese medical speech recognition... Entropy-Centric Explainable AI for Remote Sensing Image Segmentation. This paper proposes an entropy-centric explainable AI (XAI) method for semantic... Contextual Information Policy Optimization for Search Agents. The paper, authored by Xingyu Guo, Wei Chen, Linlin Yang, and Baochang Zhang... Dual-Primal Graph VAEs for Noisy Label Aggregation. Dual-Primal Graph VAEs for Noisy Label Aggregation proposes a graph VAE... RLMOpt: Adaptive Prompt Optimization via Recursive Language Models. RLMOpt is a prompt optimizer that makes the search policy itself... Actionable Hallucination Detection: Translating Latent Uncertainty into Agentic Critique. Actionable Hallucination Detection: Translating Latent Uncertainty into Agentic... ELVAE: Evidential Learning-Based Variational Autoencoder for Uncertainty-Aware Generation. ELVAE: Evidential Learning-Based Variational Autoencoder for Uncertainty-Aware... FUSE: Frame-Unified Stress Estimation from Facial Video. FUSE (Frame-Unified Stress Estimation) is a facial-video stress detection... Reasoning Shortcuts and Value Symmetries: What Symmetry Permits, Architecture Realizes, and Optimization Selects. This paper addresses the problem of reasoning shortcuts in neurosymbolic... Beyond Forecasting: Recasting Volatility Control as a Routing Problem. Beyond Forecasting: Recasting Volatility Control as a Routing Problem Core... Continuous Interaction Diffusion: A Diffusion-Native Runtime for Asynchronous Tool-Augmented Reasoning. Continuous Interaction Diffusion (CID) is a diffusion-native model–runtime... Riemann GeoResolver: A Non-Euclidean Attention Framework from Euclidean Resolver to Hyperbolic-Spherical Geometry. Riemann GeoResolver: A Non-Euclidean Attention Framework from Euclidean... Threat-guided Policy-aware Scene Perturbation for Safe Autonomous Driving with Online Reinforcement Learning. TPSP introduces a policy-aware scene encoder to capture the interaction between... Basin: Efficient and Extensible Numerical Optimization in Rust. Basin is a numerical optimization library for the Rust programming language... InSight-doc: Agentic Visual Perception for Long-Document Understanding. InSight-doc: Agentic Visual Perception for Long-Document Understanding proposes... Hardware-Aware Deployment of Joint SAR Compression and Despeckling on FPGA. This paper presents the first deployment of a joint SAR Despeckling and Data... Click2Poly: A VLM for vector mapping buildings and walls. Click2Poly is a human-in-the-loop AI assistant designed to speed up the manual... BooST: Bridging Semantics and Motions for Efficient Skill Transfer. BooST: Bridging Semantics and Motions for Efficient Skill Transfer introduces a... Language-Structured Relational Q-Learning for Threat-Aware Control in Safety-Critical Driving. The paper introduces Language-Structured Relational Q-Learning, instantiated... Retrieval-Augmented Vision Foundation Models for Robust Leukemia Cell Classification across Multiple Microscopy Datasets. This work presents a robust framework for leukemia classification across... TideRL: Boosting Agentic RL Goodput with Readiness-Aware Scheduling. TIDE RL: Boosting Agentic RL Goodput with Readiness-Aware Scheduling Abstract... UniProbe: A Learnable Token-Level Hallucination Detector for Large VLMs using Multi-Structural Internal Representations. UniProbe is a lightweight, unified, learnable detector for token-level... IADD-TR: Intervention-Aware Dynamics Decoupling with Targeted Regularization for Model-Based Reinforcement Learning. IADD-TR: Intervention-Aware Dynamics Decoupling with Targeted Regularization... Forward Trajectory Steering for Hamilton-Jacobi Reachability Analysis. This paper introduces STEER2REACH (S2R), a physics-informed neural network... Diffusion-Based Data-Driven Assortment Optimization. Diffusion-Based Data-Driven Assortment Optimization proposes D3AO, a... RelShap: Relationally Consistent Shapley Explanations. RelShap is a framework that incorporates relational constraints and data... Convergence Guarantees of Gradient Descent for Neural Networks via Generalized Lipschitz Smoothness. The paper establishes convergence guarantees for gradient descent applied to... Large-scale AI-Ready Data for Anti-Cancer Drug Response Modeling. The paper presents a substantial expansion of the IMPROVE benchmark dataset for... -VAEs as Effective Theories: Tolerance-Dependent Dimension. In a beta-VAE, increasing the regularization strength acts as a spectral cutoff... Generator-Guided Inverse Sampling for Lévy-Driven Generative Models. This paper studies inverse sampling for Lévy-driven generative models from the... Invertible Logits Transformation for Accuracy-Preserving Post-Hoc Uncertainty Calibration. The paper introduces Invertible Logits Transformation (InvLT), a post-hoc... Accelerated Learning of High Dimensional Functions with a Tensor-Featured Training Network. This work presents a method to accelerate the optimization of learning high... Efficient Weak-Entropy PINN for Solving Hyperbolic Conservation Laws. The paper introduces a novel physics-informed neural network framework called... MUSE: A Full-Text Cross-Domain Knowledge Base of Scientific Problems, Solutions, and Rationales. MUSE: A Full-Text Cross-Domain Knowledge Base of Scientific Problems,... VoxSumm: A Multilingual Corpus of Long-Form Spoken News for Joint Summarization and Translation. The paper introduces VoxSumm, a multilingual corpus and benchmark for joint... Stigma and Support in Online Sexual Violence Narratives on Reddit. This paper introduces the SCOPE (Stigma and COmmunity Peer Expressions)... Decomposition-Induced Context-Memory Conflict: When Fact-Checking Pipelines Contradict Their Own Source Text. Decomposition-Induced Context-Memory Conflict: When Fact-Checking Pipelines... REAP: Relation-Aware Elicitation and Parsing for Closed-Book Knowledge Base Construction from LLMs. The paper presents R EAP (Relation-aware Elicitation And Parsing), a system for... Assessing Reliability of BERT-Based Models on Question Answering Tasks. This study evaluates the reliability of four BERT-based models—RoBERTa,... Evo-Bench: Can Language Models Improve Agent Harness?. Evo-Bench is the first benchmark designed to evaluate large language models'... ConRub-Med: Reinforcement Learning with Consensus Rubrics for Open-Ended Medical Question Answering. ConRub-Med: Reinforcement Learning with Consensus Rubrics for Open-Ended... ASR-Roundtrip Evaluation Can Mask Context- and Convention-Dependent Reading Errors in Chinese News TTS. ASR-roundtrip evaluation is widely used as a scalable proxy for text-to-speech... Who Gets Heeded? An Obligation-Level Audit of Responsiveness in EPA Rulemaking. This paper introduces obligation-level responsiveness auditing, an auditable,... Stability of Finite-Batch Particle Mean-Field Variational Inference Beyond Strong Convexity. This paper studies the stability of finite-batch particle mean-field... Emotion2Skill: Model-Internal Emotion Signals for Adaptive Skill Selection and Evolution. Emotion2Skill is a framework that extracts LLM-internal emotion vectors and... sLTN: Structural Logic Tensor Networks. sLTN: Structural Logic Tensor Networks introduces an extension of Logic Tensor... A Comparative Evaluation of Deep Learning Object Detection Models on a Real-World Multi-Plant Dataset from Africa. This study presents a comparative evaluation of six object detection... DACRI: Decision-Aware Causal Intervention Ranking for Critical Supply Chains. DACRI: Decision-Aware Causal Intervention Ranking for Critical Supply Chains... Efficient Hypergradient Descent for Inverse Reinforcement Learning. This paper addresses the computational challenges of inverse reinforcement... Test-Time Self-Evolving GUI Visual Grounding via Reflection-Guided On-Policy Self-Distillation. This paper introduces a Test-Time Self-Evolving framework for GUI visual... Beyond a Bag of Features: Set-Level Instability in Sparse Autoencoders. This paper investigates whether the interpretability of individual sparse... ReRound: Reconstructive Rounding to Resolve Midpoint Ambiguity in Calibration-Free LLM Quantization. ReRound (Reconstructive Rounding) is a post-training quantization method that... Spectral Embeddings of Degree- Laplacians in Random Dot Product Graphs. This paper studies a continuum of degree-normalized spectral embeddings for... Path Integral Value Matching for Linear Quadratic Stochastic Optimal Control. The paper addresses the computational challenges in solving Linear Quadratic... Decoding-Level Taboo: A Diagnostic Stress Test for LLM Robustness. Decoding-Level Taboo is a zero-prompt diagnostic stress test that intervenes... Long-Horizon Forecasting of Complete Financial Statements with Forma. Core Contribution and Problem Statement The paper introduces ProForma-20Q, a... Statistically-Secure Bit Commitment and Coin Flipping Protocols Based on Quantum Hardware Assumptions. This paper introduces the first statistically secure bit commitment and coin... Enhanced Filtering Algorithms for the Euclidean Traveling Salesperson Problem and its variants in Constraint Logic Programming. The paper "Enhanced Filtering Algorithms for the Euclidean Traveling... Efficient Reinforcement Learning for Long-Horizon Tool-Use Agentic Tasks. S INK F LEX-RL is a modular training system for reinforcement learning (RL) in... Spectral graph clustering with inhomogeneous latent geometry. Spectral graph clustering with inhomogeneous latent geometry Authors:... BPG: Balancing Plasticity and Generalization for Domain Incremental Learning. BPG: Balancing Plasticity and Generalization for Domain Incremental Learning... EvoMem: Memory-Augmented Evolution for Code Optimization. EvoMem is a persistent memory architecture for LLM-based evolutionary program... Iterative Erasure Count Is Not an Affine-Invariant Concept Dimension. Core Claim This paper argues that iterative erasure count is not an... Benchmarking LLM-Guided Control-Plane Policies for Backend Fault Isolation in HAProxy. This paper investigates whether a Large Language Model (LLM) can replace the... When and Where Faults Matter: A Study of Transient Errors in CKKS Multiplication. This paper presents an in-depth analysis of the resilience of homomorphic... Battlefield 5G: Dual-PKI and TPM-Based UE Attestation for Tactical 5G Standalone Networks. The paper presents Battlefield 5G, a pre-authentication framework for tactical... Goodness-of-Fit Tests and Calibration Machine-Learning Algorithms for Logistic Regression with Sparse Data. Based on the paper "Goodness-of-Fit Tests and Calibration Machine-Learning... Improved cross-validated distances for multivariate pattern analysis. The paper "Improved cross-validated distances for multivariate pattern... MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment. MultiModal Code-Switching: Interleaving Visual Objects into Language for... Beyond Detection Accuracy: Measuring Explanation Cost, Stability, and Utility for Resource-Aware IoT Intrusion Detection. This study jointly evaluates predictive effectiveness, explanation cost, local... VisEditBench: Can Vision-Language Models Edit Visualization Code from Multimodal Feedback?. VisEditBench is a benchmark for evaluating vision-language models (VLMs) on the... Persona Conditioning as an Assessor-Sensitivity Probe for LLM-Based IR Evaluation. The paper studies persona conditioning as a diagnostic mechanism for exposing... Quantum Incremental Learning with Mixed State Prototypes. Quantum Incremental Learning with Mixed State Prototypes Abstract Incremental... RevCRN: Reversible Analog Computation using Chemical Reaction Networks. This paper introduces the Reversible Chemical Reaction Network (RevCRN) model... Predicting Space Groups of Double Perovskites by LLM with Dynamic Few-Shot Learning. DyRIS, an LLM-agent-based framework, predicts ranked space-group (SG)... Automatic Field-of-View Adjustment for a View-Expansive Microscope via LSTM-Based Gaze and Pipette Motion Interpretation. Intracytoplasmic sperm injection (ICSI) operators frequently adjust the... Benchmarking Time Series Generation Methods for Privacy-Preserving Forecasting. the paper "Benchmarking Synthetic Time Series Generation Methods for... Self-Knowledge Retrieval Augmented Generation Framework for Patent Matching. This paper proposes a self-knowledge retrieval-augmented generation (RAG)... Two-stage Odd Residual Flows for Mean-Preserving Probabilistic Time Series Forecasting. Two-stage Odd Residual Flows for Mean-Preserving Probabilistic Time Series... ReOrder-OPD:Reliability-Aware Prompt Ordering for On-Policy Distillation. On-policy distillation (OPD) applies token-level teacher supervision to... Hierarchical Empirical-Bayes Naive Bayes: Minimax Smoothing and Calibration with AODE Extension. The paper proposes hierarchical empirical-Bayes Naive Bayes (HEB-NB), a... Certify or Refuse: A Cross-Model Map for Selective Risk Control with Coverage Floors under Covariate Shift. The paper introduces the Floor Certification Map, a theoretical framework for... 3D Weighted Geometric Graph Neural Networks for Sheep Facial Pain Assessment. This paper presents the 3D Sheep Pain Facial Expression System (3D-SPFES), a... GeoForge: Non-Parametric Self-Evolving Agents for Earth-Observation Reasoning. GeoForge is a training-free, self-evolving framework for Earth observation (EO)... CLEAR: Class-wise Expert Aggregation with Structured Sampling for Long-Tailed Classification. CLEAR (Class-wise reLiability-aware Expert Aggregation for long-tailed... A Lightweight Fault-Detection Scheme for Barrett Modular Multiplication Using Multiple Conditional Reduction Paths. This paper proposes a lightweight fault-detection scheme for Barrett Modular... Narrative Keyframing for Generative Creative Writing. The paper introduces narrative keyframing, a new interaction technique for... On the Sensitivity to Errors in Homomorphic Computing: Single Transient Bit-flip Client-side Error Characterization. This paper analyzes the sensitivity of Homomorphic Encryption (HE) to bit-level... A Gateway Architecture for Enterprise MCP Authentication: Unifying Heterogeneous Auth, Identity Delegation, and the User / Non-User Persona Problem. The paper reports on a production deployment of a centralized MCP (Model... Trigger the Straggler: Load Hijack on Mixture-of-Experts LLMs. This paper introduces Load Hijack, a supply-chain attack against... A Quantum Roadmap for Softmax Attention: Exact Born-Rule Analogs for Softmax Attention on the Probability Simplex. The paper presents a theoretical construction establishing an exact equivalence... Physics-Informed Implicit Neural Representations for Improved Myocardial Perfusion MRI Quantification. This paper presents a physics-informed neural network (PINN) framework with... Dynamics Models for Offline Hyperparameter Selection in Real-World RL. A key obstacle to deploying reinforcement learning in real-world systems is... Unmasking Toxic Mimicry in Medical Offline Reinforcement Learning for ICU Sepsis Management via Counterfactual Clinical Audits. Offline reinforcement learning (RL) offers considerable promise for optimizing... AutoGrable: What Is a Good Graph for a Table?. The paper addresses the fundamental question of graph construction for tabular... TEAMMix: Taxonomy Enrichment Augmentation and Minority-augmented Mixing Strategy for LLM-enhanced Weak-Supervised Hierarchical Text Classification. TEAMMix: Taxonomy Enrichment Augmentation and Minority-augmented Mixing... Gaussian Meta-Space Augmentation for Stacking Ensembles in Multimodal IPMN Risk Stratification. The paper introduces cUPMI (calibrated Upstream Probabilistic Meta-Imputation),... Self-Normalized Inference for Constant-Stepsize Temporal-Difference Learning under Markovian Sampling. This paper develops inferential methods for constant-stepsize... Attention-Path Fragility as an Uncertainty Signal in Large Language Models. The paper proposes that a model's uncertainty about a token is reflected not... Mechanism Design for Generative Engines: From Exploitation toward Win-Win Outcomes. Mechanism Design for Generative Engine: From Exploitation to Win-Win... ODE-Based Transformer Decoders for Iterative Sign Language Translation. Sign language translation has achieved strong results with Transformer... Self-Evolving Embodied Agents via Skill-Harness Evolution. Core Contribution The paper introduces SHAPER, a self-evolving framework for... ELBench: A Multi-Dimensional Benchmark for Education-Facing Large Language Models. ELBench: A Multi-Dimensional Benchmark for Education-Facing Large Language... An Empirical Study of Output-to-Input Loops for Black-Box Backdoor Detection in Fine-Tuned Open-Weight LLMs. An Empirical Study of Output-to-Input Loops for Black-Box Backdoor Detection in... Self-Correcting Long-Horizon Search Agents via Tree-Structured Memory. ReTree is a self-correcting tree-structured memory mechanism for LLM-based... MIRA: Medical Image Reflection for Agentic Diagnosis. MIRA (Medical Image Reflection for Agentic Diagnosis) is a medical visual... ThinkRetrieve: Retrieval-Augmented Reasoning Traces for Test-Time Scaling. ThinkRetrieve is a test-time scaling framework that augments the reasoning... RTSKG: Building a Rail Transit Station Knowledge Graph Dataset. RTSKG is a new rail transit station knowledge graph dataset that explicitly... Towards an approach to multivariate outlier detection for District Heating System data. This paper tests different methods for multivariate detection of outliers in... Surgical WAM: A World-Action Model for Data-Efficient Surgical Robot Learning. Surgical WAM: A World-Action Model for Data-Efficient Surgical Robot Learning... FedCGR: Federated Cross-Domain Generative Recommendation. FedCGR: Federated Cross-Domain Generative Recommendation proposes a federated... Socioduality: A Relational Process Framework for Human-AI Interaction. Socioduality is defined as "a sequential, reciprocal, and history-carrying... IO Factory: Simulating AI-Enabled Influence Campaigns at Scale. IO Factory is an AI-driven framework for simulating information and influence... FaithformBench: Benchmarking Faithfulness of Mathematical Chain-of-Thought Autoformalisation. The paper introduces FaithformBench, a benchmark for evaluating the... -SUB: A Physics-Informed Synthetic Underwater Benchmark Dataset for Underwater Image Enhancement. This paper presents π-SUB, a physics-informed framework for generating... Longitudinal Evidence That General-Purpose Chatbots Actively Foster Relational Engagement. This paper presents a pre-registered four-week longitudinal study (N = 72,... Cross-View Sequential Visual Localization with Spatio-Temporal Context Modeling for Autonomous Driving. This paper proposes a temporal-context-enhanced framework for cross-view... Benchmarking Cyberattack Detection in Electric Vehicle Charging Infrastructure with Benign User Updates. This paper develops a leakage-controlled session-level benchmark for... Variational Parameter Calibration with Physics-Aware Latent-Space Surrogates. This paper introduces a physics-aware neural-network-based latent-space... Information Bottleneck under Perfect Privacy. The paper studies the information bottleneck problem under a perfect privacy... Long-Time Trajectory Approximation via SA-NODEs: Model Predictive and Floquet Strategies. This paper studies the approximation of dynamical systems by semi-autonomous... SCOUT: Symmetric Consensus Outlier Detection for Failure Localization in LLM Pre-Training. SCOUT is a unified runtime failure-localization framework for LLM pre-training,... What Iterated Self-Feeding Probes of Language Models Measure, and a test that separates the construction from the model. This paper asks what iterated self-feeding probes of language models measure,... Link-adaptive digital twin for robust physical-layer modeling in hybrid-amplified ultra-wideband optical networks. Link-Adaptive Digital Twin for Robust Physical-Layer Modeling in... CARB: A Characterization-Guided Framework for CNN Inference Cost Prediction and Deployment Screening. CARB: A Characterization-Guided Framework for CNN Inference Cost Prediction and... Dreamer-SAC: Off-Policy Learning in Latent World Models for Sample-Efficient Autonomous Driving. Dreamer-SAC: Off-Policy Learning in Latent World Models for Sample-Efficient... Can Bayesian Optimization Efficiently Find a Strong Single Expert in Neural Thickets?. The paper investigates whether Bayesian optimization (BO) can efficiently find... MemSpec: Memory-Aware Runtime for Adaptive Draft Scheduling in Speculative Decoding on Edge Devices. MemSpec: Memory-Aware Runtime for Adaptive Draft Scheduling in Speculative... Expert-Guided g-computation with Large Language Models for Estimating Causal Effects on Timings: Applications to Hospital Quality Improvement. This paper introduces expert-guided g-computation (egg-computation), a novel... MazzikaAI: A knowledge-based performance-to-prompt compiler for real-time Arabic maqam accompaniment with a streaming text-to-music model. MazzikaAI is a knowledge-based system for real-time Arabic maqam accompaniment... Strengthening Full Justified Representation: Efficient Verification and Computation. This paper introduces FJR+, a strict strengthening of the full justified... Let it Cook: Learning to Wait in Sequential Decision Making. The paper addresses the question of whether agents in sequential decision... From Numbers to Judgment: Specialist LLM Agents and Reinforcement Learning for European Listed Real Estate. The paper "From Numbers to Judgment: Specialist LLM Agents and Reinforcement... Contextual Quality-Diversity Evolutionary Reinforcement Learning for HVAC Control in Tropical Commercial Buildings. This paper proposes CQD-ERL, a contextual quality-diversity evolutionary... V-FiLLM: Verified Financial LLM Reasoning Benchmark. V-FiLLM is a framework that generates financial reasoning benchmarks from... Do Influence Tactics Matter? Investigating Prompt Framing Effects in LLM Code Generation. This paper presents the first large-scale empirical study investigating whether... From Prompting to Behavioral Alignment: Personalized LLM Judges for Recommendation Evaluation. Traditional offline recommendation evaluation relies heavily on complex,... Social Chain of Thought: A Multi-Agent Architecture Grounded in Medical Differential Diagnosis Methodology. Social Chain of Thought: A Multi-Agent Architecture Grounded in Medical... Gaze Target Estimation Anywhere with Concepts. This paper introduces the Promptable Gaze Target Estimation (PGE) task and the... AlbumentationsX: One Augmentation Pipeline for Images and Related Annotations. AlbumentationsX is a data augmentation library that stores the transform list,... Herding End-to-End Autonomous Driving via Neuro-Symbolic Safety Guards. The paper introduces a neuro-symbolic safety guard for end-to-end autonomous... Can Frontier LLMs Match Natively Multimodal Embeddings? A Comparison on Hard-Negative Text-to-Image Retrieval. This paper presents the first direct comparison of natively multimodal... Coordinating the Unknown Lipschitz Constant in Multiplayer Bandits. The paper studies cooperative multi-agent bandits in continuous (Lipschitz)... Uncertainty-Aware and Explainable Ensemble Deep Learning Framework for Multi-Class Skin Lesion Classification. This paper proposes an uncertainty-aware and explainable deep learning... VERDICT: Training-Free Step-Wise Verification of Multimodal Reasoning via Disagreement-Aware Consensus. This paper introduces VERDICT (VERification via Disagreement-Informed Coupled... Dynamic Context Adapters: Efficiently Infusing History into Vision-and-Language Models. Dynamic Context Adapters: Efficiently Infusing History into Vision-and-Language... Evaluating Rational Contracting in Natural Language. and Motivation The paper addresses the challenge of evaluating how well... A Cost-Efficient Routing Pipeline for Multilingual Short-Text Classification Using Small Language Models. This paper presents a cost-efficient routing pipeline for multilingual... TACTICL: Task-Aware Compression of Tabular ICL Models. TACTICL: Task-Aware Compression of Tabular ICL Models Abstract Summary: The... Tree-of-Ideas: Automated Research Ideation via Cross-Trajectory Reasoning over Scholarly Evolution. The paper addresses a fundamental limitation in automated research ideation... SPIEval: Evaluating Large Language Models as Mobile Assistants over Scattered Personal Information. SPIEVAL is a human-curated benchmark introduced to evaluate large language... Retrieval-Corrected Conformal Prediction for Time Series. Retrieval–Corrected Conformal Prediction (RCCP) is a retrieval-augmented... Measuring Semantic Abstractness of SAE Features via Nonlocality. and Motivation The paper addresses a central challenge in Mechanistic... XCoT-VLA: Executable Chain-of-Thought for Vision-Language-Action Driving. XCoT-VLA: Executable Chain-of-Thought for Vision-Language-Action Driving... Compositional Benchmark Synthesis for Hierarchical Human Action Recognition. Compositional Benchmark Synthesis for Hierarchical Human Action Recognition... Smart Enough to Go Extinct? An Evolutionary Challenge to the Value of General Intelligence and Its Ethical Implications for AGI. The paper "Smart Enough to Go Extinct? An Evolutionary Challenge to the Value... Robust Multi-Agent Bandits with Heavy-Tailed Rewards and Information Asymmetry. The paper studies multi-agent multi-armed bandits (MAB) with heavy-tailed... ReLTEx: Reliable LLM-based Taxonomy Expansion. ReLTEx: Reliable LLM-based Taxonomy Expansion Zeinab Ghamlouch, Mehwish Alam... VibeLifeBench: Can Your Life Agent Be Proactive and Persistent in a Living World?. VibeLifeBench: Can Your Life Agent Be Proactive and Persistent in a Living... Optimal Stopping of Self-Refining Foundation Models. The paper "Optimal Stopping of Self-Refining Foundation Models" by Kim Hammar,... Your LLM, Your Style: Behavioral Mode Axes for LLM Behavioral Control. Your LLM, Your Style: Behavioral Mode Axes for LLM Behavioral Control Core... DegradeQuery: Counterfactual Tuple Pretraining for Context-Aware PROTAC Degradation Prediction. Problem and Motivation Proteolysis-targeting chimeras (PROTACs) induce protein... Never Stop Speaking: a Denial-of-Service Attack on End-to-End Speech Language Models. This paper introduces a novel denial-of-service (DoS) attack targeting... SegPAR: Class-Centric Decision-Based Sparse Attack for Semantic Segmentation. SegPAR: Class-Centric Decision-Based Sparse Attack for Semantic Segmentation... MEGA: Self-Evolving Agent Optimization Infrastructure via Wisdom Graph. MEGA (Meta Evaluation-Grounded Adaptation) is presented as a self-evolving... Lost in Reconstruction: Aligning Action Representations with Language in Vision-Language-Action Models. The paper introduces SALT, a Semantically ALigned action Tokenizer, to address... DuplexWorld: Can voice agents help you get through the day?. DUPLEX WORLD introduces a benchmark for holistically evaluating... Curate Before You Connect: Identity and Ontology Tagging in a Production Knowledge Graph. This paper describes the ingestion and ontology-tagging layer that turns a... Batch Size or Negatives? A Selection Rule for Memory-Constrained Recommender Training. Large-scale neural recommender systems are typically trained with a softmax... Fisher8: Stabilizing Neural Heteroscedastic Regression via Output-Layer Fisher Geometry. Fisher8: Stabilizing Neural Heteroscedastic Regression via Output-Layer Fisher... FiGuRO: Intrinsic Dimension Estimation for Multi-Modal Data. FiGuRO: Intrinsic Dimension Estimation for Multi-Modal Data Viktoria Schuster,... Conversational versus Dashboard Explainable AI for UAV Intrusion Detection: An Empirical Study of Operator Trust and Reliance. This paper proposes a Conversational XAI interface powered by Large Language... Evaluation Resolution Confounds Learning-Rule Comparisons in Model-Brain RSA of Early Visual Cortex. The paper investigates a methodological confound in representational similarity... Predictive Allostatic Organization in Recurrent and Spiking Agents Under Partial Observability. Adaptive behavior under partial observability depends on internal organization... Stay or Stray - A Dynamical Systems Viewpoint of Popularity Bias. The paper "Stay or Stray – A Dynamical Systems Viewpoint of Popularity Bias"... SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning. SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning... Exploration-Driven Personalized Federated Reinforcement Learning via Intrinsic Motivation. Exploration-Driven Personalized Federated Reinforcement Learning via Intrinsic... MERA: Model Evolution and Routing with Skill Adaptation for Agentic Systems at Scale. MERA: Model Evolution and Routing with Skill Adaptation for Agentic Systems at... R4DSG: Relative 4D Scene Graph Memory for Object-Centric Question Answering in Long Egocentric Video. R4DSG introduces a relative 4D scene graph memory for long egocentric video,... Reinforcement Learning-Based Laser Cutting Machine Parameter Optimization. This paper presents the Reinforcement Learning for Laser Cutting (RL2C)... CARE: Confidence-Aware Reasoning for Reliable Medical VQA. CARE: Confidence-Aware Reasoning for Reliable Medical VQA proposes a framework... Analysis of Federated Aggregation under Model Poisoning and Backdoor Attacks: A Reconstructed Cross-Dataset and Cross-Architecture Benchmark. The paper presents a reconstructed comparative benchmark and evidence audit of... Threshold Structure of Optimal Policies in Restart POMDPs. We study a Restart POMDP (Partially Observable Markov Decision Process) on a... Post-Calibration Reliability Reranking of Relevance Decisions via Label-wise Monotone Projection. Web search, product search, and question-answering retrieval systems often... On the Limitations of Cross-Lingual Consistency in Multilingual Text-to-image Generation. The paper introduces LingT2I, a new benchmark designed to evaluate... Clinical Feasibility of Low-Magnification Fluorescence Imaging for Breast Cancer Margin Detection Using Texture Analysis and Deep Learning. This study investigated how 4× and 10× magnifications affect margin-level... StreamFlow: Dynamic Memory Flows for Streaming Video Understanding. StreamFlow introduces an efficient visual memory framework for streaming video... Can Released LLM Vocabularies Support Token-Level Estimation of Hidden Corpora?. The paper "Can Released LLM Vocabularies Support Token-Level Estimation of... Benchmarking LLM Judges for Mobile Agent Evaluation. M OBILE J UDGE B ENCH is introduced as "to our knowledge the first benchmark... DSAgentBench: Can Agents Automate End-to-End Data-Science Workflows in Real Computer Environments?. DSAgentBench: Can Agents Automate End-to-End Data-Science Workflows in Real... Data Attribution of Emergent Misalignment with Persona Features. Core Research Question This paper investigates emergent misalignment (EM) in... X2-Turn: Frame-Synchronous Dual-Head Modeling for Joint Streaming ASR and Turn State Prediction. X2-Turn presents a frame-synchronous turn state prediction method via... When Self-Consistency Backfires: Majority Vote Hurts the Majority of Hard Science Problems for Small LLMs. Title: When Self-Consistency Backfires: Majority Vote Hurts the Majority of... A Study of Kernel Telemetry Options for Security-Oriented Provenance. This paper studies kernel telemetry options for building the capture layer of... Knowledge-Graph-Guided Retrieval-Augmented LLMs for Explainable Root Cause Analysis in Automotive HiL Validation. This paper proposes a knowledge-graph-guided retrieval-augmented large language... How Robust Are LLMs to Vietnamese Dialects?. This paper introduces VialectBench, the first end-to-end human-annotated... DonorRank: Donor Language Selection for Low-Resource Cross-Lingual Speech Recognition. DonorRank: Donor Language Selection for Low-Resource Cross-Lingual Speech... Gloss-Free Representation Learning for Cross-Dataset Sign Spotting. Sign-language research for resource-constrained languages is often limited by... Derivative Computation in PINNs: Automatic Differentiation, Finite Differences and Beyond. The paper systematically investigates finite-difference (FD) derivative... Topological Feasibility Guarantees for Differentiable Predictive Control. This paper establishes deterministic feasibility guarantees for differentiable... Inverse Theory of Mind Modeling for Content Recommendation: From Web Browsing to Dynamic Intelligent Interfaces. The paper proposes an Inverse Theory of Mind (IToM) pipeline for content... CARD: Controlled Agentic Reddit Discussions for Credit Card Simulation. CARD: Controlled Agentic Reddit Discussions for Credit Card Simulation This... Federated Learning for Distributed CNC Tool Wear Prediction. This paper investigates federated learning for CNC tool wear prediction using... EliSeg: Verified Target Construction for Report-Grounded Abnormality Segmentation. Problem Statement The paper addresses report-grounded abnormality segmentation,... ECHO: A Locally-Deployable Agentic Health Assistant with Temporal Memory, Safety Guardrails, and Speech Assessment. the Paper Authors: Abdulkadir Küçe, Alihan Esen, Çağla Fikir, Berke Kurt,... XGBoost "is all you need": the case of forecasting transmitted heat energy in District Heating Systems. This paper presents a comparative study of two distinct approaches, XGBoost and... Blast Radius. by M.Y. Strategies to Avoid Illegal Data Access. This study examines technology solutions, personnel training, and policy... Multiclass Sentiment Analysis for Identifying Political Viewpoints. The paper investigates multiclass sentiment analysis of political viewpoints on... Reoptimization Algorithms for Contextual Bandits with Knapsack Constraints. The paper "Reoptimization Algorithms for Contextual Bandits with Knapsack... Decision-Aware Approximation of Belief Functions for Evidential Combinatorial Optimization. The paper introduces a decision-aware approximation method for belief functions... When Agents Talk: Honeytokens under Shared Memory. The paper "When Agents Talk: Honeytokens under Shared Memory" by Joshua S. Modelling Geographic Atrophy Progression using Implicit Neural Representations. Age-related Macular Degeneration (AMD) is the major cause of blindness in the... Every pooling rule has its world: matching probability combination rules to situations and stakes. Every pooling rule has its world: matching probability combination rules to... The Signal Rail: A Deterministic Motion Grammar for Communicating Conversational Agent State in Terminal Interfaces. The Signal Rail: A Deterministic Motion Grammar for Communicating... FITTER: Vocabulary-Agnostic Cross-Domain Inference on Temporal Knowledge Graphs. FITTER: Vocabulary-Agnostic Cross-Domain Inference on Temporal Knowledge Graphs... Why Post-Norm Transformers Collapse: Attention Amplification and Gradient Repair Failure. The paper "Why Post-Norm Transformers Collapse: Attention Amplification and... SkillZip: Evaluation-Free Skill Compression for Self-Evolving Agents by Discovering Reusable Structure. SkillZip: Evaluation-Free Skill Compression for Self-Evolving Agents by... MVTrack: Ultrafast Appearance-Free Moving Object Tracking from Compressed Bitstreams. MVTrack is an ultrafast tracker for moving objects that operates directly on... REDAgentBench: Executable Red Teaming and Faithful Measurement of LLM Agent Systems. REDAgentBench: Executable Red Teaming and Faithful Measurement of LLM Agent... A HamNoSys-Guided Dataset and Baselines for Fine-Grained Isolated Handshape Recognition in Sign Language. This paper introduces a new dataset and baseline models for fine-grained... A Runtime Decentralized Attestation and Coordinated Repair Framework for Securing Automotive ECUs. The paper introduces DACER, a runtime decentralized attestation and coordinated... A Modular Agentic Framework for Synthetically Constrained Multi-Objective Hit-to-Lead Optimization. SABLE (Synthetically-accessible Agentic Bayesian Ligand Exploration) is an... A Systematic Sample Size Analysis of ML-Based Path Loss Prediction for LPWAN. This paper investigates how training-set size affects the accuracy of machine... VIDS-Seg: Towards Reliable Uncertainty Quantification in Pediatric Cardiac Ultrasound Segmentation. VIDS-Seg: Towards Reliable Uncertainty Quantification in Pediatric Cardiac...

The papers