001 $β$-OPSD: Deriving with Policy Optimization, Training with Self-Distillation arXiv Paper Jiawei Xu et al. Yesterday 002 (Towards) Scalable Reliable Automated Evaluation with Large Language Models arXiv Paper Bertil Braun et al. Yesterday 003 A Distributed Acoustic Sensing Dataset for Vessel Detection and Localization in Submarine Cable Protection arXiv Paper Erick Eduardo Ramirez-Torres et al. Yesterday 004 A Montage-Agnostic Encoder for Calibration-Light Cross-User Gesture Recognition from Surface Electromyography arXiv Paper Jethro Odeyemi, Christine Zhang Yesterday 005 A Query-Efficient Stochastic Volume Rendering Framework for Time-Varying Implicit Neural Volumes arXiv Paper Alper Sahistan et al. Yesterday 006 APO: Unsupervised Atomic Policy Optimization for 3D Structure Prediction of Atomic Systems arXiv Paper Shentong Mo et al. Yesterday 007 Albilich: Steerable Proof-State Orchestration for LLM-Based Mathematical Research with CAS Integration arXiv Paper Ting Gong, M. Zeng et al. Yesterday 008 AskChem: Claim-Centered Infrastructure for Chemistry Literature Synthesis arXiv Paper Bing Yan et al. Yesterday 009 AutoPref: Automatic Discovery of Task-Specific Preference Objectives for Neural Combinatorial Optimization arXiv Paper Shengda Gu, Kai Li et al. Yesterday 010 Back from the Future: Key-Value Cache Management by Counter-Causal Surprise arXiv Paper Stephen Gould, Anton van den Hengel Yesterday 011 Baikal: Structured Search for Deep Research over Data Lakes arXiv Paper Dhruv Agarwal, Rishitha Guttapalle Mohan et al. Yesterday 012 Beyond Binary Rewards: A Comparative Study of Reward Design for Reinforcement Unlearning arXiv Paper Efstratios Zaradoukas, Davide Gabrielli et al. Yesterday 013 Beyond Geometric Complementarity: Coherent Overlap in Sparse Mixture-of-Experts Routing arXiv Paper Huiyuan Tian et al. Yesterday 014 Beyond the Best Teacher: Expanding and Compressing the Reasoning Solution Manifold arXiv Paper Songshuo Lu, Zhi Chen et al. Yesterday 015 Building a User Foundation Model for the Open Web arXiv Paper Solal Vernier et al. Yesterday 016 CACHE-UK: A Stability-Aware Memory Editor for Sequentially Updated Quantized LLMs in Finance arXiv Paper Anubhav Lakra et al. Yesterday 017 Causal Discovery with Inverted Self-attention for Multivariate Time Series arXiv Paper Yusen Liu et al. Yesterday 018 Certifying when decision-time information justifies adaptive experimentation arXiv Paper Jia Bi, Samuel Pinilla et al. Yesterday 019 Change2Task: From Repository Changes to Executable Coding Agent Tasks and Environments arXiv Paper Haomin Qi et al. Yesterday 020 Chem World: A Large-Scale Benchmark and Physics-Informed Framework for Trustworthy Chemical Property Prediction arXiv Paper Tianyou Bai et al. Yesterday 021 Class-Aware Reinforcement Learning for Counterfactual Explanation Generation arXiv Paper Muhammad Adil Saleem, S. Raza et al. Yesterday 022 ClawTrack: Towards Trace-Level Evaluation and Improvement of Real-World Autonomous Agents arXiv Paper Xingjian Wu et al. Yesterday 023 Complementary Matrix-Gated QKAN Fast-Weight Programmers for Quantum Dynamics Forecasting arXiv Paper Kuo-Chung Peng, Samuel Yen-Chi Chen et al. Yesterday 024 Compliance2LoRA: On-Demand Safety Alignment on Arbitrary Policy Subsets via Hypernetwork-Generated LoRA Adapters arXiv Paper Pankayaraj Pathmanathan, Furong Huang Yesterday 025 Contrastive Concept Importance: Explaining Pairwise Class Decisions Through Automatically Extracted Concept Representations arXiv Paper Roel W. Visser, Isaac Roberts et al. Yesterday 026 Contrastive Reinforced Policy Optimization via Privileged Self-Distillation arXiv Paper Xingjian Wu et al. Yesterday 027 Cross-Embodiment Transfer via Behavior-Aligned Representations arXiv Paper A. Sridhar, Jensen Gao et al. Yesterday 028 Cybersecurity Detection Classification with Reasoning-enabled Language Models arXiv Paper Amol Khanna et al. Yesterday 029 DAS-PMVC: A Framework for Partial Multi-View Clustering via Dual Alignment and Structure Enhancement arXiv Paper Shubin Ma, Liang Zhao et al. Yesterday 030 DS@GT ARC at ImageCLEFmedical 2026: Architectural Diversity for Concept Detection and Foundation-Model Scaling for Caption Prediction in Medical Image Analysis arXiv Paper Bowen Wang, Youwen Zhang et al. Yesterday 031 Doubly Robust Functional Representation Learning for Longitudinal Causal Inference with Irregular Histories arXiv Paper Mengfei Ran et al. Yesterday 032 Driving up Inference Energy on SNNs: Per-Sample and Universal Sponge Attacks arXiv Paper Spyridon Raptis et al. Yesterday 033 Dynamic Spectral Filtering for Temporal Graph Learning: Learning Evolving Propagation Operators arXiv Paper Yang Kong Yesterday 034 EMBL AI Librarian: Life-Sciences Knowledge Layer for AI Agents arXiv Paper Luigi Sigillo et al. Yesterday 035 Echoverse: Deep, Evolving Environments for Training Computer-Use Agents at Scale arXiv Paper Yash Pandya et al. Yesterday 036 Encryption-Compatible Clustered Federated Learning via Distributed Expectation-Maximization over Metadata arXiv Paper Michael Ben Ali et al. Yesterday 037 Enhancing Irregular Time Series Forecasting with Continuous-Time Modeling Framework arXiv Paper Tianen Shen et al. Yesterday 038 Error Analysis of Neural-Network-Based Engression arXiv Paper Juntong Chen, Zijian Guo et al. Yesterday 039 Evaluation Protocols and Cross-Subject Generalization in EEG Emotion Recognition arXiv Paper Hanting Suo, Yuwen Li Yesterday 040 Event-Structured Physics-Informed Neural Networks for Differentiable Critical Clearing Boundaries arXiv Paper Baoli Hao, Chenxi Hu et al. Yesterday 041 Exact Action Values Are Not Enough: Rollout-Verified Reinforcement Fine-Tuning of a Reasoning Model for Multi-Zone VAV Control arXiv Paper Takumi Shioda, Kohei Terashima et al. Yesterday 042 Fairness Pruning: Locating Demographic Bias in GLU-MLP Layers via Differential Activations arXiv Paper Pere Martra et al. Yesterday 043 FeatFix: Reuse What You Verify through Local Exact-Feature Correction for Faster Cached Diffusion Inference arXiv Paper Hanshuai Cui, Zhiqing Tang et al. Yesterday 044 FedOGL: Combating Catastrophic Forgetting in Federated Open-World Multimodal Graph Learning arXiv Paper Zekai Chen, Haodong Lu et al. Yesterday 045 Filling the Pareto-Optimal Front for Affordance Segmentation on Embedded Devices Using RGB-D Cameras arXiv Paper Edoardo Ragusa et al. Yesterday 046 FinSMART: Financial Sentiment Analysis for Algorithmic Trading through Market-Aligned Reinforcement Learning arXiv Paper Giorgos Iacovides et al. Yesterday 047 First-order Constrained Trilevel Optimization Over Distributed Networks for Robust Coreset Selection arXiv Paper Yang Jiao, Kaixuan Jiao et al. Yesterday 048 Flux-OPD: On-Policy Distillation with Evolving Contexts arXiv Paper Yuran Wang et al. Yesterday 049 From Expert Reduction to Behavioral Divergence: Tracing Numerical State through Sparse MoE Inference arXiv Paper Tianyang Zhu Yesterday 050 Fully Inductive Cardinality Estimation arXiv Paper Tim Schwabe et al. Yesterday 051 GVR-Coder: A Visual-Feedback Framework for Structured SVG Generation in Complex Document and Meeting Scenarios arXiv Paper Yiming Xu et al. Yesterday 052 Generalization Bounds on Optimal Control for Transformer Training and Wasserstein Distributional Robustness arXiv Paper Kagan Akman, Naci Saldi et al. Yesterday 053 Generalization and Trade-off in Adversarial Training: An RKHS Perspective via Kernel Integral Operators arXiv Paper Yiling Xie et al. Yesterday 054 Gradient-free Task-Conditioned Retrieval for On-Device In-Context Learning arXiv Paper Xinyu Luo, Hui Liu et al. Yesterday 055 Graph Neural Multilevel Preconditioners for Iterative Solvers arXiv Paper Zechen Zhang et al. Yesterday 056 Graph Neural Network Force Fields for Spin Dynamics in Metallic Magnets arXiv Paper Ali Rayat et al. Yesterday 057 Group-Reflective Self-Distillation for Agentic Reinforcement Learning arXiv Paper Binbin Zheng et al. Yesterday 058 GyRot: Leveraging Hidden Synergy between Rotation and Fine-grained Group Quantization for Low-bit LLM Inference arXiv Paper Sangjin Kim, Yuseon Chou et al. Yesterday 059 HARGO: Heterogeneity-Aware Reward-Guided Optimization for RL Post-Training of LLMs on HPC Tasks arXiv Paper Tiangang Li et al. Yesterday 060 Harnessing the Potential of Optimizing Data Mixtures via Bayesian Domain Reweighting arXiv Paper Xiang Yuan, Kaiqing Lei et al. Yesterday