001 LAVIFT: Latent-Action-Guided Vision Fine-Tuning for Surgical Interaction Recognition arXiv Paper Jiajun Cheng, Subarna Tripathi et al. Jul 22 002 SpEmoC: A Balanced Speaker-Segment Multimodal Emotion Benchmark arXiv Paper Sania Bano, Shahzad Ahmad et al. Jul 20 003 LLMs as judges: Toward the LLM-assisted review of GSN-compliant assurance cases OpenAlex Paper Gerhard Yu, Mithila Sivakumar et al. Jul 17 004 Model Merging for Medical LVLMs: A Benchmark and a Winner-Take-All Approach arXiv Paper Lichao Mou, Shilan Zhang et al. Jul 17 005 Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge arXiv Paper Guoliang You, Hongming Li et al. Jul 12 006 SHTA: Semantic Hard Token Correction and Center Alignment for Semi-Supervised Medical Image Segmentation arXiv Paper Zhuo Zhang, Yiheng Zhong et al. Jul 8 007 TACoS: Weakly Supervised Learning of Two-Dimensional Materials from Scribble Annotations to Precise Segmentation arXiv Paper Jiabei Chen, Liping Zhang et al. Jul 8 008 Aura: Consistent Multi-Subject Video Generation via VLM-Grounded Semantic Alignment arXiv Paper Zixiang Zhou, Zhentao Yu et al. Jul 5 009 ProxyUp: Training-Free Proxy-Conditioned Video Generation for Controllable Dynamics arXiv Paper Zanwei Zhou, Jiazhong Cen et al. Jul 4 010 Reward Lightning: Fast Video Generation via Homologous Preference Distillation arXiv Paper Jiaxiang Cheng, Bing Ma et al. Jul 4 011 DetailAnywhere: Fashion Detail Generation via Cross-Modal Feature Alignment Distillation arXiv Paper Zijun Li, Yimin Zhou et al. Jul 2 012 AlterAtlas: Shifting Travel Planning from AI Generation to Validation via Persona-Driven Simulations arXiv Paper William Huang, Ruofei Du et al. Jul 1 013 Beyond Confidence in AI-Assisted Colonoscopy: A Spatial, Temporal and Quality-Aware Audit Framework for Endoscopic AI Review arXiv Paper Roberto Alcaraz Machado, Ana Lucía Rodríguez Blanco Jul 1 014 CoPersona: Collaborative Persona Graphs for Robust LLM Personalization arXiv Paper Yangtian Zhang, Leyao Wang et al. Jul 1 015 Emergence of Preferential Attachment and Glass-Ceiling Effects in Autonomous Networks of LLMs arXiv Paper Yiming Zhang, Vikram Krishnamurthy Jul 1 016 FedMark-FM: Auditable, Risk-Adjusted Data Markets for Federated Foundation-Model Adaptation arXiv Paper Phat T. Tran-Truong, X. Le et al. Jul 1 017 Flow-PIN: A Two-Stage Power-Flow-Guided Method for System-Wide Multivariate Profile Inpainting in Distribution Networks arXiv Paper Zhenghao Zhou, Yiyan Li et al. Jul 1 018 From Chaos to Clarity: A Framework for Program-Level AI Learning Outcomes arXiv Paper Grace Barkhuff, Ian Pruitt et al. Jul 1 019 GeoSearcher: Anchor-Guided Progressive Reasoning for Remote Sensing Visual Grounding with Process Supervision arXiv Paper Di Wang, Peirong Zhang et al. Jul 1 020 Anchoring on Reality: Breaking the Pseudo-Target Ceiling in Makeup Transfer arXiv Paper B. Wei, Xianhui Lin et al. Jun 30 021 Clearer Sight, Fewer Lies: Oriented Pickup Preference Optimization for Multimodal Hallucination Mitigation arXiv Paper Xin Zou, Hao Deng et al. Jun 29 022 The Verification Benchmarking Standard (Verification Intelligence series, Paper 11 of 12) OpenAlex Paper Darren Wright Jun 27 023 Tactile-WAM: Touch-Aware World Action Model with Tactile Asymmetric Attention arXiv Paper Siyu Wu, Linjing You et al. Jun 25 024 Progressive Pixel-Neighborhood Deformable Cross-Attention for Multispectral Object Detection arXiv Paper Tian Qiu, Jifeng Shen et al. Jun 23 025 Vortex: Multi-Modal Fusion System for Intelligent Video Retrieval arXiv Paper Duc M. Nguyen, Hieu-Hoc Tran-Minh et al. Jun 18 026 Admittance-Based Surface Alignment for Human-in-the-Loop Robotic Visual Inspection arXiv Paper Antara Banerjee, Colin Acton et al. Jun 17 027 V2P-Manip: Learning Dexterous Manipulation from Monocular Human Videos arXiv Paper Kai Chen, Yanming Shao et al. Jun 15 028 Keep It in Mind: User Centric Continual Spatial Intelligence Reasoning in Egocentric Video Streams arXiv Paper Yun Wang, Junbin Xiao et al. Jun 13 029 FoleyGenEx: Unified Video-to-Audio Generation with Multi-Modal Control, Temporal Alignment, and Semantic Precision arXiv Paper Shiyao Wang, Xijuan Zeng et al. Jun 12 030 AnimaSpark: A Feed-Forward Method for Animating Arbitrary 3D Objects arXiv Paper Yiming Zhao, Haoyu Sun et al. Jun 9 031 Benchmarking stereo reconstruction for 3D printable Martian terrain models arXiv Paper Joseph Wang Jun 9 032 The Consistency Illusion: How Multi-Agent Debate Hides Reasoning Misalignment arXiv Paper Xiaoyang Wang, Christopher C. Yang Jun 7 033 Toward Human-Centered Multi-Agent Systems: Integrating Cognition, Culture, Values, and Cooperation in AI Agents arXiv Paper Safia Baloch, Rahemeen Khan Jun 6 034 Discrete-WAM: Unified Discrete Vision-Action Token Editing for World-Policy Learning arXiv Paper Ziyang Yao, Haochen Liu et al. Jun 4 035 ReSAGE-PAR: Representational Similarity Assessment for Generative Expansion in Pedestrian Attribute Recognition arXiv Paper Pablo Ayuso-Albizu, Pablo Carballeira et al. Jun 4 036 A Pathology Foundation Model for Gastric Cancer with Real-World Validation arXiv Paper Li Liang, Jiabo Ma et al. Jun 3 037 Follow-Your-Preference++: Rethinking Preference Alignment for Image Inpainting arXiv Paper Junkun Yuan, YuTao Shen et al. Jun 2 038 A Sheaf Framework for Strategic Multi-Agent Systems: From Consensus to Nash Equilibria arXiv Paper M. Hernandez,, E. Sánchez-Soto Jun 1 039 Belief at Risk: Quantifying Agentic AI Model Risk with LLM-Inferred Bayesian State Filters arXiv Paper Matthew F. Dixon Jun 1 040 Embedding Semantic Risk into Distance Fields and CBFs for Online Monocular Safe Control arXiv Paper Dawei Zhang, Nuo Chen et al. Jun 1 041 How is Latin America engaging with responsible metrics? A systematic review comparing regional and global scientific production arXiv Paper Daniela Oyarz'un-Cristi, 'Alvaro Cabezas-Clavijo et al. Jun 1 042 Lattice Aggregation in Distributed Verification under Crash and Byzantine Failures arXiv Paper Gilde Valeria Rodríguez, B. Bonakdarpour et al. Jun 1 043 Matter to Mechanism: A Benchmark for AI Co-Scientists in Materials and Battery Research arXiv Paper Shashwat Sourav, Tanjin He et al. Jun 1 044 Mind Companion: An Embodied Conversational Agent for Process-Based Psychotherapy arXiv Paper Sofie Kamber, Lukas Diebold et al. Jun 1 045 Parity Selection Rule for Information and Dissipation in Driven Steady States arXiv Paper Mengqi Li, Lixin Li et al. Jun 1 046 Reflexivity as Prompt: Does Awareness of Self-Reinforcing Market Dynamics Improve LLMs as Financial Market Forecasters? arXiv Paper Eugene W Park Jun 1 047 Vibe Coding for Visualization Implementation: An Empirical Study of Practices and Challenges arXiv Paper Zhengyu Sun, Xiaolin Wen et al. Jun 1 048 Wearable Single-Lead ECG Detects Fine-Grained Structural Heart Disease Through Echo-Report Supervision arXiv Paper Chenyang He, Qinghao Zhao et al. Jun 1 049 From General Vision to Reliable Traversability Estimation: Adapting Vision Foundation Models for Unstructured Outdoor Environments arXiv Paper Ji-Hoon Hwang, Jisung Bae et al. May 28 050 Compositional Text-to-Image Generation Via Region-aware Bimodal Direct Preference Optimization arXiv Paper Zhuohan Liu, Wujian Peng et al. May 27 051 RAG-Match: Retrieval-Augmented Knowledge Injection and Hierarchical Reasoning for Calibrated Semantic Relevance arXiv Paper He Jiang, Liansheng Sun et al. May 25 052 Towards Anatomically Plausible Human Image Generation via Synthetic Localized Preferences arXiv Paper Baolu Li, Yuliang Xiu et al. May 25 053 QwenSafe: Multimodal Content Rating Description Identification via Preference-Aligned VLMs arXiv Paper Dishanika Denipitiyage, Aruna Seneviratne et al. May 20 054 Aero-World: Action-Conditioned Aerial Video Generation from Inertial Controls arXiv Paper A. Radi, Kunyang Li et al. May 19 055 Self-Creative Text-to-Object Generation using Semantic-Aware Spatial Weighting arXiv Paper Yue Yu, Haibo Chen et al. May 19 056 Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction arXiv Paper Chaoqun He, Ming Xiang et al. May 17 057 Multi-LLM Systems Exhibit Robust Semantic Collapse arXiv Paper Weiyi Kong, Shiyang Lai et al. May 16 058 Training-Free Occluded Text Rendering via Glyph Priors and Attention-Guided Semantic Blending arXiv Paper Jingqi Hou, Hongtian Wang May 16 059 WOW-Seg: A Word-free Open World Segmentation Model arXiv Paper Danyang Li, Tianhao Wu et al. May 16 060 Do Less, Achieve More: Do We Need Every-Step Optimization for RL Fine-tuning of Diffusion Models? arXiv Paper Renye Yan, Jikang Cheng et al. May 15