Library

Subject
Tags

4,925 matches · Robustness

#
001SpEmoC: A Balanced Speaker-Segment Multimodal Emotion BenchmarkarXivPaperSania Bano, Shahzad Ahmad et al.Jul 20
002Can Experts Adapt Without Training? On Test-Time Modality Generalization in MVLMsarXivPaperRaza Imam, Darakshan Rashid et al.Jul 18
003Learning Robust Execution in Robotic Manipulation with Agentic Reinforcement LearningarXivPaperXiaopeng Zhang, Yueyang Weng et al.Jul 15
004The Effect of Multi-Lingual and Keyword Adversarial Injection on LLM Relevance JudgmentarXivPaperNguyen-Thanh-Thao Vo, Duy Duong Tuong et al.Jul 11
005Dive Into the Implicit Biases of Low-rank Vision-language AlignmentarXivPaperMingjia Shi, Shuo Wang et al.Jul 9
006Data Augmentation for L2 English Speaking Assessment using TTSarXivPaperStefano Bannò, Penny Karanasou et al.Jul 1
007Dual-BEATs: Unlocking Zero-Shot Stereo Audio Perception in Audio Large Language Models via DitheringarXivPaperShuo-Chun Lin, Hen-Hsen HuangJul 1
008Optimal Transport-based Semantic Alignment for LLM-based Audio-Visual Speech RecognitionarXivPaperXugang Lu, Peng Shen et al.Jul 1
009SCOPE: Leveraging Subgoal Critiques for Code GenerationarXivPaperYueke Zhang, Yifan Zhang et al.Jul 1
010VIP-MINGLE: A Corpus for Videoconference and In-Person Multimodal Interaction in Group Language EngagementarXivPaperAndrew Chang, Abhinay K Bodi et al.Jul 1
011Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation LearningarXivPaperYanzhe Tang, Xinyu Shao et al.Jun 24
012SAGE-Nav: Leveraging LLM Planning and Alignment Fusion for Hierarchical Scene Graph-Guided NavigationarXivPaperHao Su, Yuehao Huang et al.Jun 24
013Towards Fast and Effective Long Video Understanding of Multimodal Large Language Models via Adaptive Quasi-Gaussian SamplingarXivPaperKun Zhang, Chenxin Fang et al.Jun 23
014Vortex: Multi-Modal Fusion System for Intelligent Video RetrievalarXivPaperDuc M. Nguyen, Hieu-Hoc Tran-Minh et al.Jun 18
015Object-Centric Residual RL for Zero-Shot Sim-to-Real VLA EnhancementarXivPaperKinam Kim, Namiko Saito et al.Jun 17
016EventDrive: Event Cameras for Vision-Language Driving IntelligencearXivPaperDongyue Lu, Rong Li et al.Jun 16
017ACE-Ego-0: Unifying Egocentric Human and Robotic Data for VLA PretrainingarXivPaperHao Li, Ganlong Zhao et al.Jun 15
018DifferAD-R1: A Difference-Guided IndustrialAnomaly Localization with Multimodal LargeLanguage ModelsarXivPaperDingrong Wang, Xian Tao et al.Jun 15
019GaussTrace: Provenance Analysis of 3D Gaussian Splatting Models with Evidence-based LLM ReasoningarXivPaperHaoliang Han, Ziyuan Luo et al.Jun 9
020MemoryVLA++: Temporal Modeling via Memory and Imagination in Vision-Language-Action ModelsarXivPaperHaochen Shi, Weiye Li et al.Jun 8
021Personalized and Robust Proactive Robot Assistance with Uncertainty-Guided LLM ReasoningarXivPaperÁ. González, M. H. Shovo et al.Jun 7
022Multimodal Sexism Identification and Characterization using Large Language Models and Gradient BoostingarXivPaperKyriakos Chaviaras, Maria Lymperaiou et al.Jun 4
023Drift-Augmented Scoring: Text-Derived Noise Robustness for Zero-Shot Audio-Language ClassificationarXivPaperTu Vo, S. Zaheer et al.Jun 3
024OpenEAI-Platform: An Open-source Embodied Artificial Intelligence Hardware-Software Unified PlatformarXivPaperJinyu Zhang, Luoyi Fan et al.Jun 2
025BashCoder-R1: Towards Robust and Explainable Bash Code Generation with Robustness-Aware Group Relative Policy OptimizationarXivPaperLei Yu, Peng Wang et al.Jun 1
026Missing-Token Prompted Reliability-Aware Fusion for Robust Polyglot Speaker IdentificationarXivPaperPeng Jia, Likun Dai et al.Jun 1
027Qiskit Code Migration with LLMsarXivPaperJosé Manuel Suárez, Luís Mariano Bibbó et al.Jun 1
028QoEReasoner: An Agentic Reasoning Framework for Automated and Explainable QoE Diagnosis in RANsarXivPaperQizhe Li, Hao Chen et al.Jun 1
029SemCEB: A Cardinality Estimation Benchmark for Semantic OperatorsarXivPaperAndreas Zimmerer, Claudius Kuhn et al.Jun 1
030SpeechEditBench: A Bilingual Multi-Attribute Benchmark for Instruction-Guided Speech EditingarXivPaperHanlin Zhang, Daxin Tan et al.Jun 1
031Toward Open-Set Speaker Attribute Prediction with Keyword-Appended LLM EmbeddingsarXivPaperB. So, Jaejun Lee et al.Jun 1
032FinCom: A Financial Multi-Agent Demo with Disagree-or-Commit DeliberationarXivPaperChaojie Yang, Zixiao Tan et al.May 31
033Agentic Language-to-Objective Synthesis for Optofluidic AssemblyarXivPaperI. Saraev, E. Erben et al.May 26
034STAMBRIDGE: Spectral-Temporal Amplitude-aware Mid-Feature Bridge for EEG Visual DecodingarXivPaperJiahe Meng, Weiming Zeng et al.May 22
035CLUE: Adaptively Prioritized Contextual Cues by Leveraging a Unified Semantic Map for Effective Zero-Shot Object-Goal NavigationarXivPaperTaeyun Kim, Alvin Jinsung Choi et al.May 19
036RoVLA: Multi-Consistency Constraints for Robust Vision-Language-Action ModelsarXivPaperJing-Hua Luo, Yifan Wen et al.May 19
037LISA: Language-guided Interference-aware Spatial-Frequency Attention for Driver Gaze EstimationarXivPaperJun Ma, Zhenye Yang et al.May 17
038BlockVLA: Accelerating Autoregressive VLA via Block Diffusion FinetuningarXivPaperRui-Huan Wang, Shuanghao Bai et al.May 13
039CA-GCL: Cross-Anatomy Global-Local Contrastive Learning for Robust 3D Medical Image UnderstandingarXivPaperHanwen Zhang, Yaofang Liu et al.May 13
040WinDeskGround: A Benchmark for Robust GUI Grounding in Complex Multi-Window Desktop EnvironmentsarXivPaperHao Zhao, Tianyi Chen et al.May 13
041CheXTemporal: A Dataset for Temporally-Grounded Reasoning in Chest RadiographyarXivPaperE. Prakash, Yunhe Gao et al.May 11
042A General Framework for Multimodal LLM-Based Multimedia Understanding in Large-Scale Recommendation SystemsarXivPaperYiming Zhu, Xu Liu et al.May 10
043Beyond Accuracy: Evaluating Strategy Diversity in LLM Mathematical ReasoningarXivPaperXia Yang, Xuanyi Zhang et al.May 10
044DeltaRubric: Generative Multimodal Reward Modeling via Joint Planning and VerificationarXivPaperRui Liu, Dian Yu et al.May 10
045How Much is Brain Data Worth for Machine Learning?arXivPaperLane Lewis, Zhixin Wang et al.May 10
046NEXUS: Continual Learning of Symbolic Constraints for Safe and Robust Embodied PlanningarXivPaperTiehan Cui, Peipei Liu et al.May 10
047Budgeted Subset Refinement for Execution-Aware LLM Research IdeationarXivPaperMichael ZhangMay 9
048POETS: Uncertainty-Aware LLM Optimization via Compute-Efficient Policy EnsemblesarXivPaperNicolas Menet, Andreas Krause et al.May 8
049Beyond Negative Rollouts: Positive-Only Policy Optimization with Implicit Negative GradientsarXivPaperMingwei Xu, Hao FangMay 7
050FedAttr: Towards Privacy-preserving Client-Level Attribution in Federated LLM Fine-tuningarXivPaperSuanyuan Zhang, Junfeng Guo et al.May 7
051From Storage to Experience: A Survey on the Evolution of LLM Agent Memory MechanismsarXivPaperJing Luo, Yuchen Tian et al.May 7
052Layer Collapse in Diffusion Language ModelsarXivPaperA. Conzelmann, Albert Catalan-Tatjer et al.May 7
053MELD: Multi-Task Equilibrated Learning Detector for AI-Generated TextarXivPaperChenjun Li, Chengcheng Wan et al.May 7
054On the Role of Language Representations in Auto-Bidding: Findings and ImplicationsarXivPaperGuanyu Zhu, Jining Luan et al.May 7
055Reward Shaping and Action Masking for Compositional Tasks using Behavior Trees and LLMsarXivPaperNicholas Potteiger, Ankita Samaddar et al.May 7
056CANDI: Contextual Alignment for Niche Domains Question AnsweringarXivPaperMegha Chakraborty, Darssan Eswaramoorthi et al.May 6
057Automatic Reflection Level Classification in Hungarian Student EssaysarXivPaperZsolt Csibi, Mónika Sándor et al.May 4
058Understanding Asynchronous Inference Methods for Vision-Language-Action ModelsarXivPaperAyoub AgouzoulMay 4
059FIRCE: A Framework for Intrusion Response and Conformal EvaluationarXivPaperSeth Barrett, Lin Li et al.May 3
060Self-Captioning Multimodal Interaction Tuning: Amplifying Exploitable Redundancies for Robust Vision Language ModelsarXivPaperYuriel Ryan, Hei Man Ip et al.May 3

Showing 60 of 4,925 documents · scroll for more