Library

Subject
Tags

29,479 matches · Reinforcement Learning in Robotics

#
001Motion Planning in Urban Environments via Self-Play Reinforcement LearningOpenAlexPaperIslem Kobbi, Rocha Gonçalves, Tiago et al.Sep 15
002When to Parallelize Stochastic Exploration of Rare Rewards in Reinforcement LearningOpenAlexPaperErnesto García, Daniel Mastropietro et al.Aug 31
003Evaluating Trajectories in Agentic SystemsOpenAlexPaperNagraj NaiduAug 15
004Mathematical Foundations of Reinforcement Learning and Control TheoryOpenAlexPaperDr.J. Nagaraj4 days ago
005ROEP: A Robotics-Oriented Evaluation Protocol for Deployment-Facing Vision–Language–Action Manipulation PoliciesOpenAlexPaperSangwoo Han, Hyunguk Choi4 days ago
006Squirrel OS: A Deterministically Governed Probabilistic Neural Mesh (DGPNM) — Technical Paper SuiteOpenAlexPaperLeon Long4 days ago
007Squirrel OS: A Deterministically Governed Probabilistic Neural Mesh (DGPNM) — Technical Paper SuiteOpenAlexPaperLeon Long4 days ago
008The Weight of Silence: A Causal Case for Weights Over the Scratchpad in Latent Chess ReasoningOpenAlexPaperKshirsagar Ishan4 days ago
009The Weight of Silence: A Causal Case for Weights Over the Scratchpad in Latent Chess ReasoningOpenAlexPaperKshirsagar Ishan4 days ago
010A Hybrid Owl Search Algorithm with Lévy Flights, Opposition-Based Learning, and Adaptive Memory for Continuous OptimizationOpenAlexPaperAwaz Ahmed Shaban, Subhi R. M. Zeebaree5 days ago
011Deterministic Spiral-Time Execution Filtering for LLM-Assisted Legged Robots: A MuJoCo Quadruped Proof of ConceptOpenAlexPaperMarcel Krüger, Don Feeney5 days ago
012Narsi Regression and Narsi Intelligence: A Unified Theoretical Framework for Dynamic Representation Evolution, Recursive Cognitive Adaptation, and Self-Evolving Artificial IntelligenceOpenAlexPaperA Chaudhary5 days ago
013Narsi Regression and Narsi Intelligence: A Unified Theoretical Framework for Dynamic Representation Evolution, Recursive Cognitive Adaptation, and Self-Evolving Artificial IntelligenceOpenAlexPaperA Chaudhary5 days ago
014Reinforcement Learning with Verifiable Rewards (RLVR) and GRPO for ReasoningOpenAlexPaperDheiver Francisco Santos5 days ago
015Reinforcement Learning with Verifiable Rewards (RLVR) and GRPO for ReasoningOpenAlexPaperDheiver Francisco Santos5 days ago
016ActionShift: A Benchmark for Hidden Compositional Action-Interface Adaptation in ManipulationOpenAlexPaperKrishi Attri6 days ago
017Addressing the Sim-to-Real Gap in Reinforcement Learning for UAVs by Recovering Markovian Properties with Applications to Moving Window TraversalOpenAlexPaperNorhan Mohsen Elocla, Mohamad Chehadeh et al.6 days ago
018Impact of Architectural Heterogeneity and Reward Design on Coordination in IPPO-Based Multi-Agent Systems for Sequential SAR TasksOpenAlexPaperJulia Wróbel, Wojciech Owczarek et al.6 days ago
019P3LA: P-Gated Predictive-Policy-Language ArchitectureOpenAlexPaperChristopher D. Pang6 days ago
020P3LA: P-Gated Predictive-Policy-Language ArchitectureOpenAlexPaperChristopher D. Pang6 days ago
021Agreement Without Progress: A Preregistered Null for Multi-Agent Drift DetectionOpenAlexPaperDiego Rincon7 days ago
022Agreement Without Progress: A Preregistered Null for Multi-Agent Drift DetectionOpenAlexPaperDiego Rincon7 days ago
023DFGP: Computational framework for makespan-aware multi-robot task allocation in obstacle-rich environmentsOpenAlexPaperJangHo Seo, Joonwoo Lee7 days ago
024Hayula Research PaperOpenAlexPaperYahya Saqban7 days ago
025Hayula Research PaperOpenAlexPaperYahya Saqban7 days ago
026Orchestration Over Scale: Five Strategies for Frontier-Level AI Without Trillion-Parameter ModelsOpenAlexPaperYahya Saqban7 days ago
027Orchestration Over Scale: Five Strategies for Frontier-Level AI Without Trillion-Parameter ModelsOpenAlexPaperYahya Saqban7 days ago
028Reinforcement learningOpenAlexPaperJae-yoon Choi7 days ago
029Runtime Invariant SupervisionOpenAlexPaperYashvi Patel7 days ago
030AI & Decision Making in Game ScenariosOpenAlexPaperXinyi LiJul 23
031Fate-Coupling: A Runtime Governance Primitive for AI AlignmentOpenAlexPaperGabriel CassadyJul 23
032Filtro de Realidad v5 + Anti-Sycophancy: A model-agnostic conduct protocol for AI coding assistantsOpenAlexPaperItan Homero Ruiz-HernandezJul 23
033Filtro de Realidad v5 + Anti-Sycophancy: A model-agnostic conduct protocol for AI coding assistantsOpenAlexPaperItan Homero Ruiz-HernandezJul 23
034History-Aware C-FAME: Context-Aligned Frontier Assignment for Multi-Robot ExplorationOpenAlexPaperXueyan Yao, Matilda Isaac et al.Jul 23
035Infinite Ends From Finite Samples: Open‐Ended Goal Inference as Top‐Down Bayesian Filtering of Bottom‐Up ProposalsOpenAlexPaperTan Zhi-Xuan, Gloria Kang et al.Jul 23
036Robustness-Aware Physical AI for Multi-Functional Humanoid Robot Team Concurrency Control Under Imperfect Digital Twin InformationOpenAlexPaperRubab Anwar, Won-Tae KimJul 23
037Solving Deep Reinforcement Learning Tasks with Evolution Strategies and Linear Policy NetworksOpenAlexPaperAnnie Wong, Jacob de Nobel et al.Jul 23
038State-Fraction Coupled Quantile Network Based Distributional Reinforcement Learning for Uncertainty-Aware Decision-Making of Autonomous DrivingOpenAlexPaperYu Qiu, Zhicheng He et al.Jul 23
039When Mamba Needs AttentionOpenAlexPaperUmar AslamJul 23
040When Mamba Needs AttentionOpenAlexPaperUmar AslamJul 23
041Beyond imitation: Robots that learn to work in the real worldOpenAlexPaperJohannes A. StorkJul 22
042Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory PoliciesOpenAlexPaperJisheng Liu, Jing ZhangJul 22
043Decoupling what, how, and when for observing decision-making context in autonomous robotsOpenAlexPaperJohannes Ernst, David Lennart Risch et al.Jul 22
044Filtro de Realidad v5 + Anti-Sycophancy: A model-agnostic conduct protocol for AI coding assistantsOpenAlexPaperItan Homero Ruiz-HernandezJul 22
045Fold Go: Exact Counted Legality, Certified Solves, and Zero-Parameter Competitive PlayOpenAlexPaperMaria SmithJul 22
046From the Self-Proven Theorem to Master-Level Chess - and the Law Inside Neural NetworksOpenAlexPaperMaria SmithJul 22
047Hierarchical Multi-Agent Systems with Dynamic Task Decomposition: A Novel Adaptive FrameworkOpenAlexPaperAhmed Anifowose, Oluwole Ilesanmi et al.Jul 22
048Hierarchical Multi-Agent Systems with Dynamic Task Decomposition: A Novel Adaptive FrameworkOpenAlexPaperAhmed Anifowose, Oluwole Ilesanmi et al.Jul 22
049Kilo Code におけるカスタムAI エージェントの作成方法OpenAlexPaperKatsuhiko KawaiJul 22
050Kilo Code におけるカスタムAI エージェントの作成方法OpenAlexPaperKatsuhiko KawaiJul 22
051Performant robotic manipulation with real-world reinforcement learningOpenAlexPaperKun Lei, Hui Li et al.Jul 22
052Proximal Policy Optimization in Autonomous Driving: A Systematic Review of Methods, Imitation Learning, and Evaluation PracticesOpenAlexPaperAnas Bayu Kusuma, Yogiek Indra Kurniawan et al.Jul 22
053Safe Reinforcement Learning under Regularized Probabilistic Counterexample GuidanceOpenAlexPaperXiaotong Ji, Antonio FilieriJul 22
054WARA: A Closed-Loop Multi-Agent Framework for Wireless Optimization AutoresearchOpenAlexPaperYuan Guo, Yilong Chen et al.Jul 22
055Adaptive self-supervised learning for real-time problem solving in autonomous systemsOpenAlexPaperNagunuri Rajender, Girish Reddy Ginni et al.Jul 21
056Model-Based Prior for Model-Free Reinforcement Learning in Process Control: An Offline Methodology for Semiconductor ManufacturingOpenAlexPaperYanrong Li, Fugee Tsung et al.Jul 21
057Modeling network evolution by multi-agent reinforcement learningOpenAlexPaperDan Li, Tianwei Lin et al.Jul 21
058ORBIT: Robust Offline Reinforcement Learning via Integrated Observation Recovery with Belief-Based Decision-Making Under Observation Corruption and DelayOpenAlexPaperYibo Zhao, Zijian Cao et al.Jul 21
059Reinforcement Learning Method for Large Language Models: A Comparative Study and Empirical AnalysisOpenAlexPaperJie ZhouJul 21
060Safety-Aware Event-Triggered Intervention for Motion Planning and Decision Making in Diffusion-Based Autonomous DrivingOpenAlexPaperXuerui Fang, Hui Li et al.Jul 21

Showing 60 of 29,479 documents · scroll for more