Robuta

https://rldm.org/ RLDM | The Multi-disciplinary Conference on Reinforcement Learning and Decision Making multi disciplinaryreinforcement learningrldmconference https://pwnagotchi.ai/ Pwnagotchi - Deep Reinforcement Learning instrumenting bettercap for WiFi pwning. deep reinforcement learningpwnagotchibettercapwifipwning https://arxiv.org/abs/2511.19399 [2511.19399] DR Tulu: Reinforcement Learning with Evolving Rubrics for Deep Research Abstract page for arXiv paper 2511.19399: DR Tulu: Reinforcement Learning with Evolving Rubrics for Deep Research reinforcement learning https://openreview.net/forum?id=tsE5HLYtYg SafeDreamer: Safe Reinforcement Learning with World Models | OpenReview The deployment of Reinforcement Learning (RL) in real-world applications is constrained by its failure to satisfy safety criteria. Existing Safe Reinforcement... reinforcement learningworld modelssafeopenreview https://blog.bluedot.org/p/rlhf-limitations-for-ai-safety Problems with Reinforcement Learning from Human Feedback (RLHF) for AI safety Reinforcement Learning from Human Feedback (RLHF) is the primary technique currently used to align the outputs of Large Language Models (LLMs) with human... learning from human feedbackproblems withfor aireinforcement https://data.mendeley.com/datasets/zbg64f6vb3/1 Data for: Solving the Paint Shop Problem Using Reinforcement Learning - Mendeley Data This repository contains the instances of the paint shop problem for our paper "Solving the Paint Shop Problem Using Reinforcement Learning". The repository... the paint shopreinforcement learningdatasolving https://research.tudelft.nl/en/publications/reinforcement-learning-for-smart-mobile-factory-operation-in-line/ Reinforcement Learning for Smart Mobile Factory Operation in Linear Infrastructure Projects - TU... reinforcement learningsmart mobile https://insights.princeton.edu/tag/reinforcement-learning/ Reinforcement learning Archives - Princeton Insights Princeton Insights covers a variety of research, including psychology, neuroscience, chemistry, computer science, and more. reinforcement learningarchivesprincetoninsights https://zwang4.github.io/publication/taslp21/ Reinforcement Learning-based Dialogue Guided Event Extraction to Exploit Argument Relations | Zheng... reinforcement learning https://vivo.colorado.edu/display/pubid_252364 TOWARDS REINFORCEMENT LEARNING TECHNIQUES FOR SPACECRAFT AUTONOMY | CU Experts | CU Boulder reinforcement learningtowardstechniquesspacecraftautonomy https://huggingface.co/papers/2508.02091 Paper page - CRINN: Contrastive Reinforcement Learning for Approximate Nearest Neighbor Search Join the discussion on this paper page paper pagereinforcement learningnearest neighbor https://www.hrl.uni-bonn.de/publications/2022/deep-reinforcement-learning-for-next-best-view-planning-in-agricultural-applications Deep Reinforcement Learning for Next-Best-View Planning in Agricultural Applications deep reinforcement learningbest viewnext https://research.ugent.be/web/result/project/4b5d0313-60d8-11e9-aa53-555acf89d448/details/3s013819-deep-reinforcement-learning-as-a-control-strategy-for-wastewater-treatment-plants/en Research Explorer - (3S013819) Deep reinforcement learning as a control strategy for wastewater... Research Explorer - Basic information about research project Deep reinforcement learning as a control strategy for wastewater treatment plants (3S013819). -... deep reinforcement learningresearch explorer https://www.mathworks.com/help/reinforcement-learning/ug/use-visualization-to-configure-exploration.html Configure Exploration for Reinforcement Learning Agents - MATLAB & Simulink Use visualization to configure exploration in reinforcement learning agents. reinforcement learningconfigureexplorationagentsmatlab https://arxiv.org/abs/2604.05808 [2604.05808] Hierarchical Reinforcement Learning with Augmented Step-Level Transitions for LLM... Abstract page for arXiv paper 2604.05808: Hierarchical Reinforcement Learning with Augmented Step-Level Transitions for LLM Agents reinforcement learning https://arxiv.org/abs/2510.08763 [2510.08763] Reinforcement Learning-Based Optimization of CT Acquisition and Reconstruction... Abstract page for arXiv paper 2510.08763: Reinforcement Learning-Based Optimization of CT Acquisition and Reconstruction Parameters Through Virtual Imaging... reinforcement learningbased https://open.library.ubc.ca/soa/cIRcle/collections/ubctheses/24/items/1.0444963 Physics-based simulation and reinforcement learning control of the heating phase in the... Learning, knowledge, research, insight: welcome to the world of UBC Library, the second-largest academic research library in Canada. reinforcement learning https://www.mdpi.com/1424-8220/22/5/1746 Multi-Agent Reinforcement Learning Based Fully Decentralized Dynamic Time Division Configuration... Future network services must adapt to the highly dynamic uplink and downlink traffic. To fulfill this requirement, the 3rd Generation Partnership Project... multi agentreinforcement learningbased https://openresearch.surrey.ac.uk/esploro/outputs/journalArticle/Multi-Objective-Deep-Reinforcement-Learning-Assisted-Resource/99814065602346 Multi-Objective Deep Reinforcement Learning Assisted Resource Allocation for MEC-Caching-coexist... Multi-Objective Deep Reinforcement Learning Assisted Resource Allocation for MEC-Caching-coexist System - University of Surrey - Journal article deep reinforcement learning https://www.cs.utexas.edu/~pstone/Papers/bib2html/b2hd-MLJ11-shivaram.html Peter Stone: Characterizing Reinforcement Learning Methods through Parameterized Learning Problems peter stonereinforcement learningcharacterizingmethodsproblems https://nrc-publications.canada.ca/eng/view/object/?id=e02634fa-53d9-4666-8876-5db877efe04a Hierarchical reinforcement learning for vehicle routing problems with time windows - NRC... Hierarchical reinforcement learning for vehicle routing problems with time windows reinforcement learningvehicle routingproblems withtime windowshierarchical https://openresearch.surrey.ac.uk/esploro/outputs/conferenceProceeding/Deep-Reinforcement-Learning-for-Control-of/99522222202346 Deep Reinforcement Learning for Control of Probabilistic Boolean Networks - University of Surrey Jan 5, 2021 - Probabilistic Boolean Networks (PBNs) were introduced as a computational model for the study of complex dynamical systems, such as Gene Regulatory Networks... deep reinforcement learningcontrol https://chienfeng-hub.github.io/meow/ Maximum Entropy Reinforcement Learning via Energy-Based Normalizing Flow Maximum Entropy Reinforcement Learning via Energy-Based Normalizing Flow reinforcement learningvia energymaximumentropybased https://www.ri.cmu.edu/event/robust-adaptive-reinforcement-learning-for-safety-critical-applications-via-curricular-learning/ Robust Adaptive Reinforcement Learning for Safety Critical Applications via Curricular Learning -... Mar 13, 2023 - Abstract: Reinforcement Learning (RL) presents great promises for autonomous agents. However, when using robots in a safety critical domain, a system has to be... reinforcement learningsafety criticalrobustadaptiveapplications https://www.coursera.org/courses?query=reinforcement%20learning&page=393 Top Reinforcement Learning Courses - Learn Reinforcement Learning Online Reinforcement Learning courses from top universities and industry leaders. Learn Reinforcement Learning online with courses like MOST from a Conceptual... reinforcement learningtopcoursesonline https://intra.kth.se/en/aktuellt/kalender/multi-agent-reinforcement-learning-for-enhanced-turbulence-control-in-bluff-bodies-1.1369885?date=2024-12-05&orgdate=2024-04-19&length=1&orglength=257 Multi-agent reinforcement learning for enhanced turbulence control in bluff bodies | KTH multi agentreinforcement learning https://ben-eysenbach.github.io/women-in-rl/ Notable Women+ in Reinforcement Learning | Benjamin Eysenbach notable womenreinforcement learningbenjamin https://arxiv.org/abs/2306.16021 [2306.16021] Structure in Deep Reinforcement Learning: A Survey and Open Problems Abstract page for arXiv paper 2306.16021: Structure in Deep Reinforcement Learning: A Survey and Open Problems deep reinforcement learning https://edu.epfl.ch/coursebook/fr/deep-reinforcement-learning-CS-456 Deep reinforcement learning - CS-456 - EPFL This course provides an overview and introduces modern methods for reinforcement learning (RL.) The course starts with the fundamentals of RL, such as... deep reinforcement learningcsepfl https://la.mathworks.com/help/reinforcement-learning/ref/rl.logging.filelogger.html FileLogger - Log reinforcement learning training data to MAT files - MATLAB Use a FileLogger object to log data to MAT files, within the train function or inside a custom training loop. reinforcement learningtraining datalogmatfiles https://fr.mathworks.com/help/reinforcement-learning/ref/rl.env.abstractenv.runepisode.html runEpisode - Simulate reinforcement learning environment against policy or agent - MATLAB Use runEpisode to simulate an environment with a policy or agent for a whole episode. reinforcement learningsimulateenvironmentpolicyagent https://in.mathworks.com/help/reinforcement-learning/ug/optimize-queue-selection-strategy-using-reinforcement-learning.html Optimizing Queue Selection Strategies Using Reinforcement Learning - MATLAB & Simulink Train a DQN agent to optimally route customers through a multi-queue checkout system. reinforcement learningoptimizingqueueselectionstrategies https://debjitpaul.github.io/blog/tag/reinforcement-learning/ reinforcement-learning | Debjit Paul Updating it, Work in Progress, My personal website reinforcement learningpaul https://mediaspace.illinois.edu/media/t/1_yit4akqd CS 440/ECE448 Fall 2024 (Reinforcement Learning 2) - Illinois Media Space CS 440/ECE448 Fall 2024 reinforcement learningcsfall https://nl.mathworks.com/help/reinforcement-learning/ug/reinforcement-learning-environments.html Reinforcement Learning Environments - MATLAB & Simulink Model environment dynamics using a MATLAB object that generates rewards and observations in response to agents actions. reinforcement learningenvironmentsmatlabsimulink https://www.ovhcloud.com/en-ca/learn/what-is-rlhf/ What is reinforcement learning from human feedback (RLHF)? | OVHcloud Canada Discover Reinforcement Learning from Human Feedback: AI trained from human feedback for more relevant decisions. learning from human feedbackwhat isreinforcementrlhfovhcloud https://cwatson1998.github.io/publication/f2-2025 F2: Offline Reinforcement Learning for Hamiltonian Simulation via Free-Fermionic Subroutine... reinforcement learningoffline https://ai.g2.com/marketplace?tag=reinforcement-learning Best AI Tools for Reinforcement Learning | G2 Discover the best AI tools and agents for reinforcement-learning. Browse verified tools with pricing, features, and reviews on G2's AI Marketplace. best ai toolsreinforcement learning https://jp.mathworks.com/help/reinforcement-learning/ref/rl.env.future.html Future - Object that supports deferred outputs for reinforcement learning environment simulations... When runEpisode runs in the background it returns a Future object as a result. reinforcement learningfutureobjectsupportsdeferred https://adelaide.edu.au/study/courses/math-6019/ Reinforcement Learning | Adelaide University reinforcement learningadelaideuniversity https://au.mathworks.com/videos/what-is-reinforcement-learning-toolbox-1561618843923.html What Is Reinforcement Learning Toolbox? - MATLAB Reinforcement Learning Toolbox provides MATLAB functions and Simulink blocks for training policies using reinforcement learning algorithms including DQN, A2C,... what isreinforcement learningtoolboxmatlab https://pubs.lib.uiowa.edu/dhm/article/id/31782/print/ Reinforcement learning with digital human models of varying visual characteristics | Proceedings of... Digital Human Modelling (DHM) is rapidly emerging as one of the most cost-effective tools for generating computer-based virtual human-in-the-loop simulations.... reinforcement learningdigital human https://kth.diva-portal.org/smash/record.jsf?faces-redirect=true&language=en&searchType=SIMPLE&query=&af=%5B%5D&aq=%5B%5B%5D%5D&aq2=%5B%5B%5D%5D&aqe=%5B%5D&pid=diva2%3A1804631&noOfRows=50&sortOrder=author_sort_asc&sortOrder2=title_sort_asc&onlyFullText=false&sf=all Model-Based Reinforcement Learning for Cavity Filter Tuning model basedreinforcement learningcavity filtertuning https://repository.gatech.edu/entities/publication/2a35fbb1-9bbc-40f8-8914-1c302b74b45c Integrating independent and centralized multi-agent reinforcement learning for traffic signal... Traffic congestion in metropolitan areas is a world-wide problem that can be ameliorated by traffic lights that respond dynamically to real-time conditions.... multi agentreinforcement learningintegratingindependentcentralized https://publications-cnrc.canada.ca/eng/view/object/?id=60096670-c232-416f-985e-a28e4ec2998e Reinforcement learning-based controller with NMPC-assisted training for autonomous surface vessels... Reinforcement learning-based controller with NMPC-assisted training for autonomous surface vessels reinforcement learning https://research-portal.uu.nl/en/publications/model-based-reinforcement-learning-for-evolving-soccer-strategies/ Model-Based Reinforcement Learning for Evolving Soccer Strategies - Utrecht University model basedreinforcement learningevolvingsoccerstrategies https://www.coursera.org/courses?query=reinforcement%20learning&page=587 Top Reinforcement Learning Courses - Learn Reinforcement Learning Online Reinforcement Learning courses from top universities and industry leaders. Learn Reinforcement Learning online with courses like Analyze Data Using R for... reinforcement learningtopcoursesonline https://mlinscience.gitlab.io/events/250304_adv_rl/ Advances in Reinforcement Learning: Chaos and efficient search and rescue missions | Machine... reinforcement learningrescue missionsadvanceschaos https://rlph-workshop.github.io/index.html Reinforcement Learning and Philosophy Workshop Home page for Reinforcement Learning and Philosophy Workshop reinforcement learningphilosophyworkshop https://research.tue.nl/nl/studentTheses/reinforcement-learning-in-lifecycle-investment/ Reinforcement Learning in Lifecycle Investment - Onderzoeksportaal Eindhoven University of... reinforcement learninglifecycleinvestmentonderzoeksportaaleindhoven https://www.frontiersin.org/journals/psychiatry/articles/10.3389/fpsyt.2022.966369/full Frontiers | Editorial: Computational accounts of reinforcement learning and decision making in... Many psychiatric disorders are associated with aberrations in decision making (1). As well as having implications for patients' quality of life, such dif... reinforcement learningdecision makingfrontierseditorialcomputational https://andreadelprete.github.io/talk/combining-reinforcement-learning-and-trajectory-optimization/ Combining Reinforcement Learning and Trajectory Optimization | Andrea Del Prete Jul 16, 2025 - Invited lecture at the Optimization for robotics summer school in Patras reinforcement learningtrajectory optimizationcombiningandreadel https://albertometelli.github.io/publication/0040-2024-No-Regret-Reinforcement-Learning-in-Smooth-MDPs No-Regret Reinforcement Learning in Smooth MDPs - Alberto Maria Metelli, Ph.D. no regretreinforcement learning https://arxiv.org/abs/2510.17431 [2510.17431] Agentic Reinforcement Learning for Search is Unsafe Abstract page for arXiv paper 2510.17431: Agentic Reinforcement Learning for Search is Unsafe reinforcement learningagenticsearchunsafe https://www.repository.cam.ac.uk/items/76787786-0c75-4916-8952-c763d74842f0 Reinforcement learning optimization of reaction routes on the basis of large, hybrid organic... Computer-assisted synthesis planning (CASP) accelerates the development of organic synthesis routes of complex functional molecules. Computer-assisted... reinforcement learning https://cordis.europa.eu/project/id/306638/es "Scaling Up Reinforcement Learning: Structure Learning, Skill Acquisition, and Reward Shaping" |... "Learning how to act optimally in high-dimensional stochastic dynamic environments is a fundamental problem in many areas of engineering and computer science.... scaling upreinforcement learningskill acquisitionstructurereward https://www.kth.se/om/upptack/kalender/disputationer/towards-safe-aligned-and-efficient-reinforcement-learning-from-human-feedback-1.1405316?date=2025-06-05&orgdate=2025-06-01&length=1&orglength=30 Towards safe, aligned, and efficient reinforcement learning from human feedback | KTH learning from human feedbacktowardssafealignedefficient https://la.mathworks.com/help/reinforcement-learning/ref/rl.env.rlmultiagentfunctionenv.html rlMultiAgentFunctionEnv - Create custom multiagent reinforcement learning environment - MATLAB Use rlMultiAgentFunctionEnv to create a custom multiagent reinforcement learning environment in which all agents execute in the same step. create customreinforcement learningmultiagentenvironmentmatlab https://www.mathworks.com/help/reinforcement-learning/ref/rl.agent.rlmbpoagent.html rlMBPOAgent - Model-based policy optimization (MBPO) reinforcement learning agent - MATLAB A model-based policy optimization (MBPO) agent is a model-based, off-policy, reinforcement learning method for environment with a discrete or continuous action... model basedpolicy optimizationreinforcement learningagentmatlab https://aws.amazon.com/blogs/machine-learning/optimize-customer-engagement-with-reinforcement-learning/ Optimize customer engagement with reinforcement learning | Artificial Intelligence Mar 23, 2022 - This is a guest post co-authored by Taylor Names, Staff Machine Learning Engineer, Dev Gupta, Machine Learning Manager, and Argie Angeleas, Senior Product... customer engagementreinforcement learningoptimizeartificialintelligence https://kth.diva-portal.org/smash/record.jsf?faces-redirect=true&language=no&searchType=SIMPLE&query=&af=%5B%5D&aq=%5B%5B%5D%5D&aq2=%5B%5B%5D%5D&aqe=%5B%5D&pid=diva2%3A1804631&noOfRows=50&sortOrder=author_sort_asc&sortOrder2=title_sort_asc&onlyFullText=false&sf=all Model-Based Reinforcement Learning for Cavity Filter Tuning model basedreinforcement learningcavity filtertuning https://uwspace.uwaterloo.ca/items/32729fb4-fa94-4185-86fb-45de19d0e590 The Reinforcement Learning Kelly Strategy The full Kelly portfolio strategy's deficiency in the face of estimation errors in practice can be mitigated by fractional or shrinkage Kelly strategies. This... reinforcement learningkellystrategy https://arxiv.org/abs/2505.08827 [2505.08827] RLSR: Reinforcement Learning from Self Reward Abstract page for arXiv paper 2505.08827: RLSR: Reinforcement Learning from Self Reward reinforcement learningselfreward https://nn.cs.utexas.edu/?AAAI21-jiang Temporal-Logic-Based Reward Shaping for Continuing Reinforcement Learning Tasks temporal logicreinforcement learningbasedrewardshaping https://artificialintelligencesystemsauthority.com/reinforcement-learning-systems/ Reinforcement Learning Systems: Concepts and Applications Reinforcement learning RL represents a distinct paradigm within machine learning in artificial intelligence... reinforcement learningsystemsconceptsapplications https://jp.mathworks.com/help/reinforcement-learning/ref/rl.env.abstractenv.runepisode.html runEpisode - Simulate reinforcement learning environment against policy or agent - MATLAB Use runEpisode to simulate an environment with a policy or agent for a whole episode. reinforcement learningsimulateenvironmentpolicyagent https://collaborate.princeton.edu/en/publications/safety-and-liveness-guarantees-through-reach-avoid-reinforcement-/fingerprints/ Safety and Liveness Guarantees through Reach-Avoid Reinforcement Learning - Fingerprint - Princeton... reinforcement learningsafetylivenessguarantees https://ch.mathworks.com/help/reinforcement-learning/ref/rl.agent.rltd3agent.html rlTD3Agent - Twin-delayed deep deterministic (TD3) policy gradient reinforcement learning agent -... The twin-delayed deep deterministic (TD3) policy gradient algorithm is an off-policy actor-critic method for environments with a continuous action-space. policy gradientreinforcement learningtwindelayeddeep https://www.kth.se/math/kalender/lina-palmborg-premium-control-with-reinforcement-learning-1.1201587?date=2022-10-26&orgdate=2022-09-27&length=1&orglength=0 Lina Palmborg: Premium control with reinforcement learning | KTH reinforcement learninglinapremiumcontrolkth https://www.mathworks.com/help/reinforcement-learning/ug/create-custom-agents.html Create Custom Reinforcement Learning Agents - MATLAB & Simulink Create custom agents. create customreinforcement learningagentsmatlabsimulink https://par.nsf.gov/biblio/10613089-experiential-explanations-reinforcement-learning Experiential Explanations for Reinforcement Learning | NSF Public Access Repository This page contains metadata information for the record with PAR ID 10613089 reinforcement learningpublic accessexperientialexplanationsnsf https://arxiv.org/abs/2512.04302 [2512.04302] Towards better dense rewards in Reinforcement Learning Applications Abstract page for arXiv paper 2512.04302: Towards better dense rewards in Reinforcement Learning Applications reinforcement learningtowardsbetterdenserewards https://ramagazine.ieee.org/2023/06/28/tumbling-robot-control-using-reinforcement-learning-an-adaptive-control-policy-that-transfers-well-to-the-real-world/ Tumbling Robot Control Using Reinforcement Learning: An Adaptive Control Policy That Transfers Well... Sep 21, 2023 - Tumbling robots are simple platforms that are able to traverse large obstacles relative to their size, at the cost of being difficult to control. Existing... robot controlreinforcement learning https://bahh723.github.io/rl2025fa/ Reinforcement Learning (Fall 2025) reinforcement learningfall https://cordis.europa.eu/project/id/948671/es Characterizing information integration in reinforcement learning: a neuro-computational... Reinforcement learning (RL) characterizes how we adaptively learn, by trial and errors, to select actions that maximize the occurrence of rewards, and minimize... information integrationreinforcement learningcharacterizingneurocomputational https://www.ovhcloud.com/en-au/learn/what-is-reinforcement-learning/ What is reinforcement learning? | OVHcloud Australia Reinforcement learning is an AI technique where agents learn to make decisions by trial and error, maximizing rewards in dynamic environments. what isreinforcement learningovhcloudaustralia https://eric.ed.gov/?id=EJ1044693 ERIC - EJ1044693 - Reinforcement Learning in Information Searching, Information Research: An... Introduction: The study seeks to answer two questions: How do university students learn to use correct strategies to conduct scholarly information searches... reinforcement learningericinformationsearchingresearch https://vivo.colorado.edu/display/pubid_364752 A comparative analysis of reinforcement learning algorithms for earth-observing satellite... comparative analysisreinforcement learning https://repository.gatech.edu/entities/publication/94bbb00c-3d8d-45f6-901f-e0b34918491e Deep Reinforcement Learning Framework for Autonomous Surface Vehicles in Environmental Cleanup The water pollution from floating plastics poses significant environmental threats that require efficient solutions. ASV presents a promising solution to... deep reinforcement learningframework https://www.ideals.illinois.edu/items/109884 Improving cache replacement policy using deep reinforcement learning | IDEALS deep reinforcement learningreplacement policyimprovingcacheusing https://la.mathworks.com/help/reinforcement-learning/ref/rl.agent.rlppoagent.html rlPPOAgent - Proximal policy optimization (PPO) reinforcement learning agent - MATLAB Proximal policy optimization (PPO) is an on-policy, policy gradient reinforcement learning method for environments with a discrete or continuous action space. proximal policy optimizationreinforcement learningppoagentmatlab https://pmc.ncbi.nlm.nih.gov/articles/PMC12111634/ Cyber security Enhancements with reinforcement learning: A zero-day vulnerabilityu identification... A zero-day vulnerability is a critical security weakness of software or hardware that has not yet been found and, for that reason, neither the vendor nor the... cyber securityreinforcement learningzero dayenhancements https://seg.inf.unibe.ch/keywords/reinforcement-learning/ Keywords: Reinforcement Learning | Software Engineering Group Official Website of the Software Engineering Group, Institute of Computer Science, University of Bern. reinforcement learningsoftware engineeringkeywordsgroup https://huggingface.co/papers/2507.13158 Paper page - Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics,... Join the discussion on this paper page large language modelpaper pagereinforcement learning https://collaborate.princeton.edu/en/publications/optimizing-multidocument-summarization-by-blending-reinforcement-/fingerprints/ Optimizing Multidocument Summarization by Blending Reinforcement Learning Policies - Fingerprint -... reinforcement learningoptimizingsummarizationblendingpolicies https://github.com/OpenMOSE/RWKV-LM-RLHF GitHub - OpenMOSE/RWKV-LM-RLHF: Reinforcement Learning Toolkit for RWKV.(v6,v7,ARWKV)... Reinforcement Learning Toolkit for RWKV.(v6,v7,ARWKV) Distillation,SFT,RLHF(DPO,ORPO), infinite context training, Aligning. Exploring the possibilities for... reinforcement learning https://collaborate.princeton.edu/en/publications/stochastic-policy-gradient-reinforcement-learning-on-a-simple-3d-/ Stochastic policy gradient reinforcement learning on a simple 3D biped - Princeton University policy gradientreinforcement learning https://collaborate.princeton.edu/en/publications/meta-reinforcement-learning-for-trajectory-design-in-wireless-uav-2/ Meta-Reinforcement Learning for Trajectory Design in Wireless UAV Networks - Princeton University reinforcement learning https://research.ibm.com/publications/optimistic-exploration-for-risk-averse-constrained-reinforcement-learning Optimistic Exploration for Risk-Averse Constrained Reinforcement Learning for ECAI 2025 - IBM... Optimistic Exploration for Risk-Averse Constrained Reinforcement Learning for ECAI 2025 by Radu Marinescu et al. reinforcement learningoptimisticexplorationriskaverse https://impact.ornl.gov/en/publications/interactive-reinforcement-learning-and-error-related-potential-cl/fingerprints/ Interactive reinforcement learning and error-related potential classification for implicit feedback... reinforcement learninginteractiveerror https://www.mathworks.com/help/robotics/ug/avoid-obstacles-using-reinforcement-learning-for-mobile-robots.html Avoid Obstacles Using Reinforcement Learning for Mobile Robots - MATLAB & Simulink Use DDPG based reinforcement learning to develop an obstacle avoidance strategy for a mobile robot. reinforcement learningmobile robotsavoidobstaclesusing https://research.tue.nl/en/studentTheses/reinforcement-learning-algorithms-tailored-for-the-reset-applicat/ Reinforcement Learning Algorithms Tailored for the Reset Application - Research portal Eindhoven... reinforcement learningthe resetapplication researchalgorithmstailored https://fr.mathworks.com/matlabcentral/answers/853165-how-to-interpret-the-learnableparameters-reinforcement-learning-toolbox How to interpret the learnableParameters (Reinforcement Learning Toolbox)? - MATLAB Answers -... How to interpret the learnableParameters... Learn more about reinforcement learning Simulink, Reinforcement Learning Toolbox how toreinforcement learninginterprettoolboxmatlab https://pure.psu.edu/en/projects/national-science-foundation-award-501/ CAREER: Securing Deep Reinforcement Learning - Penn State deep reinforcement learningcareersecuringpennstate https://researchconnect.buffalo.edu/en/publications/robust-multi-agent-reinforcement-learning-with-state-uncertainty/ Robust Multi-Agent Reinforcement Learning with State Uncertainty - SUNY University at Buffalo multi agentreinforcement learning https://dare.uva.nl/id/42e0a1e6-fb39-42b6-a4c8-e42aa6ee83c8 UvA DARE | Robustness challenges in Reinforcement Learning based time-critical cloud resource... reinforcement learning https://collaborate.princeton.edu/en/publications/teamwork-reinforcement-learning-with-concave-utilities/ Teamwork Reinforcement Learning With Concave Utilities - Princeton University reinforcement learningteamworkconcaveutilitiesprinceton https://kth.diva-portal.org/smash/record.jsf?pid=diva2:1804631 Model-Based Reinforcement Learning for Cavity Filter Tuning model basedreinforcement learningcavity filtertuning https://arxiv.org/abs/2511.02286v1 [2511.02286v1] Reinforcement learning based data assimilation for unknown state model Abstract page for arXiv paper 2511.02286v1: Reinforcement learning based data assimilation for unknown state model reinforcement learningdata assimilationbased https://eprints.soton.ac.uk/503689/ COLERGs-constrained safe reinforcement learning for realising MASS's risk-informed collision... reinforcement learning