https://rldm.org/
RLDM | The Multi-disciplinary Conference on Reinforcement Learning and Decision Making
multi disciplinaryreinforcement learningrldmconference
https://pwnagotchi.ai/
Pwnagotchi - Deep Reinforcement Learning instrumenting bettercap for WiFi pwning.
deep reinforcement learningpwnagotchibettercapwifipwning
https://arxiv.org/abs/2511.19399
[2511.19399] DR Tulu: Reinforcement Learning with Evolving Rubrics for Deep Research
Abstract page for arXiv paper 2511.19399: DR Tulu: Reinforcement Learning with Evolving Rubrics for Deep Research
reinforcement learning
https://openreview.net/forum?id=tsE5HLYtYg
SafeDreamer: Safe Reinforcement Learning with World Models | OpenReview
The deployment of Reinforcement Learning (RL) in real-world applications is constrained by its failure to satisfy safety criteria. Existing Safe Reinforcement...
reinforcement learningworld modelssafeopenreview
https://blog.bluedot.org/p/rlhf-limitations-for-ai-safety
Problems with Reinforcement Learning from Human Feedback (RLHF) for AI safety
Reinforcement Learning from Human Feedback (RLHF) is the primary technique currently used to align the outputs of Large Language Models (LLMs) with human...
learning from human feedbackproblems withfor aireinforcement
https://data.mendeley.com/datasets/zbg64f6vb3/1
Data for: Solving the Paint Shop Problem Using Reinforcement Learning - Mendeley Data
This repository contains the instances of the paint shop problem for our paper "Solving the Paint Shop Problem Using Reinforcement Learning". The repository...
the paint shopreinforcement learningdatasolving
https://research.tudelft.nl/en/publications/reinforcement-learning-for-smart-mobile-factory-operation-in-line/
Reinforcement Learning for Smart Mobile Factory Operation in Linear Infrastructure Projects - TU...
reinforcement learningsmart mobile
https://insights.princeton.edu/tag/reinforcement-learning/
Reinforcement learning Archives - Princeton Insights
Princeton Insights covers a variety of research, including psychology, neuroscience, chemistry, computer science, and more.
reinforcement learningarchivesprincetoninsights
https://zwang4.github.io/publication/taslp21/
Reinforcement Learning-based Dialogue Guided Event Extraction to Exploit Argument Relations | Zheng...
reinforcement learning
https://vivo.colorado.edu/display/pubid_252364
TOWARDS REINFORCEMENT LEARNING TECHNIQUES FOR SPACECRAFT AUTONOMY | CU Experts | CU Boulder
reinforcement learningtowardstechniquesspacecraftautonomy
https://huggingface.co/papers/2508.02091
Paper page - CRINN: Contrastive Reinforcement Learning for Approximate Nearest Neighbor Search
Join the discussion on this paper page
paper pagereinforcement learningnearest neighbor
https://www.hrl.uni-bonn.de/publications/2022/deep-reinforcement-learning-for-next-best-view-planning-in-agricultural-applications
Deep Reinforcement Learning for Next-Best-View Planning in Agricultural Applications
deep reinforcement learningbest viewnext
https://research.ugent.be/web/result/project/4b5d0313-60d8-11e9-aa53-555acf89d448/details/3s013819-deep-reinforcement-learning-as-a-control-strategy-for-wastewater-treatment-plants/en
Research Explorer - (3S013819) Deep reinforcement learning as a control strategy for wastewater...
Research Explorer - Basic information about research project Deep reinforcement learning as a control strategy for wastewater treatment plants (3S013819). -...
deep reinforcement learningresearch explorer
https://www.mathworks.com/help/reinforcement-learning/ug/use-visualization-to-configure-exploration.html
Configure Exploration for Reinforcement Learning Agents - MATLAB & Simulink
Use visualization to configure exploration in reinforcement learning agents.
reinforcement learningconfigureexplorationagentsmatlab
https://arxiv.org/abs/2604.05808
[2604.05808] Hierarchical Reinforcement Learning with Augmented Step-Level Transitions for LLM...
Abstract page for arXiv paper 2604.05808: Hierarchical Reinforcement Learning with Augmented Step-Level Transitions for LLM Agents
reinforcement learning
https://arxiv.org/abs/2510.08763
[2510.08763] Reinforcement Learning-Based Optimization of CT Acquisition and Reconstruction...
Abstract page for arXiv paper 2510.08763: Reinforcement Learning-Based Optimization of CT Acquisition and Reconstruction Parameters Through Virtual Imaging...
reinforcement learningbased
https://open.library.ubc.ca/soa/cIRcle/collections/ubctheses/24/items/1.0444963
Physics-based simulation and reinforcement learning control of the heating phase in the...
Learning, knowledge, research, insight: welcome to the world of UBC Library, the second-largest academic research library in Canada.
reinforcement learning
https://www.mdpi.com/1424-8220/22/5/1746
Multi-Agent Reinforcement Learning Based Fully Decentralized Dynamic Time Division Configuration...
Future network services must adapt to the highly dynamic uplink and downlink traffic. To fulfill this requirement, the 3rd Generation Partnership Project...
multi agentreinforcement learningbased
https://openresearch.surrey.ac.uk/esploro/outputs/journalArticle/Multi-Objective-Deep-Reinforcement-Learning-Assisted-Resource/99814065602346
Multi-Objective Deep Reinforcement Learning Assisted Resource Allocation for MEC-Caching-coexist...
Multi-Objective Deep Reinforcement Learning Assisted Resource Allocation for MEC-Caching-coexist System - University of Surrey - Journal article
deep reinforcement learning
https://www.cs.utexas.edu/~pstone/Papers/bib2html/b2hd-MLJ11-shivaram.html
Peter Stone: Characterizing Reinforcement Learning Methods through Parameterized Learning Problems
peter stonereinforcement learningcharacterizingmethodsproblems
https://nrc-publications.canada.ca/eng/view/object/?id=e02634fa-53d9-4666-8876-5db877efe04a
Hierarchical reinforcement learning for vehicle routing problems with time windows - NRC...
Hierarchical reinforcement learning for vehicle routing problems with time windows
reinforcement learningvehicle routingproblems withtime windowshierarchical
https://openresearch.surrey.ac.uk/esploro/outputs/conferenceProceeding/Deep-Reinforcement-Learning-for-Control-of/99522222202346
Deep Reinforcement Learning for Control of Probabilistic Boolean Networks - University of Surrey
Jan 5, 2021 - Probabilistic Boolean Networks (PBNs) were introduced as a computational model for the study of complex dynamical systems, such as Gene Regulatory Networks...
deep reinforcement learningcontrol
https://chienfeng-hub.github.io/meow/
Maximum Entropy Reinforcement Learning via Energy-Based Normalizing Flow
Maximum Entropy Reinforcement Learning via Energy-Based Normalizing Flow
reinforcement learningvia energymaximumentropybased
https://www.ri.cmu.edu/event/robust-adaptive-reinforcement-learning-for-safety-critical-applications-via-curricular-learning/
Robust Adaptive Reinforcement Learning for Safety Critical Applications via Curricular Learning -...
Mar 13, 2023 - Abstract: Reinforcement Learning (RL) presents great promises for autonomous agents. However, when using robots in a safety critical domain, a system has to be...
reinforcement learningsafety criticalrobustadaptiveapplications
https://www.coursera.org/courses?query=reinforcement%20learning&page=393
Top Reinforcement Learning Courses - Learn Reinforcement Learning Online
Reinforcement Learning courses from top universities and industry leaders. Learn Reinforcement Learning online with courses like MOST from a Conceptual...
reinforcement learningtopcoursesonline
https://intra.kth.se/en/aktuellt/kalender/multi-agent-reinforcement-learning-for-enhanced-turbulence-control-in-bluff-bodies-1.1369885?date=2024-12-05&orgdate=2024-04-19&length=1&orglength=257
Multi-agent reinforcement learning for enhanced turbulence control in bluff bodies | KTH
multi agentreinforcement learning
https://ben-eysenbach.github.io/women-in-rl/
Notable Women+ in Reinforcement Learning | Benjamin Eysenbach
notable womenreinforcement learningbenjamin
https://arxiv.org/abs/2306.16021
[2306.16021] Structure in Deep Reinforcement Learning: A Survey and Open Problems
Abstract page for arXiv paper 2306.16021: Structure in Deep Reinforcement Learning: A Survey and Open Problems
deep reinforcement learning
https://edu.epfl.ch/coursebook/fr/deep-reinforcement-learning-CS-456
Deep reinforcement learning - CS-456 - EPFL
This course provides an overview and introduces modern methods for reinforcement learning (RL.) The course starts with the fundamentals of RL, such as...
deep reinforcement learningcsepfl
https://la.mathworks.com/help/reinforcement-learning/ref/rl.logging.filelogger.html
FileLogger - Log reinforcement learning training data to MAT files - MATLAB
Use a FileLogger object to log data to MAT files, within the train function or inside a custom training loop.
reinforcement learningtraining datalogmatfiles
https://fr.mathworks.com/help/reinforcement-learning/ref/rl.env.abstractenv.runepisode.html
runEpisode - Simulate reinforcement learning environment against policy or agent - MATLAB
Use runEpisode to simulate an environment with a policy or agent for a whole episode.
reinforcement learningsimulateenvironmentpolicyagent
https://in.mathworks.com/help/reinforcement-learning/ug/optimize-queue-selection-strategy-using-reinforcement-learning.html
Optimizing Queue Selection Strategies Using Reinforcement Learning - MATLAB & Simulink
Train a DQN agent to optimally route customers through a multi-queue checkout system.
reinforcement learningoptimizingqueueselectionstrategies
https://debjitpaul.github.io/blog/tag/reinforcement-learning/
reinforcement-learning | Debjit Paul
Updating it, Work in Progress, My personal website
reinforcement learningpaul
https://mediaspace.illinois.edu/media/t/1_yit4akqd
CS 440/ECE448 Fall 2024 (Reinforcement Learning 2) - Illinois Media Space
CS 440/ECE448 Fall 2024
reinforcement learningcsfall
https://nl.mathworks.com/help/reinforcement-learning/ug/reinforcement-learning-environments.html
Reinforcement Learning Environments - MATLAB & Simulink
Model environment dynamics using a MATLAB object that generates rewards and observations in response to agents actions.
reinforcement learningenvironmentsmatlabsimulink
https://www.ovhcloud.com/en-ca/learn/what-is-rlhf/
What is reinforcement learning from human feedback (RLHF)? | OVHcloud Canada
Discover Reinforcement Learning from Human Feedback: AI trained from human feedback for more relevant decisions.
learning from human feedbackwhat isreinforcementrlhfovhcloud
https://cwatson1998.github.io/publication/f2-2025
F2: Offline Reinforcement Learning for Hamiltonian Simulation via Free-Fermionic Subroutine...
reinforcement learningoffline
https://ai.g2.com/marketplace?tag=reinforcement-learning
Best AI Tools for Reinforcement Learning | G2
Discover the best AI tools and agents for reinforcement-learning. Browse verified tools with pricing, features, and reviews on G2's AI Marketplace.
best ai toolsreinforcement learning
https://jp.mathworks.com/help/reinforcement-learning/ref/rl.env.future.html
Future - Object that supports deferred outputs for reinforcement learning environment simulations...
When runEpisode runs in the background it returns a Future object as a result.
reinforcement learningfutureobjectsupportsdeferred
https://adelaide.edu.au/study/courses/math-6019/
Reinforcement Learning | Adelaide University
reinforcement learningadelaideuniversity
https://au.mathworks.com/videos/what-is-reinforcement-learning-toolbox-1561618843923.html
What Is Reinforcement Learning Toolbox? - MATLAB
Reinforcement Learning Toolbox provides MATLAB functions and Simulink blocks for training policies using reinforcement learning algorithms including DQN, A2C,...
what isreinforcement learningtoolboxmatlab
https://pubs.lib.uiowa.edu/dhm/article/id/31782/print/
Reinforcement learning with digital human models of varying visual characteristics | Proceedings of...
Digital Human Modelling (DHM) is rapidly emerging as one of the most cost-effective tools for generating computer-based virtual human-in-the-loop simulations....
reinforcement learningdigital human
https://kth.diva-portal.org/smash/record.jsf?faces-redirect=true&language=en&searchType=SIMPLE&query=&af=%5B%5D&aq=%5B%5B%5D%5D&aq2=%5B%5B%5D%5D&aqe=%5B%5D&pid=diva2%3A1804631&noOfRows=50&sortOrder=author_sort_asc&sortOrder2=title_sort_asc&onlyFullText=false&sf=all
Model-Based Reinforcement Learning for Cavity Filter Tuning
model basedreinforcement learningcavity filtertuning
https://repository.gatech.edu/entities/publication/2a35fbb1-9bbc-40f8-8914-1c302b74b45c
Integrating independent and centralized multi-agent reinforcement learning for traffic signal...
Traffic congestion in metropolitan areas is a world-wide problem that can be ameliorated by traffic lights that respond dynamically to real-time conditions....
multi agentreinforcement learningintegratingindependentcentralized
https://publications-cnrc.canada.ca/eng/view/object/?id=60096670-c232-416f-985e-a28e4ec2998e
Reinforcement learning-based controller with NMPC-assisted training for autonomous surface vessels...
Reinforcement learning-based controller with NMPC-assisted training for autonomous surface vessels
reinforcement learning
https://research-portal.uu.nl/en/publications/model-based-reinforcement-learning-for-evolving-soccer-strategies/
Model-Based Reinforcement Learning for Evolving Soccer Strategies - Utrecht University
model basedreinforcement learningevolvingsoccerstrategies
https://www.coursera.org/courses?query=reinforcement%20learning&page=587
Top Reinforcement Learning Courses - Learn Reinforcement Learning Online
Reinforcement Learning courses from top universities and industry leaders. Learn Reinforcement Learning online with courses like Analyze Data Using R for...
reinforcement learningtopcoursesonline
https://mlinscience.gitlab.io/events/250304_adv_rl/
Advances in Reinforcement Learning: Chaos and efficient search and rescue missions | Machine...
reinforcement learningrescue missionsadvanceschaos
https://rlph-workshop.github.io/index.html
Reinforcement Learning and Philosophy Workshop
Home page for Reinforcement Learning and Philosophy Workshop
reinforcement learningphilosophyworkshop
https://research.tue.nl/nl/studentTheses/reinforcement-learning-in-lifecycle-investment/
Reinforcement Learning in Lifecycle Investment - Onderzoeksportaal Eindhoven University of...
reinforcement learninglifecycleinvestmentonderzoeksportaaleindhoven
https://www.frontiersin.org/journals/psychiatry/articles/10.3389/fpsyt.2022.966369/full
Frontiers | Editorial: Computational accounts of reinforcement learning and decision making in...
Many psychiatric disorders are associated with aberrations in decision making (1). As well as having implications for patients' quality of life, such dif...
reinforcement learningdecision makingfrontierseditorialcomputational
https://andreadelprete.github.io/talk/combining-reinforcement-learning-and-trajectory-optimization/
Combining Reinforcement Learning and Trajectory Optimization | Andrea Del Prete
Jul 16, 2025 - Invited lecture at the Optimization for robotics summer school in Patras
reinforcement learningtrajectory optimizationcombiningandreadel
https://albertometelli.github.io/publication/0040-2024-No-Regret-Reinforcement-Learning-in-Smooth-MDPs
No-Regret Reinforcement Learning in Smooth MDPs - Alberto Maria Metelli, Ph.D.
no regretreinforcement learning
https://arxiv.org/abs/2510.17431
[2510.17431] Agentic Reinforcement Learning for Search is Unsafe
Abstract page for arXiv paper 2510.17431: Agentic Reinforcement Learning for Search is Unsafe
reinforcement learningagenticsearchunsafe
https://www.repository.cam.ac.uk/items/76787786-0c75-4916-8952-c763d74842f0
Reinforcement learning optimization of reaction routes on the basis of large, hybrid organic...
Computer-assisted synthesis planning (CASP) accelerates the development of organic synthesis routes of complex functional molecules. Computer-assisted...
reinforcement learning
https://cordis.europa.eu/project/id/306638/es
"Scaling Up Reinforcement Learning: Structure Learning, Skill Acquisition, and Reward Shaping" |...
"Learning how to act optimally in high-dimensional stochastic dynamic environments is a fundamental problem in many areas of engineering and computer science....
scaling upreinforcement learningskill acquisitionstructurereward
https://www.kth.se/om/upptack/kalender/disputationer/towards-safe-aligned-and-efficient-reinforcement-learning-from-human-feedback-1.1405316?date=2025-06-05&orgdate=2025-06-01&length=1&orglength=30
Towards safe, aligned, and efficient reinforcement learning from human feedback | KTH
learning from human feedbacktowardssafealignedefficient
https://la.mathworks.com/help/reinforcement-learning/ref/rl.env.rlmultiagentfunctionenv.html
rlMultiAgentFunctionEnv - Create custom multiagent reinforcement learning environment - MATLAB
Use rlMultiAgentFunctionEnv to create a custom multiagent reinforcement learning environment in which all agents execute in the same step.
create customreinforcement learningmultiagentenvironmentmatlab
https://www.mathworks.com/help/reinforcement-learning/ref/rl.agent.rlmbpoagent.html
rlMBPOAgent - Model-based policy optimization (MBPO) reinforcement learning agent - MATLAB
A model-based policy optimization (MBPO) agent is a model-based, off-policy, reinforcement learning method for environment with a discrete or continuous action...
model basedpolicy optimizationreinforcement learningagentmatlab
https://aws.amazon.com/blogs/machine-learning/optimize-customer-engagement-with-reinforcement-learning/
Optimize customer engagement with reinforcement learning | Artificial Intelligence
Mar 23, 2022 - This is a guest post co-authored by Taylor Names, Staff Machine Learning Engineer, Dev Gupta, Machine Learning Manager, and Argie Angeleas, Senior Product...
customer engagementreinforcement learningoptimizeartificialintelligence
https://kth.diva-portal.org/smash/record.jsf?faces-redirect=true&language=no&searchType=SIMPLE&query=&af=%5B%5D&aq=%5B%5B%5D%5D&aq2=%5B%5B%5D%5D&aqe=%5B%5D&pid=diva2%3A1804631&noOfRows=50&sortOrder=author_sort_asc&sortOrder2=title_sort_asc&onlyFullText=false&sf=all
Model-Based Reinforcement Learning for Cavity Filter Tuning
model basedreinforcement learningcavity filtertuning
https://uwspace.uwaterloo.ca/items/32729fb4-fa94-4185-86fb-45de19d0e590
The Reinforcement Learning Kelly Strategy
The full Kelly portfolio strategy's deficiency in the face of estimation errors in practice can be mitigated by fractional or shrinkage Kelly strategies. This...
reinforcement learningkellystrategy
https://arxiv.org/abs/2505.08827
[2505.08827] RLSR: Reinforcement Learning from Self Reward
Abstract page for arXiv paper 2505.08827: RLSR: Reinforcement Learning from Self Reward
reinforcement learningselfreward
https://nn.cs.utexas.edu/?AAAI21-jiang
Temporal-Logic-Based Reward Shaping for Continuing Reinforcement Learning Tasks
temporal logicreinforcement learningbasedrewardshaping
https://artificialintelligencesystemsauthority.com/reinforcement-learning-systems/
Reinforcement Learning Systems: Concepts and Applications
Reinforcement learning RL represents a distinct paradigm within machine learning in artificial intelligence...
reinforcement learningsystemsconceptsapplications
https://jp.mathworks.com/help/reinforcement-learning/ref/rl.env.abstractenv.runepisode.html
runEpisode - Simulate reinforcement learning environment against policy or agent - MATLAB
Use runEpisode to simulate an environment with a policy or agent for a whole episode.
reinforcement learningsimulateenvironmentpolicyagent
https://collaborate.princeton.edu/en/publications/safety-and-liveness-guarantees-through-reach-avoid-reinforcement-/fingerprints/
Safety and Liveness Guarantees through Reach-Avoid Reinforcement Learning - Fingerprint - Princeton...
reinforcement learningsafetylivenessguarantees
https://ch.mathworks.com/help/reinforcement-learning/ref/rl.agent.rltd3agent.html
rlTD3Agent - Twin-delayed deep deterministic (TD3) policy gradient reinforcement learning agent -...
The twin-delayed deep deterministic (TD3) policy gradient algorithm is an off-policy actor-critic method for environments with a continuous action-space.
policy gradientreinforcement learningtwindelayeddeep
https://www.kth.se/math/kalender/lina-palmborg-premium-control-with-reinforcement-learning-1.1201587?date=2022-10-26&orgdate=2022-09-27&length=1&orglength=0
Lina Palmborg: Premium control with reinforcement learning | KTH
reinforcement learninglinapremiumcontrolkth
https://www.mathworks.com/help/reinforcement-learning/ug/create-custom-agents.html
Create Custom Reinforcement Learning Agents - MATLAB & Simulink
Create custom agents.
create customreinforcement learningagentsmatlabsimulink
https://par.nsf.gov/biblio/10613089-experiential-explanations-reinforcement-learning
Experiential Explanations for Reinforcement Learning | NSF Public Access Repository
This page contains metadata information for the record with PAR ID 10613089
reinforcement learningpublic accessexperientialexplanationsnsf
https://arxiv.org/abs/2512.04302
[2512.04302] Towards better dense rewards in Reinforcement Learning Applications
Abstract page for arXiv paper 2512.04302: Towards better dense rewards in Reinforcement Learning Applications
reinforcement learningtowardsbetterdenserewards
https://ramagazine.ieee.org/2023/06/28/tumbling-robot-control-using-reinforcement-learning-an-adaptive-control-policy-that-transfers-well-to-the-real-world/
Tumbling Robot Control Using Reinforcement Learning: An Adaptive Control Policy That Transfers Well...
Sep 21, 2023 - Tumbling robots are simple platforms that are able to traverse large obstacles relative to their size, at the cost of being difficult to control. Existing...
robot controlreinforcement learning
https://bahh723.github.io/rl2025fa/
Reinforcement Learning (Fall 2025)
reinforcement learningfall
https://cordis.europa.eu/project/id/948671/es
Characterizing information integration in reinforcement learning: a neuro-computational...
Reinforcement learning (RL) characterizes how we adaptively learn, by trial and errors, to select actions that maximize the occurrence of rewards, and minimize...
information integrationreinforcement learningcharacterizingneurocomputational
https://www.ovhcloud.com/en-au/learn/what-is-reinforcement-learning/
What is reinforcement learning? | OVHcloud Australia
Reinforcement learning is an AI technique where agents learn to make decisions by trial and error, maximizing rewards in dynamic environments.
what isreinforcement learningovhcloudaustralia
https://eric.ed.gov/?id=EJ1044693
ERIC - EJ1044693 - Reinforcement Learning in Information Searching, Information Research: An...
Introduction: The study seeks to answer two questions: How do university students learn to use correct strategies to conduct scholarly information searches...
reinforcement learningericinformationsearchingresearch
https://vivo.colorado.edu/display/pubid_364752
A comparative analysis of reinforcement learning algorithms for earth-observing satellite...
comparative analysisreinforcement learning
https://repository.gatech.edu/entities/publication/94bbb00c-3d8d-45f6-901f-e0b34918491e
Deep Reinforcement Learning Framework for Autonomous Surface Vehicles in Environmental Cleanup
The water pollution from floating plastics poses significant environmental threats that require efficient solutions. ASV presents a promising solution to...
deep reinforcement learningframework
https://www.ideals.illinois.edu/items/109884
Improving cache replacement policy using deep reinforcement learning | IDEALS
deep reinforcement learningreplacement policyimprovingcacheusing
https://la.mathworks.com/help/reinforcement-learning/ref/rl.agent.rlppoagent.html
rlPPOAgent - Proximal policy optimization (PPO) reinforcement learning agent - MATLAB
Proximal policy optimization (PPO) is an on-policy, policy gradient reinforcement learning method for environments with a discrete or continuous action space.
proximal policy optimizationreinforcement learningppoagentmatlab
https://pmc.ncbi.nlm.nih.gov/articles/PMC12111634/
Cyber security Enhancements with reinforcement learning: A zero-day vulnerabilityu identification...
A zero-day vulnerability is a critical security weakness of software or hardware that has not yet been found and, for that reason, neither the vendor nor the...
cyber securityreinforcement learningzero dayenhancements
https://seg.inf.unibe.ch/keywords/reinforcement-learning/
Keywords: Reinforcement Learning | Software Engineering Group
Official Website of the Software Engineering Group, Institute of Computer Science, University of Bern.
reinforcement learningsoftware engineeringkeywordsgroup
https://huggingface.co/papers/2507.13158
Paper page - Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics,...
Join the discussion on this paper page
large language modelpaper pagereinforcement learning
https://collaborate.princeton.edu/en/publications/optimizing-multidocument-summarization-by-blending-reinforcement-/fingerprints/
Optimizing Multidocument Summarization by Blending Reinforcement Learning Policies - Fingerprint -...
reinforcement learningoptimizingsummarizationblendingpolicies
https://github.com/OpenMOSE/RWKV-LM-RLHF
GitHub - OpenMOSE/RWKV-LM-RLHF: Reinforcement Learning Toolkit for RWKV.(v6,v7,ARWKV)...
Reinforcement Learning Toolkit for RWKV.(v6,v7,ARWKV) Distillation,SFT,RLHF(DPO,ORPO), infinite context training, Aligning. Exploring the possibilities for...
reinforcement learning
https://collaborate.princeton.edu/en/publications/stochastic-policy-gradient-reinforcement-learning-on-a-simple-3d-/
Stochastic policy gradient reinforcement learning on a simple 3D biped - Princeton University
policy gradientreinforcement learning
https://collaborate.princeton.edu/en/publications/meta-reinforcement-learning-for-trajectory-design-in-wireless-uav-2/
Meta-Reinforcement Learning for Trajectory Design in Wireless UAV Networks - Princeton University
reinforcement learning
https://research.ibm.com/publications/optimistic-exploration-for-risk-averse-constrained-reinforcement-learning
Optimistic Exploration for Risk-Averse Constrained Reinforcement Learning for ECAI 2025 - IBM...
Optimistic Exploration for Risk-Averse Constrained Reinforcement Learning for ECAI 2025 by Radu Marinescu et al.
reinforcement learningoptimisticexplorationriskaverse
https://impact.ornl.gov/en/publications/interactive-reinforcement-learning-and-error-related-potential-cl/fingerprints/
Interactive reinforcement learning and error-related potential classification for implicit feedback...
reinforcement learninginteractiveerror
https://www.mathworks.com/help/robotics/ug/avoid-obstacles-using-reinforcement-learning-for-mobile-robots.html
Avoid Obstacles Using Reinforcement Learning for Mobile Robots - MATLAB & Simulink
Use DDPG based reinforcement learning to develop an obstacle avoidance strategy for a mobile robot.
reinforcement learningmobile robotsavoidobstaclesusing
https://research.tue.nl/en/studentTheses/reinforcement-learning-algorithms-tailored-for-the-reset-applicat/
Reinforcement Learning Algorithms Tailored for the Reset Application - Research portal Eindhoven...
reinforcement learningthe resetapplication researchalgorithmstailored
https://fr.mathworks.com/matlabcentral/answers/853165-how-to-interpret-the-learnableparameters-reinforcement-learning-toolbox
How to interpret the learnableParameters (Reinforcement Learning Toolbox)? - MATLAB Answers -...
How to interpret the learnableParameters... Learn more about reinforcement learning Simulink, Reinforcement Learning Toolbox
how toreinforcement learninginterprettoolboxmatlab
https://pure.psu.edu/en/projects/national-science-foundation-award-501/
CAREER: Securing Deep Reinforcement Learning - Penn State
deep reinforcement learningcareersecuringpennstate
https://researchconnect.buffalo.edu/en/publications/robust-multi-agent-reinforcement-learning-with-state-uncertainty/
Robust Multi-Agent Reinforcement Learning with State Uncertainty - SUNY University at Buffalo
multi agentreinforcement learning
https://dare.uva.nl/id/42e0a1e6-fb39-42b6-a4c8-e42aa6ee83c8
UvA DARE | Robustness challenges in Reinforcement Learning based time-critical cloud resource...
reinforcement learning
https://collaborate.princeton.edu/en/publications/teamwork-reinforcement-learning-with-concave-utilities/
Teamwork Reinforcement Learning With Concave Utilities - Princeton University
reinforcement learningteamworkconcaveutilitiesprinceton
https://kth.diva-portal.org/smash/record.jsf?pid=diva2:1804631
Model-Based Reinforcement Learning for Cavity Filter Tuning
model basedreinforcement learningcavity filtertuning
https://arxiv.org/abs/2511.02286v1
[2511.02286v1] Reinforcement learning based data assimilation for unknown state model
Abstract page for arXiv paper 2511.02286v1: Reinforcement learning based data assimilation for unknown state model
reinforcement learningdata assimilationbased
https://eprints.soton.ac.uk/503689/
COLERGs-constrained safe reinforcement learning for realising MASS's risk-informed collision...
reinforcement learning