https://www.learnwitharobot.com/p/training-a-reinforcement-learning
Training a Reinforcement Learning Policy and running it on Bittle Robot
reinforcement learningtrainingpolicyrunningbittle
https://scholar.hit.edu.cn/en/publications/reinforcement-learning-based-adaptive-stateless-routing-for-ambie/
Reinforcement Learning-Based Adaptive Stateless Routing for Ambient Backscatter Wireless Sensor...
reinforcement learningbasedadaptivestatelessrouting
https://bytez.com/docs/arxiv/1901.01365/paper
Hierarchical Reinforcement Learning via Advantage-Weighted Information Maximization | Read Paper on...
Jan 5, 2019 - Real-world tasks are often highly structured. Hierarchical reinforcement learning (HRL) has attracted research interest as an approach for leveraging the...
reinforcement learningread paperviaadvantageweighted
https://pure.hud.ac.uk/en/publications/an-efficient-resource-allocation-model-in-iiot-using-federated-re/
An Efficient Resource Allocation Model in IIoT Using Federated Reinforcement Learning - University...
efficient resourcereinforcement learningallocationmodeliiot
https://openreview.net/forum?id=tbFBh3LMKi
Uni-O4: Unifying Online and Offline Deep Reinforcement Learning with Multi-Step On-Policy...
Combining offline and online reinforcement learning (RL) is crucial for efficient and safe learning. However, previous approaches treat offline and online...
reinforcement learningstep onunionlineoffline
https://scienceportal.tecnalia.com/en/publications/collaborative-training-of-heterogeneous-reinforcement-learning-ag/
Collaborative training of heterogeneous reinforcement learning agents in environments with sparse...
reinforcement learningcollaborativetrainingagentsenvironments
https://www.mpib-berlin.mpg.de/events/42687/2549
LIP External Colloquium: Johannes Niediek, TU Berlin, Germany - Reinforcement learning to model...
Max Planck Institute for Human Development
tu berlinreinforcement learninglipexternalcolloquium
https://www.nisheetpatel.me/
Nisheet Patel | Home | Neuroscience & Reinforcement Learning
Nisheet Patel is a PhD grad student in theoretical neuroscience and machine learning based in Geneva, Switzerland, in the lab of Alexandre Pouget.
reinforcement learningpatelneuroscience
https://www.mdpi.com/1424-8220/23/14/6510
Optimizing Forecasted Activity Notifications with Reinforcement Learning
In this paper, we propose the notification optimization method by providing multiple alternative times as a reminder for a forecasted activity with and without...
reinforcement learningoptimizingforecastedactivitynotifications
https://ece.northsouth.edu/capstone-design/warehouse-agent-using-deep-reinforcement-learning/
Warehouse Agent using Deep Reinforcement Learning - Department of Electrical and Computer...
reinforcement learningdepartment ofwarehouseagentusing
https://arxiv.org/abs/2309.02976v2
[2309.02976v2] Natural and Robust Walking using Reinforcement Learning without Demonstrations in...
Abstract page for arXiv paper 2309.02976v2: Natural and Robust Walking using Reinforcement Learning without Demonstrations in High-Dimensional Musculoskeletal...
reinforcement learningnaturalrobustwalkingusing
https://ilkogretim-online.org/index.php/pub/article/view/3288/3204
View of INTELLIGENT MPPT CONTROL FOR WIND ENERGY CONVERSION SYSTEMS BASED REINFORCEMENT LEARNING
wind energyreinforcement learningviewintelligentmppt
https://proceedings.neurips.cc/paper_files/paper/2019/file/0e900ad84f63618452210ab8baae0218-Reviews.html
Reviews: Adaptive Auxiliary Task Weighting for Reinforcement Learning
reinforcement learningreviewsadaptiveauxiliarytask
https://www.itm-conferences.org/articles/itmconf/ref/2026/05/itmconf_issf2026_03003/itmconf_issf2026_03003.html
Reinforcement Learning Based Energy Harvesting and Data Aggregation in Wireless Sensor Networks |...
ITM Web of Conferences, open-access proceedings in information technology, computer science and mathematics
wireless sensor networksreinforcement learningenergy harvestingdata aggregationbased
https://chatpaper.com/de/paper/79615?from=subpath-venues
Reinforcement Learning Under Latent Dynamics: Toward Statistical and Algorithmic Modularity
We study the statistical requirements and algorithmic principles for reinforcement learning under general latent dynamics
reinforcement learninglatentdynamicstowardstatistical
https://pure.seoultech.ac.kr/en/publications/a-survey-on-deep-reinforcement-learning-driven-task-offloading-in/
A Survey on Deep Reinforcement Learning-driven Task Offloading in Aerial Access Networks - Seoul...
a surveyreinforcement learningaccess networksdeepdriven
https://groovesquid.com/paper/summary-of-safe-rl-saliency-aware-counterfactual-explainer-for-deep-reinforcement-learning-policies-by-amir-samadi-et-al/
Summary of Safe-rl: Saliency-aware Counterfactual Explainer For Deep Reinforcement Learning...
Jul 13, 2025 - SAFE-RL: Saliency-Aware Counterfactual Explainer for Deep Reinforcement Learning Policies by Amir Samadi, Konstantinos Koufos, Kurt Debattista, Mehrdad Dianati
reinforcement learningsummarysaferlaware
https://se.mathworks.com/help/reinforcement-learning/ref/rl.env.abstractenv.sim.html
sim - Simulate trained reinforcement learning agents within specified environment - MATLAB
This MATLAB function simulates one or more reinforcement learning agents within an environment, using default simulation options.
reinforcement learningsimtrainedagentswithin
https://mro.massey.ac.nz/items/523129da-6426-4e08-a295-8e77344d1a02
Analysis of reinforcement learning strategies for predation in a mimic-model prey environment
In this paper we propose a mathematical learning model for a stochastic automaton simulating the behaviour of a predator operating in a random environment...
reinforcement learninganalysisstrategiespredationmimic
https://www.azooptics.com/News.aspx?newsID=28401
New Simulated Photonic Reinforcement Learning Method
Aug 22, 2023 - A multidisciplinary research team at the University of Tokyo, under the direction of Hiroaki Shinkawa created an extended photonic reinforcement learning...
reinforcement learningnewsimulatedphotonicmethod
https://scholars.duke.edu/publication/1640791
Scholars@Duke publication: Behaviorally diverse traffic simulation via reinforcement learning
traffic simulationreinforcement learningscholarsdukepublication
https://researchportal.rma.ac.be/de/activities/adptsim-an-active-directory-simulator-for-reinforcement-learning-/
ADPTSim: An Active Directory Simulator for Reinforcement Learning Agents - Royal Military Academy
active directoryreinforcement learningmilitary academysimulatoragents
https://www.amazon.science/publications/reinforcement-learning-assisted-dynamic-large-scale-graph-learning
Reinforcement learning assisted dynamic large scale graph learning - Amazon Science
Graph Neural Networks (GNNs) have proven to be highly effective for link and edge prediction across domains ranging from social networks to drug discovery....
reinforcement learninglarge scaleamazon scienceassisteddynamic
https://repositum.tuwien.at/handle/20.500.12708/224882
reposiTUm: Multi-Agent Deep Reinforcement Learning for Mobile Wireless Systems: From Distributed...
multi agentreinforcement learningwireless systemsdeepmobile
https://proceedings.mlr.press/v162/sheikh22a.html
DNS: Determinantal Point Process Based Neural Network Sampler for Ensemble Reinforcement Learning
Jun 28, 2022 - DNS: Determinantal Point Process Based Neural Network Sampler for Ensemble Reinforcement LearningHassam Sheikh, Kizza Frisbee, Mariano PhielippThe ...
neural networkreinforcement learningdnspointprocess
https://proceedings.nips.cc/paper_files/paper/2016/file/c3395dd46c34fa7fd8d729d8cf88b7a8-Reviews.html
Reviews: Cooperative Inverse Reinforcement Learning
reinforcement learningreviewscooperativeinverse
https://www.e3s-conferences.org/articles/e3sconf/abs/2023/33/e3sconf_iaqvec2023_04018/e3sconf_iaqvec2023_04018.html
Comparing model predictive control and reinforcement learning for the optimal operation of...
E3S Web of Conferences, open access proceedings in environment, energy and earth sciences
model predictive controlreinforcement learningthe optimalcomparingoperation
https://huggingface.co/papers/2507.06181
Paper page - CriticLean: Critic-Guided Reinforcement Learning for Mathematical Formalization
Join the discussion on this paper page
reinforcement learningpapercriticguidedmathematical
https://journals.flvc.org/FLAIRS/article/view/141779/147209
View of Training Ethical Language Models via Reinforcement Learning from AI Feedback
language modelsreinforcement learningai feedbackviewtraining
https://research.hub.ku.edu.tr/entities/publication/a9665856-6aaa-47c9-8be2-334e12c9b993
Training socially engaging robots: modeling backchannel behaviors with batch reinforcement learning
A key aspect of social human-robot interaction is natural non-verbal communication. In this work, we train an agent with batch reinforcement learning to...
reinforcement learningtrainingengagingrobotsmodeling
https://mcml.ai/publications/ka24/
MCML - Provable Traffic Rule Compliance in Safe Reinforcement Learning on the Open Sea
Details on publication KA24
reinforcement learningon thetrafficrulecompliance
https://www.catalyzex.com/paper/discrete-control-in-real-world-driving
Discrete Control in Real-World Driving Environments using Deep Reinforcement Learning
Discrete Control in Real-World Driving Environments using Deep Reinforcement Learning: Paper and Code. Training self-driving cars is often challenging since...
real worldreinforcement learningdiscretecontroldriving
https://ai-search.io/papers/visual-cog-stage-aware-reinforcement-learning-with-chain-of-guidance-for-text-to-image-generation
Visual-CoG: Stage-Aware Reinforcement Learning with Chain of Guidance for Text-to-Image Generation...
This paper focuses on improving how well AI models create images from text descriptions, specifically when those descriptions are complex or open to...
text to imagereinforcement learningvisualcogstage
https://www.bacancytechnology.com/qanda/qa-automation/batch-size-in-background-of-deep-reinforcement-learning
Understanding Batch Size in Deep Reinforcement Learning
Discover the meaning of batch size in deep reinforcement learning. Learn how it impacts training performance and decision-making.
batch sizein deepreinforcement learningunderstanding
https://studiegids.universiteitleiden.nl/courses/79023/reinforcement-learning
Reinforcement Learning, 2018-2019 - Studiegids - Universiteit Leiden
reinforcement learninguniversiteit leidenstudiegids
https://deepai.org/publication/hierarchical-reinforcement-learning-for-sequencing-behaviors
Hierarchical Reinforcement Learning for Sequencing Behaviors | DeepAI
Mar 5, 2018 - 03/05/18 - Recent literature in the robot learning community has focused on learning robot skills that abstract out lower-level details of ro...
reinforcement learningsequencingbehaviorsdeepai
https://tldr.takara.ai/p/2509.06949
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models | Takara TLDR
We propose TraceRL, a trajectory-aware reinforcement learning framework for diffusion language models (DLMs) that incorporates preferred inference trajectory...
large language modelsreinforcement learningrevolutionizingframeworkdiffusion
https://proceedings.iclr.cc/paper_files/paper/2024/hash/10e400a587ff6925e4e26333b419ff55-Abstract-Conference.html
Robust Adversarial Reinforcement Learning via Bounded Rationality Curricula
reinforcement learningbounded rationalityrobustadversarialvia
https://www.easychair.org/publications/preprint/MhtN
Exploring Deep Reinforcement Learning for Android Malware Detection
reinforcement learningandroid malwareexploringdeepdetection
https://learningsuccess.blog/parenting/parenting-decode-behavior-foster-emotional-growth-with-positive-reinforcement/
Parenting: Decode Behavior, Foster Emotional Growth with Positive Reinforcement - Learning Success
Sep 14, 2025 - Discover the secret to understanding your child's behavior! Learn how to communicate effectively by focusing on positive actions. Transform challenges into...
emotional growthpositive reinforcementparentingdecodebehavior
https://etextpdf.com/product/foundations-of-deep-reinforcement-learning-theory-and-practice-in-python/
Foundations of Deep Reinforcement Learning: Theory and Practice in Python - eTextPdf
eBook details Authors: Laura Graesser, Wah Loon Keng File Size: 6 MB Format: PDF Length: 416 Pages Publisher: Addison-Wesley Professional; 1st edition...
theory and practicereinforcement learningfoundationsdeeppython
https://ir.cwi.nl/pub/33844
Centrum Wiskunde & Informatica: Multi-Agent Reinforcement Learning for power grid topology...
multi agentreinforcement learningpower gridcentrumwiskunde
https://www.page-meeting.org/Abstracts/population-model-enhanced-reinforcement-learning-to-enable-precision-dosing-case-study-with-dosing-of-propofol-for-sedation/
Population model-enhanced reinforcement learning to enable precision dosing: case study with dosing...
Oct 21, 2025 - Oral: Methodology - New Modelling Approaches - Population model-enhanced reinforcement learning to enable precision dosing: case study with dosing of propofol...
population modelreinforcement learningcase studyenhancedenable
https://ieeetv.ieee.org/ondemand/robust-control-in-the-worst-case-using-continuous-time-reinforcement-learning
Robust Control in the Worst Case Using Continuous Time Reinforcement Learning | IEEETV
robust controlthe worstreinforcement learningcaseusing
https://www.sciepublish.com/article/pii/267
Multi-Robot Cooperative Target Search Based on Distributed Reinforcement Learning Method in 3D...
SCIEPublish is an international open-access journal publishing service provider run by SCIE Publishing Limited, dedicated to supporting and inspiring...
target searchbased onreinforcement learningmultirobot
https://silice.es/publication/133dc678-6d6b-4c80-92c4-c96c53a92f69
Reinforcement Learning of Bipedal Walking with Musculoskeletal Models and Reference Motions |...
In this paper, we introduce a method to obtain high-quality results at a low cost for simulating musculoskeletal characters based on data from the reference...
reinforcement learningbipedalwalkingmusculoskeletalmodels
https://liner.com/review/estimating-maximum-expected-value-in-continuous-reinforcement-learning-problems
Estimating the Maximum Expected Value in Continuous Reinforcement Learning Problems [Quick Review]
Regarding this AAAI 2017 paper, this review summarizes a novel method for estimating maximum expected value in continuous reinforcement learning problems.
expected valuereinforcement learningquick reviewestimatingmaximum
https://www.knowledge-sourcing.com/report/reinforcement-learning-market
Reinforcement Learning Market Insights: Trends, Forecast 2030
reinforcement learningmarket insightstrendsforecast
https://economics.yale.edu/undergraduate/tobin-ra/2024/decoding-gamer-behavior-leveraging-inverse-reinforcement-learning-unveil-personalized-reward
Decoding Gamer Behavior: Leveraging Inverse Reinforcement Learning to Unveil Personalized Reward...
reinforcement learningdecodinggamerbehaviorleveraging
https://www.politesi.polimi.it/handle/10589/215251
A Novel Approach for Reinforcement Learning-Based Optimal Control
a novelreinforcement learningoptimal controlapproachbased
https://pure.ups.edu.ec/en/publications/deep-reinforcement-learning-based-intelligent-water-level-control/
Deep Reinforcement Learning-Based Intelligent Water Level Control: From Simulation to Embedded...
reinforcement learningwater leveldeepbasedintelligent
https://www.frontiersin.org/journals/artificial-intelligence/articles/10.3389/frai.2023.1129370/full
Frontiers | Gamma and vega hedging using deep distributional reinforcement learning
We show how D4PG can be used in conjunction with quantile regression to develop a hedging strategy for a trader responsible for derivatives that arrive stoch...
reinforcement learningfrontiersgammavegahedging
https://jobs.joinimagine.com/companies/bosch-2-18903783-a3e1-430e-b717-f6b61708c5fc/jobs/75670490-research-scientist-agentic-ai-reinforcement-learning-and-neuro-symbolic-systems-f-m-div
Research Scientist Agentic AI, Reinforcement Learning and Neuro-Symbolic Systems (f/m/div.) @ Bosch...
Search job openings across the Imagine network.
research scientistagentic aireinforcement learningneurosymbolic
https://researchconnect.buffalo.edu/en/publications/a-reinforcement-learning-based-blade-twist-angle-distribution-sea/
A reinforcement learning based blade twist angle distribution searching method for optimizing wind...
reinforcement learningbasedbladetwistangle
https://www.educative.io/courses/agentic-ai-systems/evaluating-eureka-performance-and-insights
Evaluating Eureka Agent Performance in Multi-Task Reinforcement Learning
Explore Eureka's evaluation across diverse RL environments showing its superior adaptive reward design and consistent self-improvement beyond human experts.
agent performancereinforcement learningevaluatingeurekamulti
https://jobs.hiringourheroes.org/jobs/448032716-applied-researcher-i-ai-foundations-llm-customization-finetuning-reinforcement-learning
Applied Researcher I (AI Foundations, LLM Customization, Finetuning, Reinforcement Learning) at...
ai foundationsllm customizationreinforcement learningappliedresearcher
https://twimlai.com/podcast/twimlai/deep-reinforcement-learning-for-game-testing-at-ea
Deep Reinforcement Learning for Game Testing at EA | TWIML - The Voice of Machine Learning & AI
Sep 9, 2021 - Today we're joined by Konrad Tollmar, research director at Electronic Arts and an associate professor at KTH. In our conversation, we explore his role as the...
reinforcement learninggame testingthe voicedeepmachine
https://www.eurecom.fr/en/publication/5555
Optimal trajectory of autonomous flying base stations via reinforcement learning | EURECOM
base stationsreinforcement learningoptimaltrajectoryautonomous
https://experts.umn.edu/en/publications/real-time-holding-control-for-transfer-synchronization-via-robust/fingerprints/
Real-Time Holding Control for Transfer Synchronization via Robust Multiagent Reinforcement Learning...
real timereinforcement learningholdingcontroltransfer
https://www.analogictips.com/how-do-generative-ai-deep-reinforcement-learning-and-large-language-models-optimize-eda/
How do generative AI, deep reinforcement learning, and large language models optimize EDA?
Dec 4, 2024 - Artificial intelligence (AI) and machine learning (ML) are playing an increasingly crucial role in optimizing electronic design automation (EDA) across
large language modelsgenerative aireinforcement learningdeepoptimize
https://encyclopedia.pub/entry/history/compare_revision/103121/-1
Population-Based Deep Reinforcement Learning | Encyclopedia MDPI
Encyclopedia is a user-generated content hub aiming to provide a comprehensive record for scientific developments. All content free to post, read, share and...
reinforcement learningpopulationbaseddeepencyclopedia
https://www.sintef.no/en/publications/publication/0198cc8e9521-c591d021-a37e-4590-abea-8f648df65d60/
Privacy reinforcement learning for faults detection in the smart grid - SINTEF
reinforcement learningsmart gridprivacyfaultsdetection
https://eie.khpi.edu.ua/article/view/358720
Adaptive deep reinforcement learning-based control strategy for high-performance permanent magnet...
reinforcement learninghigh performancepermanent magnetadaptivedeep
https://irep.mbzuai.ac.ae/items/b623a198-11a6-43f2-a1af-01a682eb8983
Enabling Transactive Microgrids: A Multi-Agent Reinforcement Learning Hierarchical Framework
Climate change is an ever-present global challenge, requiring a significant increase in efforts toward creating a more efficient and resilient energy grid....
multi agentreinforcement learningenablingmicrogridsframework
https://researchwith.stevens.edu/en/publications/zero-shot-reinforcement-learning-on-graphs-for-autonomous-explora/
Zero-Shot Reinforcement Learning on Graphs for Autonomous Exploration under Uncertainty - Stevens...
zero shotreinforcement learninggraphsautonomousexploration
https://best-ai.org/ai-news/agentflow-a-deep-dive-into-in-the-flow-reinforcement-learning-for-ai-agents-1759979269503
AgentFlow: A Deep Dive into In-the-Flow Reinforcement Learning for AI Agents | Best-AI.org |...
Oct 9, 2025 - AgentFlow revolutionizes AI agent training through "In-the-Flow" Reinforcement Learning, enabling continuous learning, adaptation, and efficient tool use....
a deep divefor ai agentsthe flowreinforcement learningagentflow
https://researchr.org/publication/GaoMA024/reviews
Preserving Privacy During Reinforcement Learning With AI Feedback - researchr publication reviews
learning with aipublication reviewspreservingprivacyreinforcement
https://aaltodoc.aalto.fi/items/8c2bb005-e667-445a-9a14-fc42a2b85b00/full
The application of reinforcement learning (RL) in autonomous ship and collision avoidance: A...
In recent years, research on the application of reinforcement learning (RL) to improve the intelligence of ships has increased significantly. This study...
the applicationreinforcement learningcollision avoidancerlautonomous
https://www.wevolver.com/article/soft.actor.criticdeep.reinforcement.learning.with.realworld.robots
Soft Actor Critic-Deep Reinforcement Learning with Real-World Robots
We are announcing the release of our state-of-the-art off-policy model-free reinforcement learning algorithm, soft actor-critic (SAC). This algorithm has been...
reinforcement learningreal worldsoftactorcritic
https://www.coursera.org/learn/dmrol
Decision Making and Reinforcement Learning | Coursera
Offered by Columbia University. This course is an introduction to sequential decision making and reinforcement learning. We start with a ... Enroll for free.
decision makingreinforcement learningcoursera