<?xml version="1.0" encoding="utf-8" ?><rss version="2.0"><channel><title><![CDATA[深度强化学习实验室：一个“开源开放、共享共进”的强化学习学术组织。]]></title><description><![CDATA[]]></description><link>https://blog.csdn.net/deeprl</link><language>zh-cn</language><generator>https://blog.csdn.net/</generator><copyright><![CDATA[Copyright &copy; deeprl]]></copyright><item><title><![CDATA[【总结】为什么对累积奖励减去baseline项能起到减小方差的作用？]]></title><link>https://blog.csdn.net/deeprl/article/details/119902039</link><guid>https://blog.csdn.net/deeprl/article/details/119902039</guid><author>deeprl</author><pubDate>Tue, 24 Aug 2021 08:26:35 +0800</pubDate><description><![CDATA[深度强化学习实验室官网：http://www.neurondance.com/论坛：http://deeprl.neurondance.com/来源：https://zhuanlan.zh...]]></description><category></category></item><item><title><![CDATA[【模仿学习】南京大学&港中文联合总结: 29页中文详述模仿学习完整过程]]></title><link>https://blog.csdn.net/deeprl/article/details/119814483</link><guid>https://blog.csdn.net/deeprl/article/details/119814483</guid><author>deeprl</author><pubDate>Thu, 19 Aug 2021 09:54:34 +0800</pubDate><description><![CDATA[深度强化学习实验室官网：http://www.neurondance.com/论坛：http://deeprl.neurondance.com/来源：南京大学, 香港中文大学团队作者: 许...]]></description><category></category></item><item><title><![CDATA[【重磅总结】170道强化学习面试题目汇总，助力实验室RLer冲刺求职季！]]></title><link>https://blog.csdn.net/deeprl/article/details/119621650</link><guid>https://blog.csdn.net/deeprl/article/details/119621650</guid><author>deeprl</author><pubDate>Wed, 11 Aug 2021 08:37:38 +0800</pubDate><description><![CDATA[深度强化学习实验室官网：http://www.neurondance.com/论坛：http://deeprl.neurondance.com/问题汇总蒙特卡洛、TD、动态规划的关系？DQ...]]></description><category></category></item><item><title><![CDATA[【Mava】一个分布式多智能体强化学习研究框架]]></title><link>https://blog.csdn.net/deeprl/article/details/119259582</link><guid>https://blog.csdn.net/deeprl/article/details/119259582</guid><author>deeprl</author><pubDate>Fri, 30 Jul 2021 08:03:10 +0800</pubDate><description><![CDATA[深度强化学习实验室官网：http://www.neurondance.com/论坛：http://deeprl.neurondance.com/来源：https://github.com/...]]></description><category></category></item><item><title><![CDATA[CORL: 基于变量序和强化学习的因果发现算法]]></title><link>https://blog.csdn.net/deeprl/article/details/119194518</link><guid>https://blog.csdn.net/deeprl/article/details/119194518</guid><author>deeprl</author><pubDate>Wed, 28 Jul 2021 14:49:41 +0800</pubDate><description><![CDATA[深度强化学习实验室官网：http://www.neurondance.com/论坛：http://deeprl.neurondance.com/来源：诺亚实验室华为诺亚方舟实验室、西安交通...]]></description><category></category></item><item><title><![CDATA[【Peter Dayan】自然和人工强化学习的结合、以及未来的发展方向]]></title><link>https://blog.csdn.net/deeprl/article/details/119048016</link><guid>https://blog.csdn.net/deeprl/article/details/119048016</guid><author>deeprl</author><pubDate>Fri, 23 Jul 2021 10:47:33 +0800</pubDate><description><![CDATA[深度强化学习实验室官网：http://www.neurondance.com/论坛：http://deeprl.neurondance.com/来源：AI科技评论作者：Mr Bear、青暮...]]></description><category></category></item><item><title><![CDATA[【Google最新成果】使用新的物理模拟引擎加速强化学习]]></title><link>https://blog.csdn.net/deeprl/article/details/118837165</link><guid>https://blog.csdn.net/deeprl/article/details/118837165</guid><author>deeprl</author><pubDate>Fri, 16 Jul 2021 09:45:58 +0800</pubDate><description><![CDATA[深度强化学习实验室官网：http://www.neurondance.com/论坛：http://deeprl.neurondance.com/来源：GoogleAI Blog上一篇文章我...]]></description><category></category></item><item><title><![CDATA[【最新】如何降低深度强化学习研究的计算成本(Reducing the Computational Cost of DeepRL)...]]></title><link>https://blog.csdn.net/deeprl/article/details/118742252</link><guid>https://blog.csdn.net/deeprl/article/details/118742252</guid><author>deeprl</author><pubDate>Wed, 14 Jul 2021 09:39:34 +0800</pubDate><description><![CDATA[深度强化学习实验室官网：http://www.neurondance.com/论坛：http://deeprl.neurondance.com/人们普遍认为，将传统强化学习与深度神经网络相...]]></description><category></category></item><item><title><![CDATA[【ICML2021】 9篇RL论文作者汪昭然：构建“元宇宙”和理论基础，让深度强化学习从虚拟走进现实...]]></title><link>https://blog.csdn.net/deeprl/article/details/118715894</link><guid>https://blog.csdn.net/deeprl/article/details/118715894</guid><author>deeprl</author><pubDate>Tue, 13 Jul 2021 08:48:42 +0800</pubDate><description><![CDATA[深度强化学习实验室官网：http://www.neurondance.com/论坛：http://deeprl.neurondance.com/来源：转载自AI科技评论作者 | 陈彩娴深度...]]></description><category></category></item><item><title><![CDATA[【DRL4IR】SIGIR'21 -第二届信息检索深度强化学习研讨会(7月15-1)]]></title><link>https://blog.csdn.net/deeprl/article/details/118715673</link><guid>https://blog.csdn.net/deeprl/article/details/118715673</guid><author>deeprl</author><pubDate>Tue, 13 Jul 2021 08:48:42 +0800</pubDate><description><![CDATA[深度强化学习实验室官网：http://www.neurondance.com/论坛：http://deeprl.neurondance.com/会议地址：https://drl4ir.gi...]]></description><category></category></item><item><title><![CDATA[ICML RL4RealLife｜聚焦强化学习落地难题，学术与商业巨头齐聚【7月23日，不见不散】...]]></title><link>https://blog.csdn.net/deeprl/article/details/118503015</link><guid>https://blog.csdn.net/deeprl/article/details/118503015</guid><author>deeprl</author><pubDate>Mon, 05 Jul 2021 17:40:52 +0800</pubDate><description><![CDATA[深度强化学习实验室官网：http://www.neurondance.com/论坛：http://deeprl.neurondance.com/近年来，强化学习（RL）在游戏界的成功在AI...]]></description><category></category></item><item><title><![CDATA[强化学习 | 基于Novelty-Pursuit的高效探索方法]]></title><link>https://blog.csdn.net/deeprl/article/details/118005672</link><guid>https://blog.csdn.net/deeprl/article/details/118005672</guid><author>deeprl</author><pubDate>Thu, 17 Jun 2021 10:13:24 +0800</pubDate><description><![CDATA[深度强化学习实验室官网：http://www.neurondance.com/论坛：http://deeprl.neurondance.com/Li, Ziniu, and Xiong-H...]]></description><category></category></item><item><title><![CDATA[【华为诺亚方舟实验室】2022届毕业生招聘--决策(强化学习)推理方向]]></title><link>https://blog.csdn.net/deeprl/article/details/117858185</link><guid>https://blog.csdn.net/deeprl/article/details/117858185</guid><author>deeprl</author><pubDate>Sat, 12 Jun 2021 10:20:49 +0800</pubDate><description><![CDATA[深度强化学习实验室官网：http://www.neurondance.com/论坛：http://deeprl.neurondance.com/来源：华为诺亚方舟实验室官微诺亚方舟实验室（...]]></description><category></category></item><item><title><![CDATA[【Reward is enough】Sutton、DavidSilver师徒联手：奖励机制足够实现各种目标。]]></title><link>https://blog.csdn.net/deeprl/article/details/117829616</link><guid>https://blog.csdn.net/deeprl/article/details/117829616</guid><author>deeprl</author><pubDate>Fri, 11 Jun 2021 08:58:42 +0800</pubDate><description><![CDATA[深度强化学习实验室官网：http://www.neurondance.com/论坛：http://deeprl.neurondance.com/作者：小舟、陈萍文章来源：转载自机器之心(链...]]></description><category></category></item><item><title><![CDATA[【重磅最新】163篇ICML-2021强化学习领域论文整理汇总(2021.06.07)]]></title><link>https://blog.csdn.net/deeprl/article/details/117678353</link><guid>https://blog.csdn.net/deeprl/article/details/117678353</guid><author>deeprl</author><pubDate>Mon, 07 Jun 2021 07:55:28 +0800</pubDate><description><![CDATA[深度强化学习实验室官网：http://www.neurondance.com/论坛：http://deeprl.neurondance.com/作者：深度强化学习实验室来源：整理自http...]]></description><category></category></item><item><title><![CDATA[【Easy-RL】中科院-清华-北大3位作者贡献的200页强化学习总结笔记]]></title><link>https://blog.csdn.net/deeprl/article/details/117236826</link><guid>https://blog.csdn.net/deeprl/article/details/117236826</guid><author>deeprl</author><pubDate>Mon, 24 May 2021 10:04:35 +0800</pubDate><description><![CDATA[深度强化学习实验室官网：http://www.neurondance.com/论坛：http://deeprl.neurondance.com/编辑：DeepRL核心贡献者：王琦、杨毅远、...]]></description><category></category></item><item><title><![CDATA[京东 | AI人才联合培养计划（NLP项目实战）]]></title><link>https://blog.csdn.net/deeprl/article/details/117050678</link><guid>https://blog.csdn.net/deeprl/article/details/117050678</guid><author>deeprl</author><pubDate>Wed, 19 May 2021 14:36:02 +0800</pubDate><description><![CDATA[01 京东AI项目实战课程安排覆盖了从经典的机器学习、文本处理技术、序列模型、深度学习、预训练模型、知识图谱、图神经网络所有必要的技术。项目一、京东健康智能分诊项目第一周：文本处理与特征工...]]></description><category></category></item><item><title><![CDATA[【重磅推荐: 强化学习课程】清华大学李升波老师《强化学习与控制》]]></title><link>https://blog.csdn.net/deeprl/article/details/116811400</link><guid>https://blog.csdn.net/deeprl/article/details/116811400</guid><author>deeprl</author><pubDate>Fri, 14 May 2021 09:36:59 +0800</pubDate><description><![CDATA[深度强化学习实验室官网：http://www.neurondance.com/论坛：http://deeprl.neurondance.com/编辑：DeepRL《强化学习与控制》是一门由...]]></description><category></category></item><item><title><![CDATA[【拒绝内卷】狼吃羊的AI奖励机制不合理： 内卷，如何解决？]]></title><link>https://blog.csdn.net/deeprl/article/details/115038270</link><guid>https://blog.csdn.net/deeprl/article/details/115038270</guid><author>deeprl</author><pubDate>Sat, 20 Mar 2021 16:39:37 +0800</pubDate><description><![CDATA[深度强化学习实验室官网：http://www.neurondance.com/论坛：http://deeprl.neurondance.com/本文转载自：Ai科技评论作者 |耳洞打三金...]]></description><category></category></item><item><title><![CDATA[【重磅推荐】哥大开源“FinRL”: 一个用于量化金融自动交易的深度强化学习库]]></title><link>https://blog.csdn.net/deeprl/article/details/114828024</link><guid>https://blog.csdn.net/deeprl/article/details/114828024</guid><author>deeprl</author><pubDate>Mon, 15 Mar 2021 08:05:44 +0800</pubDate><description><![CDATA[深度强化学习实验室官网：http://www.neurondance.com/论坛：http://deeprl.neurondance.com/编辑：DeepRL一、关于FinRL目前，深...]]></description><category></category></item></channel></rss>