Skip to content
项目
群组
代码片段
帮助
当前项目
正在载入...
登录 / 注册
切换导航面板
R
rl-introduction
概览
Overview
Details
Activity
Cycle Analytics
版本库
Repository
Files
Commits
Branches
Tags
Contributors
Graph
Compare
Charts
问题
0
Issues
0
列表
Board
标记
里程碑
合并请求
0
Merge Requests
0
CI / CD
CI / CD
流水线
作业
日程表
图表
维基
Wiki
代码片段
Snippets
成员
Collapse sidebar
Close sidebar
活动
图像
聊天
创建新问题
作业
提交
Issue Boards
Open sidebar
wangchenglong
rl-introduction
Commits
34810b97
Commit
34810b97
authored
Jun 28, 2026
by
wangchenglong
Browse files
Options
Browse Files
Download
Email Patches
Plain Diff
update.
parent
7689c0d3
显示空白字符变更
内嵌
并排
正在显示
3 个修改的文件
包含
21 行增加
和
0 行删除
+21
-0
.gitignore
+8
-0
rl-introduction.pdf
+0
-0
section6/section6.tex
+13
-0
没有找到文件。
.gitignore
0 → 100644
查看文件 @
34810b97
*.aux
*.blg
*.fdb_latexmk
*.fls
*.log
*.out
*.synctex.gz
*.toc
rl-introduction.pdf
查看文件 @
34810b97
No preview for this file type
section6/section6.tex
查看文件 @
34810b97
\section
{
Reinforcement Learning for LLM-based Agents
}
\section
{
Reinforcement Learning for LLM-based Agents
}
What is the LLM-based agent?
Agent Reinforcement Learning
advanced Agent like openclaw
Planning
(1) use rl to gain better memory.
(2) use rl to help agent to use skill.
\ No newline at end of file
编写
预览
Markdown
格式
0%
重试
或
添加新文件
添加附件
取消
您添加了
0
人
到此讨论。请谨慎行事。
请先完成此评论的编辑!
取消
请
注册
或者
登录
后发表评论