RUCAIBox/R1-Searcher
R1-searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning
ToolRM: Towards Agentic Tool-Use Reward Modeling
Appears on
Quick read
Latest capture 2026-09-04 03:04
0 paths
Agent instructions and tool configuration found in this repository.
No config files detected.
7 observed captures since 2026-05-23. Observed captures are shown by default.
Stars from first capture +2
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
R1-searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning
Democratizing Reinforcement Learning for LLMs
[AAAI 2026] AutoTool: Efficient Tool Selection for Large Language Model Agents
[AAAI 2026] AutoTool: Efficient Tool Selection for Large Language Model Agents
OpenClaw-RL: Train any agent simply by talking
ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning & ReCall: Learning to Reason with Tool Call for LLMs via Reinforcement Learning