-
Natural Language Reinforcement Learning
Paper • 2411.14251 • Published • 30 -
Benjamin-eecs/Llama-3.1-8B-Instruct-NLRL-TicTacToe-Value
Feature Extraction • 8B • Updated • 21 -
Benjamin-eecs/Llama-3.1-8B-Instruct-NLRL-TicTacToe-Policy
Feature Extraction • 8B • Updated • 26 -
Waterhorse/Llama-3.1-8B-Instruct-NLRL-Breakthrough-Value
Feature Extraction • 8B • Updated • 24
🤝 Open to Collab
Bo Liu
P(doom) Rejects the premise
AI & ML interests
None yet
Recent Activity
authored a paper about 23 hours ago
Evolving in Thought Space: Training a Small Model at Test Time Unlocks Better Discoveries authored a paper about 23 hours ago
Learning Stateful Predictive Knowledge From Experience authored a paper 9 days ago
ScholarCatalyst: A Benchmark for Retrieving Papers That Inspire New Research