-
Natural Language Reinforcement Learning
Paper • 2411.14251 • Published • 31 -
Benjamin-eecs/Llama-3.1-8B-Instruct-NLRL-TicTacToe-Value
Feature Extraction • 8B • Updated • 9 -
Benjamin-eecs/Llama-3.1-8B-Instruct-NLRL-TicTacToe-Policy
Feature Extraction • 8B • Updated • 8 -
Waterhorse/Llama-3.1-8B-Instruct-NLRL-Breakthrough-Value
Feature Extraction • 8B • Updated • 8
🤝 Open to Collab
Bo Liu
Benjamin-eecs
AI & ML interests
None yet
Recent Activity
published an article about 20 hours ago
Your Inference Server is Secretly a Learner: Reef Infrastructure for Continual Self-Improving Agents upvoted an article about 20 hours ago
Your Inference Server is Secretly a Learner: Reef Infrastructure for Continual Self-Improving Agents updated a model 21 days ago
spade-rl/SPADE-Qwen3-4B-Games