🔄 In a Training Loop
swappy
rycerzes
AI & ML interests
Deep learning, LLMs/VLMs, RL
Recent Activity
upvoted an article about 7 hours ago
Training a coding agent using the OpenCode harness in remote HF sandboxes with TRL and OpenEnv updated a bucket 15 days ago
rycerzes/rl-envs-101-notebooks published a bucket 15 days ago
rycerzes/rl-envs-101-notebooks