Talk decks β slides hosted as static Spaces. Source: https://github.com/adithya-s-k/RL_Envs_101
Adithya S K
AI & ML interests
None yet
Recent Activity
updated a model about 3 hours ago
HuggingEnvs/geoguesser-qwen3.5-4b-grpo-v3 published a model about 3 hours ago
HuggingEnvs/geoguesser-qwen3.5-4b-grpo-v3 updated a bucket about 18 hours ago
AdithyaSK/geoguesser-runsOrganizations
Repo2RLEnv β Verifiable RL Environments
Verifiable RL environments built from real GitHub repos. One dataset per pipeline. Source: https://github.com/huggingface/Repo2RLEnv
Code Reasoning
-
AdithyaSK/Qwen-0.5b-Code-Reasoning
Text Generation β’ 0.5B β’ Updated β’ 16 β’ 2 -
AdithyaSK/Qwen-1.5b-Code-Reasoning
Text Generation β’ 2B β’ Updated β’ 9 β’ 1 -
AdithyaSK/Qwen-0.5b-Code-Reasoning-v1
Text Generation β’ 0.5B β’ Updated β’ 15 β’ 2 -
AdithyaSK/Llama-3b-Code-Reasoning
Text Generation β’ 3B β’ Updated β’ 8 β’ 1
Youtube-cloner
Talks / Slides
Talk decks β slides hosted as static Spaces. Source: https://github.com/adithya-s-k/RL_Envs_101
Data Agent
Repo2RLEnv β Verifiable RL Environments
Verifiable RL environments built from real GitHub repos. One dataset per pipeline. Source: https://github.com/huggingface/Repo2RLEnv
Jupyter-Agent
Code Reasoning
-
AdithyaSK/Qwen-0.5b-Code-Reasoning
Text Generation β’ 0.5B β’ Updated β’ 16 β’ 2 -
AdithyaSK/Qwen-1.5b-Code-Reasoning
Text Generation β’ 2B β’ Updated β’ 9 β’ 1 -
AdithyaSK/Qwen-0.5b-Code-Reasoning-v1
Text Generation β’ 0.5B β’ Updated β’ 15 β’ 2 -
AdithyaSK/Llama-3b-Code-Reasoning
Text Generation β’ 3B β’ Updated β’ 8 β’ 1
Avalon
Youtube-cloner