This is the checkpoints and dataset for: From Trainee to Trainer: LLM-Designed Training Environment for RL with Multi-Agent Reasoning
AI & ML interests
Large Language Models
Papers
From Trainee to Trainer: LLM-Designed Training Environment for RL with Multi-Agent Reasoning
Attention Amnesia in Hybrid LLMs: When CoT Fine-Tuning Breaks Long-Range Recall, and How to Fix It
This is the checkpoints and dataset for: EnvFactory: Scaling Tool-Use Agents via Executable Environments Synthesis and Robust RL.
-
EnvFactory: Scaling Tool-Use Agents via Executable Environments Synthesis and Robust RL
Paper • 2605.18703 • Published • 52 -
LARK-Lab/EnvFactory-1.7B
Text Generation • 2B • Updated • 29 -
LARK-Lab/EnvFactory-4B
Text Generation • 4B • Updated • 19 -
LARK-Lab/EnvFactory-8B
Text Generation • 8B • Updated • 991 • 1
This is the checkpoints and dataset for: From Trainee to Trainer: LLM-Designed Training Environment for RL with Multi-Agent Reasoning
This is the checkpoints and dataset for: EnvFactory: Scaling Tool-Use Agents via Executable Environments Synthesis and Robust RL.
-
EnvFactory: Scaling Tool-Use Agents via Executable Environments Synthesis and Robust RL
Paper • 2605.18703 • Published • 52 -
LARK-Lab/EnvFactory-1.7B
Text Generation • 2B • Updated • 29 -
LARK-Lab/EnvFactory-4B
Text Generation • 4B • Updated • 19 -
LARK-Lab/EnvFactory-8B
Text Generation • 8B • Updated • 991 • 1
models 11
LARK-Lab/Trainee2Trainer
Text Generation • 4B • Updated • 9 • 1
LARK-Lab/SWITCH-Phase3-GRPO-LoRA-Qwen3-8B
Text Generation • Updated • 8
LARK-Lab/EnvFactory-8B
Text Generation • 8B • Updated • 991 • 1
LARK-Lab/EnvFactory-4B
Text Generation • 4B • Updated • 19
LARK-Lab/EnvFactory-1.7B
Text Generation • 2B • Updated • 29
LARK-Lab/FormalRx-1.7B
2B • Updated • 2
LARK-Lab/FormalRx-8b
8B • Updated • 3
LARK-Lab/FormalRx-4b
4B • Updated • 29
LARK-Lab/CodeScaler-8B
Text Classification • 8B • Updated • 28
LARK-Lab/CodeScaler-4B
Text Classification • 4B • Updated • 17
datasets 9
LARK-Lab/MAPF-FrozenLake-Benchmark
Viewer • Updated • 3.15k • 41 • 1
LARK-Lab/EnvFactory-SFT-DeepSeekV4Flash-OpenAI
Viewer • Updated • 3.27k • 29
LARK-Lab/EnvFactory-SFT-DeepSeekV4Flash
Viewer • Updated • 48.3k • 14
LARK-Lab/SWITCH-Math-Train
Viewer • Updated • 45.8k • 38
LARK-Lab/EnvFactory-RL
Viewer • Updated • 3.09k • 175 • 1
LARK-Lab/EnvFactory-SFT-FILTERED
Viewer • Updated • 26.5k • 128
LARK-Lab/EnvFactory-SFT-ALL
Viewer • Updated • 53.4k • 119
LARK-Lab/FormalRx-Test
Viewer • Updated • 7.03k • 49 • 2
LARK-Lab/CodeScalerPair-51K
Viewer • Updated • 51.1k • 32 • 1