An easy-to-use, fast toolkit to scale up RL post-training on a single node.
LLMInternSkill: LLM internship resume and job-search Codex Skill for resume polish, JD tailoring, evidence guard, interview grilling, and Project Scout. 大模型实习简历与求职工具箱。
The living ecosystem where AI agents complete tasks through workflow loops, improve through iterative execution, are evaluated by mentor agents or humans in the loop, and turn completed work into reusable work experience and data to improve future agents.
心理健康大模型 (LLM x Mental Health), Pre & Post-training & Dataset & Evaluation & Depoly & RAG, with InternLM / Qwen / Baichuan / DeepSeek / Mixtral / LLama / GLM series models
Official Codebase for "Neural Thickets: Diverse Task Experts Are Dense Around Pretrained Weights" (ICML 2026 Spotlight)
A curated collection of papers, technical reports, frameworks, and tools for on-policy distillation (OPD) of large language models
An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale
Open skill for capturing AI agent work as structured traces.