GymPod
Popular repositories Loading
-
slime-public-fork
slime-public-fork PublicForked from THUDM/slime
use this to push your changes you intend to make PRs off of
Python
-
strands-sglang
strands-sglang PublicSGLang model provider for Strands Agents for on-policy agentic RL training.
Python
-
-
-
zochaoqu-verifiers
zochaoqu-verifiers PublicForked from PrimeIntellect-ai/verifiers
Our library for RL environments + evals
Python
-
sglang
sglang PublicForked from sgl-project/sglang
SGLang is a high-performance serving framework for large language models and multimodal models.
Python
Repositories
- sglang Public Forked from sgl-project/sglang
SGLang is a high-performance serving framework for large language models and multimodal models.
- miles-upstream Public Forked from radixark/miles
Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.
- code-sandbox-bench Public
- zochaoqu-verifiers Public Forked from PrimeIntellect-ai/verifiers
Our library for RL environments + evals
- Megatron-LM Public
- Megatron-Bridge Public Forked from NVIDIA-NeMo/Megatron-Bridge
Training library for Megatron-based models with bidirectional Hugging Face conversion capability
- flash-attention Public Forked from Dao-AILab/flash-attention
Fast and memory-efficient exact attention
- slime-public-fork Public Forked from THUDM/slime
use this to push your changes you intend to make PRs off of
Top languages
Loading…
Most used topics
Loading…