Popular repositories Loading
-
trl
trl PublicForked from huggingface/trl
Train transformer language models with reinforcement learning.
Python
-
RL-Kernel
RL-Kernel PublicForked from RL-Align/RL-Kernel
High-performance RL post-training infrastructure. Designed to achieve bitwise operator-level train-inference consistency across heterogeneous engines and extreme memory efficiency for GRPO, PPO, etc.
Python
-
ai-infra
ai-infra Public从后端工程师到 AI Infra 的学习之路 — MiniMind 四阶段训练(Pretrain/SFT/DPO/LoRA) + 推理引擎 + 算子优化 + RLHF 动手实现
-
minimind
minimind PublicForked from jingyaogong/minimind
🧠「大模型」2小时完全从0训练64M的小参数LLM!Train a 64M-parameter LLM from scratch in just 2h!
Python
-
charging-customer-service
charging-customer-service Public充电桩智能客服 - Taro 3.6 + React 18 + TypeScript 微信小程序,AFK Ralph 自动化开发
TypeScript
-
vllm-omni
vllm-omni PublicForked from vllm-project/vllm-omni
A framework for efficient model inference with omni-modality models
Python
If the problem persists, check the GitHub status page or contact support.
