Club&Lab for PaddlePaddle contributors
Pinned Loading
Repositories
- mcore-bridge Public Forked from modelscope/mcore-bridge
MCore-Bridge: Providing Megatron-Core model definitions for state-of-the-art large models and making Megatron training as simple as Transformers — with support for 300+ large language models (Qwen3-Next, GLM-5.2, Deepseek-V4, MiniMax-2.7, ...) and 200+ multimodal large models (Qwen3.5, Qwen3-Omni, Gemma4, ...).
- ms-swift Public Forked from modelscope/ms-swift
Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, Phi4, ...) (AAAI 2025).
- Megatron-LM Public Forked from NVIDIA/Megatron-LM
Ongoing research training transformer models at scale
-
- ast-grep-pre-commit-mirror Public
ast-grep pre-commit hook that uses the PyPI distribution for seamless integration with Python projects.
- dochooks Public
- PaddleAPITest Public
- flash-linear-attention Public Forked from fla-org/flash-linear-attention
🚀 Efficient implementations of state-of-the-art linear attention models
Top languages
Loading…
Most used topics
Loading…