← Developers
Open on GitHub 







#1305 1
Jiarui Fang(方佳瑞)
@feifeibear · China
16.2k
Weighted score
2.15k
Contributions
14
Repos
Top repos
- Tencent/PatrickStar
- Tencent/TurboTransformers
- hpcaitech/ColossalAI
- xdit-project/xDiT
- feifeibear/long-context-attention
- feifeibear/LLMSpeculativeSampling
- Tencent-Hunyuan/HunyuanVideo
- hahnyuan/LLM-Viewer
- Oldpan/Pytorch-Memory-Utils
- AmadeusChan/Awesome-LLM-System-Papers
Ranked AI repos8
604212753
hpcaitech/ColossalAI
+0
Making large AI models cheaper, faster and more accessible
41.4k· Python· Model Development
4258253
xdit-project/xDiT
+1
xDiT: A Scalable Inference Engine for Diffusion Transformers (DiTs) with Massive Parallelism
2.74k· Python· Infrastructure
936796
Tencent/TurboTransformers
+0
a fast and user-friendly runtime for transformer inference (Bert, Albert, GPT2, Decoders, etc) on CPU and GPU.
1.55k· C++· Infrastructure
11498381
Oldpan/Pytorch-Memory-Utils
+0
pytorch memory track code
1.01k· Python· Model Development
12053426
feifeibear/LLMSpeculativeSampling
+0
Fast inference from large lauguage models via speculative decoding
926· Python· Infrastructure
13502524
Tencent/PatrickStar
+0
PatrickStar enables Larger, Faster, Greener Pretrained Models for NLP and democratizes AI for everyone.
772· Python· Model Development
14445572
feifeibear/long-context-attention
+0
USP: Unified (a.k.a. Hybrid, 2D) Sequence Parallel Attention for Long Context Transformers Model Training and Inference
695· Python· Infrastructure
145708449
hahnyuan/LLM-Viewer
+0
Analyze the inference of Large Language Models (LLMs). Analyze aspects like computation, storage, transmission, and hardware roofline model in a user-friendly interface.
686· Python· Model Development