← Developers
Open on GitHub 







#1281 5
Jiarui Fang(方佳瑞)
@feifeibear · China
16.2k
Weighted score
2.15k
Contributions
14
Repos
Top repos
- Tencent/PatrickStar
- Tencent/TurboTransformers
- hpcaitech/ColossalAI
- xdit-project/xDiT
- feifeibear/long-context-attention
- feifeibear/LLMSpeculativeSampling
- Tencent-Hunyuan/HunyuanVideo
- hahnyuan/LLM-Viewer
- Oldpan/Pytorch-Memory-Utils
- ByteDance-Seed/VeOmni
Ranked AI repos8
18673530
hpcaitech/ColossalAI
+-3
Making large AI models cheaper, faster and more accessible
41.4k· Python· Model Development
37423723
xdit-project/xDiT
+1
xDiT: A Scalable Inference Engine for Diffusion Transformers (DiTs) with Massive Parallelism
2.73k· Python· Infrastructure
8396936
Tencent/TurboTransformers
+0
a fast and user-friendly runtime for transformer inference (Bert, Albert, GPT2, Decoders, etc) on CPU and GPU.
1.55k· C++· Infrastructure
101781356
Oldpan/Pytorch-Memory-Utils
+0
pytorch memory track code
1.01k· Python· Model Development
106671422
feifeibear/LLMSpeculativeSampling
+0
Fast inference from large lauguage models via speculative decoding
927· Python· Infrastructure
120051578
Tencent/PatrickStar
+0
PatrickStar enables Larger, Faster, Greener Pretrained Models for NLP and democratizes AI for everyone.
772· Python· Model Development
128821682
feifeibear/long-context-attention
+0
USP: Unified (a.k.a. Hybrid, 2D) Sequence Parallel Attention for Long Context Transformers Model Training and Inference
695· Python· Infrastructure
1787812590
hahnyuan/LLM-Viewer
+-1
Analyze the inference of Large Language Models (LLMs). Analyze aspects like computation, storage, transmission, and hardware roofline model in a user-friendly interface.
681· Python· Model Development