← Developers
Open on GitHub 






#25724 148
十字鱼
@gluttony-10 · China
520
Weighted score
70
Contributions
17
Repos
Top repos
- zai-org/CogView4
- Francis-Rings/StableAnimator
- jiayev/GPT4V-Image-Captioner
- Francis-Rings/StableAvatar
- antgroup/echomimic_v3
- ByteDance-Seed/Bagel
- QwenLM/Qwen3-VL
- antgroup/echomimic_v2
- index-tts/index-tts
- IDEA-Research/Rex-Omni
Ranked AI repos7
1757910306
ByteDance-Seed/Bagel
+-1
Open-source unified multimodal model
6.19k· Python· Models
91755696
IDEA-Research/Rex-Omni
+0
[CVPR2026] Detect Anything via Next Point Prediction
1.60k· Jupyter Notebook· Models
9621544
Francis-Rings/StableAnimator
+0
[CVPR2025] We present StableAnimator, the first end-to-end ID-preserving video diffusion framework, which synthesizes high-quality videos without any post-processing, conditioned on a reference image and a sequence of poses.
1.43k· Python· Models
101904663
Francis-Rings/StableAvatar
+0
[NeurIPS2026] We present StableAvatar, the first end-to-end video diffusion transformer, which synthesizes infinite-length high-quality audio-driven avatar videos without any post-processing, conditioned on a reference image and audio.
1.26k· Python· Models
10854550
zai-org/CogView4
+0
CogView4, CogView3-Plus and CogView3(ECCV 2024)
1.10k· Python· Models
5196522
antgroup/echomimic_v3
+1
[AAAI 2026] EchoMimicV3: 1.3B Parameters are All You Need for Unified Multi-Modal and Multi-Task Human Animation
1.08k· Python· Models
15908498
erwold/qwen2vl-flux
+0
572· Python· Applications