Unify-Agent 🐧 Unify-Agent: An end-to-end unified multimodal agent for faithful, knowledge-grounded image generation. csfufu/FactIP Updated Apr 9 • 36 csfufu/FactIP-Full Updated Apr 9 • 22 csfufu/Unify-Agent-Toy-Data Viewer • Updated Apr 9 • 1k • 35 csfufu/Unify-Agent Any-to-Any • 15B • Updated Apr 10 • 6 • 1
Revisual-R1 🚀ReVisual-R1 is a 7B open-source multimodal language model that follows a three-stage curriculum—cold-start pre-training, multimodal reinforcement. csfufu/Revisual-R1-final Image-Text-to-Text • 8B • Updated Jul 14, 2025 • 30 • 8 csfufu/textrl Viewer • Updated Jul 14, 2025 • 32.5k • 23 • 1 csfufu/mmrl Viewer • Updated Jun 25, 2025 • 30.9k • 34 • 1
Unify-Agent 🐧 Unify-Agent: An end-to-end unified multimodal agent for faithful, knowledge-grounded image generation. csfufu/FactIP Updated Apr 9 • 36 csfufu/FactIP-Full Updated Apr 9 • 22 csfufu/Unify-Agent-Toy-Data Viewer • Updated Apr 9 • 1k • 35 csfufu/Unify-Agent Any-to-Any • 15B • Updated Apr 10 • 6 • 1
Revisual-R1 🚀ReVisual-R1 is a 7B open-source multimodal language model that follows a three-stage curriculum—cold-start pre-training, multimodal reinforcement. csfufu/Revisual-R1-final Image-Text-to-Text • 8B • Updated Jul 14, 2025 • 30 • 8 csfufu/textrl Viewer • Updated Jul 14, 2025 • 32.5k • 23 • 1 csfufu/mmrl Viewer • Updated Jun 25, 2025 • 30.9k • 34 • 1