arxiv:2608.13560
YaxinLuo PRO
YaxinLuo
AI & ML interests
AudioVisual Speaker extraction, video understanding, self-supervised large speech models
Recent Activity
upvoted a paper 3 days ago
Agentic Visual Generation: From Generative Models to Agentic Control upvoted a paper 13 days ago
LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes updated a dataset 25 days ago
figmirror/figmirror-transfer-data