About Me
I am currently a Machine Learning Engineer at Bytedance.
I am now working on Multimodal Large Language Model(MLLM) and intrested in Unified Multimodal Understanding and Generation Reaserach. If you are also interested in this filed, seeking any form of academic cooperation, or even just want discuss related ideas, please feel free to email me at zhangruo.roy@bytedance.com. We are also hiring interns in MLLM, Text to Image/Video/Speech!
Prior to that, I graduated from University College London with a master’s degree, advised by Prof YuanChang Liu.
My research interest includes Multimodal Large Language Model(MLLM), Unified Multimodal Understanding and Generation.
Selected Publications
Target-referenced Reactive Grasping for Dynamic Objects. CVPR 2023
Jirong Liu, Ruo Zhang, Haoshu Fang, Minghao Gou, Hongjie Fang, Chenxi Wang, Sheng Xu, Hengxu Yan, Cewu Lu
AirExo: Low-Cost Exoskeletons for Learning Whole-Arm Manipulation in the Wild. ICRA 2024
Hongjie Fang, Haoshu Fang, Yiming Wang, Jieji Ren, Jingjing Chen, Ruo Zhang, Weiming Wang, Cewu Lu
