Yuxiang Lin

MS ECE, Georgia Institute of Technology.

prof_pic.jpg

lin.yuxiang.contact@gmail.com

I received my M.S. in ECE from the Georgia Institute of Technology. My research broadly explores AI agents and autonomous systems, with a particular interest in building agents that can operate, learn, and collaborate over extended periods of time. My previous research spans multimodal large language models and multimodal understanding, including work on emotion recognition and reasoning.

My current research interests include:

  1. General-Purpose & Long-Horizon AI Agents
  2. Persistent Memory, State, and Continual Adaptation
  3. Agent–Harness Co-evolution & Agentic Systems
  4. Multi-Agent Collaboration & Human–Agent Interaction

I am also an open-source builder. I created Understand Anything GitHub stars, an open-source system for helping both humans and AI agents understand complex codebases through static analysis, knowledge graphs, and LLM-based reasoning.

I am currently looking for PhD opportunities in AI agents, autonomous systems, and related areas. If our research interests overlap, I would be very happy to connect and discuss potential research opportunities.

news

Aug 03, 2026 As project lead for AffectLens, I guided our team to a second-place finish in the AffectArt competition at the ACM Multimedia 2026 Grand Challenge using a multi-agent strategy.
Jun 27, 2026 Invited by Apache Flink and Alibaba Cloud to give a talk at Flink Forward Asia 2026 in Shenzhen — check out the LinkedIn post.
Jun 09, 2026 Three months after launch, Understand Anything hit 55,000 GitHub stars on June 9, growing from 0 to 55k at a remarkable pace.
Mar 20, 2026 Understand Anything reached 1,000 GitHub stars in just 5 days after launch. Deeply grateful to the open-source community for the incredible support!
Jul 01, 2025 Try MER-Factory for automatic construction of multimodal emotion recognition and reasoning datasets.
Jun 01, 2025 One paper about Benchmarking VLLM’s Emotion Interpretation ability is accepted by CVPR Workshop, NeXD (1/3 Oral).
May 01, 2025 Started my internship at Tencent.
Dec 01, 2024 One paper about Multimodal Large Language Model in Emotion Reasoning is accepted by NeurIPS (CCF-A).
Jul 01, 2024 One co-first author paper about invisible gas detection is accepted by CVIU (JCR Q1, CCF-B).
Mar 01, 2024 One paper about Conversational Emotion-Cause Pair Analysis with LLM is accepted by SemEval 2024, NAACL.
Jan 15, 2024 Started my internship at Baidu Inc.
Jan 01, 2024 Awarded the First Prize of Research and Innovation Award (3000 CNY) and Star of Craftsmanship (3000 CNY).
Aug 01, 2023 My instance segmentation tutorial has been featured in MMYOLO v0.6.0 highlight! Check out the tutorial to master the essentials of instance segmentation.
Jul 15, 2023 One paper on multimodal emotion recognition is accepted by ACM MM 2023!
Jul 01, 2023 We are the runner-up in the Grand Challenge (MER 2023) of ACM MM!

selected publications

  1. ArXiv
    emcobench.png
    Why We Feel: Breaking Boundaries in Emotional Reasoning with Multimodal Large Language Models
    Yuxiang Lin, Jingdong Sun, Zhi-Qi Cheng, and 7 more authors
    arXiv preprint arXiv:2504.07521, 2025
  2. NeurIPS
    emotionllama_framework.png
    Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning
    Zebang Cheng, Zhi-Qi Cheng, Jun-Yan He, and 5 more authors
    In Advances in Neural Information Processing Systems (NeurIPS), 2024
  3. CVIU
    gas.png
    Invisible Gas Detection: An RGB-Thermal Cross Attention Network and A New Benchmark
    Jue Wang, Yuxiang Lin, Qi Zhao, and 4 more authors
    Computer Vision and Image Understanding, 2024