About Me 🤔

01 / Background

Hi, I'm Jinxi He (何锦熙). I come from Changsha, Hunan in China. Previously, I was an undergraduate at University of Rochester, supervised by Prof. Chenliang Xu and Prof. Dongmei Li on various research projects. I got my B.S. in Computer Science and B.A. in Psychology degrees here. I've had a wonderful time collaborating with PhD students in their lab, learning and growing alongside them throughout my academic journey. Meliora!

02 / Research

My research interests include Adversarial Machine Learning, Trustworthy AI, and Vision-Language Models (VLMs). I'm passionate about developing AI systems that are both innovative and reliable. I've also contributed to several works on Chain of Thought (CoT) reasoning and multimodal systems.

My current interests are expanding into two key areas: first, MLLM hallucination analysis under different situations and MLLMs safety; and second, robot learning that incorporates vision-language models (VLMs) for planning and manipulation tasks.

I remain committed to continuous learning and growth in this rapidly evolving field.

03 / Service

Conference reviewer: CVPR, ECCV, AAAI, NeurIPS, CoRL, and COLM.

Previously, I served as Vice President of the Chinese Students and Scholars Association (CSSA) at the University of Rochester.

04 / Hobbies

Outside of academics and research, I enjoy cooking traditional Chinese dishes from my hometown and experimenting with fusion cuisine. I'm also an avid badminton player and try to play at least twice a week to stay active. These activities help me maintain a healthy work-life balance and connect with friends from different backgrounds.

Archived News

New paper EchoSafe: Evolving Contextual Safety in Multi-Modal Large Language Models via Inference-Time Self-Reflective Memory is accepted by CVPR 2026. Website Paper Code Data

I² has been accepted to CVPRW 2026. Paper

CAT-V has been awarded Best Demonstration Runner-up @ AAAI 2026.

CAT-V has been accepted to the AAAI 2026 Demonstration Program.

Why Reasoning Matters? A Survey of Advancements in Multimodal Reasoning. Paper Code

Caption Anything in Video: Fine-grained Object-centric Captioning via Spatiotemporal Multimodal Prompting. Paper Code

VERIFY: A Benchmark of Visual Explanation and Reasoning for Investigating Multimodal Reasoning Fidelity. Paper Website

Admitted to the Carnegie Mellon University Master of Science in Robotics (MSR) program!

Thrilled to be interviewed by the University of Rochester Medical Center Clinical & Translational Science Institute. UR CTSI Stories Blog

Honored to be named a Schwartz Discovery Grant winner, with sincere appreciation to Dr. Dongmei Li for her invaluable mentorship and support. Dongmei Li’s Lab News