CV
CHEN Yanxin
Ph.D. Student in Artificial Intelligence
Summary
Ph.D. student in Artificial Intelligence at Fudan University, focusing on long video understanding, video large models, retrieval-augmented generation, and large-scale VLM training infrastructure.
Education
- Ph.D. in Artificial IntelligencePresentFudan University
Work Experience
- Core Contributor2026 -MOSS-VLLarge-scale VLM pretraining and SFT workflow.
- Responsible for four-stage pretraining and the full SFT workflow, including ablation experiments and formal training runs.
- Experienced with thousand-GPU cluster training and debugging training-framework, data, NCCL communication, and deadlock issues.
Skills
AI Algorithms
- Large Language Models
- Video Large Models
- Long Video Understanding
- RAG
- SFT/RL Post-training
- Linear Attention
Training Infrastructure
- Megatron-LM
- TransformerEngine
- DeepSpeed
- FlashAttention
- FlashAttention-3
- TP/PP/DP Parallelism
- Activation Recomputation
Engineering
- Python
- C++
- PyTorch
- Linux/Bash
- NCCL Debugging
- Cluster Debugging
Publications
- ProEchoMem: Enhancing Long Video Understanding via Multi-Trace Probe-Echo Memory2026SIGIR 2026Co-first author. Proposed the PN-RAG framework for probe generation, structured annotation, and fine-grained reranking in long-video understanding.
Interests
- ResearchEfficient VLM Training, Long-context Video Modeling, VLM Post-training, Low-precision Training