Xuan Luo (罗炫)
xuan_luo AT ucsb DOT edu / GitHub / Google Scholar / Hugging Face / CV
I’m a PhD student advised by Prof. Xifeng Yan at University of California Santa Barbara. I develop adaptive and efficient language models through:
- Adaptive computation in LLMs (FlexiDepth, DiffSkip, TAR).
- Efficient Attention Mechanisms (KVPath, AHA, DAR).
- Verification-free multi-token prediction (DMTD).
Manuscripts
2026
KVPath: Connecting Key-Value Representations Across Transformer Layers
Connects KV representations across layers, reducing KV cache by 37.5% relative to MLA while improving performance.
Learning When Not to Attend Globally
Token-Adaptive Representation for Attention
Selected publications
2026
COLM 2026
Enables multi-token prediction without extra parameters, extra computation, or speculative decoding.
Dual Dimensionality for Local and Global Attention
NeurIPS 2026
Reduces KV cache by exploiting distinct roles: local KV supports both memory and prediction, while distant KV primarily supports memory.
2025
COLM 2025
Oral Presentation (top 24/418)
Enables dynamic layer skipping in pre-trained LLMs with near-baseline performance, revealing how computational demands vary across token types.
DiffSkip: Differential Layer Skipping in Large Language Models
ACL Findings 2025
2024
Bot or human? detecting chatgpt imposters with a single question
COLM 2024
2022
Progressive attentional manifold alignment for arbitrary style transfer
ACCV 2022
Computer Science Diagram Understanding with Topology Parsing
TKDD 2022
Experience
Research Scientist Intern, TikTok
Jun 2026 - Sep 2026
Lead Developer / AI Researcher, BioPACIFIC MIP (NSF Materials Innovation Platform)
Apr 2025 - Jun 2026
Professional Services
Reviewer: COLM 2026, COLM 2025, NeurIPS 2024, ACCV 2022.
