Deqing Fu

profile.jpg

This is Deqing Fu and I’m a final-year PhD candidate in Computer Science at the University of Southern California (USC). My main research interests are theoretical foundations of large language models and multimodal LLMs. I’m (co-)advised by Prof. Vatsal Sharan of USC Theory Group and Prof. Robin Jia of Allegro Lab within USC NLP Group, and I’m working closely with Prof. Mahdi Soltanolkotabi. During my Ph.D. studies, I spent time at Google and Meta as a student researcher. Before USC, I completed my undergraduate degree in Mathematics (with honors) and my master’s in Statistics at the University of Chicago.

Availability

I am on the job market in 2026–2027. Please reach out!

Research Highlights

Algorithmic Perspectives on Large Language Models
Interpretability and Alignment
Multimodal Models and Applications
  • Multimodal rewards for improving generation quality: token-level hallucination reduction (ICLR 2025) and text-to-image alignment (NAACL 2025)
  • Modality sensitivity in Multimodal LLMs (COLM 2024)
  • Large-scale dataset for visual reasoning with images (ICLR 2026)
Aug 29, 2026 Our paper Are LLM Decisions Faithful to Verbal Confidence? was accepted to EMNLP 2026 Main Conference!
Jul 21, 2026 New paper TOPL: Token-Level Off-Policy Learning for Faithful Generation Under Distribution Shift, accepted to COLM 2026!
Jul 08, 2026 Convergent Evolution and Resa are accepted to COLM 2026!
Jun 30, 2026 I contributed to TabFM, a zero-shot foundation model for tabular data, released by Google Research!
Jun 29, 2026 New preprint: Value-Aware Stochastic KV Cache Eviction for Reasoning Models.

Selected Publications

All publications →

2026

  1. ICML
    graph.png
    Transformers Provably Learn Algorithmic Solutions for Graph Connectivity, But Only with the Right Data
    ICML
    Qilin Ye*Deqing Fu*Robin Jia, and Vatsal Sharan
    In International Conference on Machine Learning (ICML), 2026
  2. arXiv
    vase.png
    Value-Aware Stochastic KV Cache Eviction for Reasoning Models
    arXiv
    In arXiv, 2026
  3. COLM
    convergent.png
    Convergent Evolution: How Different Language Models Learn Similar Number Representations
    COLM
    In Conference on Language Modeling (COLM), 2026
  4. ICLR
    zebra-cot.png
    Zebra-CoT: A Dataset for Interleaved Vision Language Reasoning
    ICLR
    In International Conference on Learning Representations (ICLR), 2026
  5. ICLR
    fone.png
    FoNE: Precise Single-Token Number Embeddings via Fourier Features
    ICLR
    In International Conference on Learning Representations (ICLR), 2026
  6. ACL
    saegull.png
    Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models
    ACL
    In Association of Computational Linguistics (ACL), 2026

2025

  1. ICLR
    tldr.png
    TLDR: Token-Level Detective Reward Model for Large Vision Language Models
    ICLR
    Deqing Fu, Tong Xiao , Rui Wang, Wang ZhuPengchuan Zhang, Guan Pang, Robin Jia, and Lawrence Chen
    In International Conference on Learning Representations (ICLR), 2025
  2. ICLR
    sensitivity.png
    Transformers Learn Low Sensitivity Functions: Investigations and Implications
    ICLR
    Bhavya Vasudeva*Deqing Fu*Tianyi Zhou, Elliot Kau , You-Qi Huang, and Vatsal Sharan
    In International Conference on Learning Representations (ICLR), 2025
  3. ICLR
    dellma.png
    DeLLMa: Decision Making Under Uncertainty with Large Language Models
    ICLR
    In International Conference on Learning Representations (ICLR), 2025
    Spotlight (Top 5.1%)
  4. NAACL
    dreamsync.png
    DreamSync: Aligning Text-to-Image Generation with Image Understanding Feedback
    NAACL
    Jiao Sun*Deqing Fu*Yushi Hu* , Su Wang, Royi Rassin, Da-Cheng Juan, Dana Alon, Charles Herrmann, Sjoerd Steenkiste, Ranjay Krishna, and Cyrus Rashtchian
    In Annual Conference of the North American Chapter of the Association for Computational Linguistics (NAACL), 2025

2024

  1. NeurIPS
    transformer-icl.png
    Transformers Learn to Achieve Second-Order Convergence Rates for In-Context Linear Regression
    NeurIPS
    Deqing Fu, Tian-Qi Chen, Robin Jia, and Vatsal Sharan
    In Conference on Neural Information Processing Systems (NeurIPS), 2024
    SoCalNLP Symposium 2023 Best Paper Award
  2. NeurIPS
    fourier.png
    Pre-trained Large Language Models Use Fourier Features to Compute Addition
    NeurIPS
    Tianyi ZhouDeqing FuVatsal Sharan, and Robin Jia
    In Conference on Neural Information Processing Systems (NeurIPS), 2024
  3. COLM
    isobench.png
    IsoBench: Benchmarking Multimodal Foundation Models on Isomorphic Representations
    COLM
    Deqing Fu*Ruohao Guo*, Ghazal Khalighinejad*Ollie Liu*Bhuwan DhingraDani YogatamaRobin Jia, and Willie Neiswanger
    In Conference on Language Modeling (COLM), 2024

* Equal contribution.