Publications

Research publications on representation learning, identifiability, and multimodal learning.

Selected publications are highlighted.

  1. Preprint’26
    masked_prediction.png
    On the Identifiability of Masked Prediction: Mode Blindness and Mask Schedules
    Yichao Cai and Javen Q. Shi
    arXiv preprint arXiv:2608.01383, 2026
  2. ICML’26
    InfoNCE_geometry.png
    The Geometric Mechanics of Contrastive Representation Learning: Alignment Potentials, Entropic Dispersion, and Cross-Modal Divergence
    Yichao Cai, Zhen Zhang, Yuhang Liu, and Javen Q. Shi
    In International Conference on Machine Learning (ICML), 2026
  3. NeurIPS’25
    misalignment.png
    On the Value of Cross-Modal Misalignment in Multimodal Representation Learning
    Yichao Cai, Yuhang Liu, Erdun Gao, Tianjiao Jiang, Zhen Zhang, Anton van den Hengel, and Javen Q. Shi
    In Advances in Neural Information Processing Systems (NeurIPS), 2025  Spotlight
  1. ICML’26
    single_cell.png
    What Makes a Representation Good for Single-Cell Perturbation Prediction?
    Wenkang Jiang, Yuhang Liu, Yichao Cai, Erdun Gao, Jiayi Dong, Ehsan Abbasnejad, Lina Yao, and Javen Q. Shi
    In International Conference on Machine Learning (ICML), 2026
  2. ICML’26
    gnn.png
    Boundary Embedding Shaping with Adaptive Contrastive Learning for Graph Structural Disentanglement
    Jiaqing Chen, Zidu Yin, Yichao Cai, Yuhang Liu, Zhen Zhang, Dong Gong, and Javen Q. Shi
    In International Conference on Machine Learning (ICML), 2026
  3. ICLR’26
    ntp_concept.png
    I Predict Therefore I Am: Is Next Token Prediction Enough to Learn Human-Interpretable Concepts from Data?
    Yuhang Liu, Dong Gong, Yichao Cai, Erdun Gao, Zhen Zhang, Biwei Huang, Mingming Gong, Anton van den Hengel, and Javen Q. Shi
    In International Conference on Learning Representations (ICLR), 2026
  4. ECCV’24
    CLAP.png
    CLAP: Isolating Content from Style through Contrastive Learning with Augmented Prompts
    Yichao Cai, Yuhang Liu, Zhen Zhang, and Javen Q. Shi
    In European Conference on Computer Vision (ECCV), 2024