Hao AI Lab @ UCSD
    • Home
    • Blogs
    • Projects
    • Talks
    • People
    • Publications
    • Contact

    Blogs

    Attention-FFN disaggregated token flow

    FastAFD: Open-Source Large-Scale Attention-FFN Disaggregation on Blackwell NVL72

    July 6, 2026

    Yichao Fu*, Yuxuan Zhang*, Ruitian Wang, Junda Chen, Hao Zhang

    JetSpec parallel tree drafting

    JetSpec: Breaking the Scaling Ceiling of Speculative Decoding with Parallel Tree Drafting

    June 22, 2026

    Lanxiang Hu, Zhaoxiang Feng, Yulun Wu, Haoran Yuan, Yujie Zhao, Yu-Yang Qian, Bojun Wang, Peng Zhao, Daxin Jiang, Yibo Zhu, Tajana Rosing, Hao Zhang

    FastWan-QAD generates a 5-second 480p video in 1.8s on a single RTX 5090

    FastWan-QAD: FastVideo generates a 5-Second Video in 1.8 Seconds on a Single NVIDIA GeForce RTX 5090 via Quantization-Aware Distillation

    June 15, 2026

    FastVideo Team

    Open-sourcing FastVideo Dreamverse: Real-Time Vibe Directing with LTX-2 on a single NVIDIA B200 GPU

    Open-sourcing FastVideo Dreamverse: Real-Time Vibe Directing with LTX-2 on a single NVIDIA B200 GPU

    May 26, 2026

    FastVideo Team

    making 4-bit attention actually work

    Attn-QAT: Making 4-Bit Attention Actually Work

    April 8, 2026

    Peiyuan Zhang*, Matthew Noto*, Wenxuan Tan*, Chengquan Jiang, Will Lin, Wei Zhou, Hao Zhang

    Into the Dreamverse: Vibe Directing in FastVideo

    Into the Dreamverse: Vibe Directing in FastVideo

    March 15, 2026

    FastVideo Team

    Create a 5s 1080p Video in 4.5s with FastVideo on a Single GPU

    Create a 5s 1080p Video in 4.5s with FastVideo on a Single GPU

    March 11, 2026

    FastVideo Team

    scientific reasoning in video world models

    From Physical Commonsense to Scientific Reasoning: Why World Modeling in Video Matters

    February 12, 2026

    Lanxiang Hu, Abhilash Shankarampeta, Yixin Huang, Zilin Dai, Haoyang Yu, Yujie Zhao, Haoqiang Kang, Daniel Zhao, Tajana Rosing, Hao Zhang

    DistCA

    CAD: Disaggregating Core Attention for Efficient Long-context Language Model Training

    December 17, 2025

    Yonghao Zhuang*, Junda Chen*, Bo Pang, Yi Gu, Yibo Zhu, Yimin Jiang, Ion Stoica, Eric Xing, Hao Zhang

    jacobi forcing decoding

    Fast and Accurate Causal Parallel Decoding using Jacobi Forcing

    December 16, 2025

    Lanxiang Hu*, Siqi Kou*, Yichao Fu, Samyam Rajbhandari, Tajana Rosing, Yuxiong He, Zhijie Deng, Hao Zhang

    • ««
    • «
    • 1
    • 2
    • 3
    • »
    • »»
    © 2026 Hao AI Lab @ UCSD Powered by Hugo & PaperMod , Adapted by Lanxiang Hu, Junda Chen & Hao Zhang