publications

Publications in reversed chronological order.

2026

  1. AIES
    From Forensics to Ecosystems: Rethinking Watermarks for Generative AI Oversight
    Susser, Daniel, Thickstun, John, and Vidan, Gili
    In AI, Ethics, and Society 2026
  2. ISMIR
    Assessing Factual Music Comprehension in Large Audio Language Models
    Lin, Daniel Chenyu, Freeman, Michael, and Thickstun, John
    In International Society for Music Information Retrieval 2026
  3. ICML
    Scaling Beyond Masked Diffusion Language Models
    Sahoo, Subham Sekhar, Lemercier, Jean-Marie, Yang, Zhihan, Deschenaux, Justin, Liu, Jingyu, Thickstun, John, and Jukic, Ante
    In International Conference on Machine Learning 2026
  4. ICML
    Esoteric Language Models: A Family of Any-Order Diffusion LLMs
    Sahoo, Subham Sekhar, Yang, Zhihan, Akhauri, Yash, Liu, Johnna, Singh, Deepansha, Cheng, Zhoujun, Liu, Zhengzhong, Xing, Eric P., Thickstun, John, and Vahdat, Arash
    In International Conference on Machine Learning 2026
  5. ICLR-W
    Esoteric Language Models: Bridging Autoregressive and Masked Diffusion LLMs
    Sahoo, Subham Sekhar, Yang, Zhihan, Akhauri, Yash, Liu, Johnna, Singh, Deepansha, Cheng, Zhoujun, Liu, Zhengzhong, Xing, Eric, Thickstun, John, and Vahdat, Arash
    In ICLR Workshop on Multimodal Intelligence: Next Token Prediction and Beyond 2026
  6. ICLR-W
    The Reliability Gap in Agentic Evidence Verification for Materials Science
    Gong, Albert, Kim, James J., Kabra, Anmol, Panigrahi, Aaditya, Wang, Jiashuo, Mulchandani, Arjun B., Freeman, Michael, Katmer, Fatmagul, Wakefield, Joshua Peters, Zhao, Linxi, Wan, Chao, Sarkar, Akanksha, Artzi, Yoav, Schoop, Leslie M., Thickstun, John, Weinberger, Kilian Q., Kim, Eun-Ah, Frazier, Peter I., and Sun, Jennifer J.
    In ICLR Workshops on Agents in the Wild (AIWILD) and Foundation Models for Science (FM4Science) 2026
  7. KDD
    Benchmark Datasets for Lead-Lag Forecasting on Social Platforms
    Kazemian, Kimia, Liu, Zhenzhen, Yang, Yangfanyu, Luo, Katie Z., Gu, Shuhan, Du, Audrey, Yang, Xinyu, Jansons, Jack, Weinberger, Kilian Q., Thickstun, John, Yin, Yian, and Dean, Sarah
    In Knowledge Discovery in Databases - Datasets and Benchmarks Track 2026

2025

  1. NeurIPS-W
    Robust Neural Audio Fingerprinting using Music Foundation Models
    Singh, Shubhr, Bhat, Kiran, Resnick, Benjamin, Riley, Xavier, Thickstun, John, and Brouwer, Walter De
    In NeurIPS Workshop on AI for Music 2025
  2. Neurips
    Linearly Constrained Diffusion Implicit Models
    Jayaram, Vivek, Kemelmacher-Shlizerman, Ira, Seitz, Steven M, and Thickstun, John
    In Advances in Neural Information Processing Systems 2025
  3. ISMIR
    Aligning Text-to-Music Evaluation with Human Preferences
    Huang, Yichen, Novack, Zachary, Saito, Koichi, Shi, Jiatong, Watanabe, Shinji, Mitsufuji, Yuki, Thickstun, John, and Donahue, Chris
    In International Society for Music Information Retrieval 2025

2024

  1. ISMIR-LBD
    Hookpad Aria: A Copilot for Songwriters
    Donahue, Chris, Wu, Shih-Lun, Kim, Yewon, Carlton, Dave, Miyakawa, Ryan, and Thickstun, John
    In International Society for Music Information Retrieval Late Breaking Demos 2024
  2. GenAICHI
    Designing Live Human-AI Collaboration for Musical Improvisation
    Becker, Nic, Louie, Ryan, Thickstun, John, and Liang, Percy
    In CHI Workshop on Generative AI and HCI 2024
  3. TMLR
    Robust distortion-free watermarks for language models
    Kuditipudi, Rohith, Thickstun, John, Hashimoto, Tatsunori, and Liang, Percy
    Transactions on Machine Learning Research 2024
  4. TMLR
    Anticipatory music transformer
    Thickstun, John, Hall, David, Donahue, Chris, and Liang, Percy
    Transactions on Machine Learning Research 2024

2023

  1. JMLR
    MAUVE Scores for Generative Models: Theory and Practice
    Pillutla, Krishna, Liu, Lang, Thickstun, John, Welleck, Sean, Swayamdipta, Swabha, Zellers, Rowan, Oh, Sewoong, Choi, Yejin, and Harchaoui, Zaid
    Journal of Machine Learning Research 2023
  2. ACL Outstanding Paper
    Backpack language models
    Hewitt, John, Thickstun, John, Manning, Christopher D., and Liang, Percy
    In Proceedings of the Association for Computational Linguistics 2023
  3. TMLR
    Evaluating Human-Language Model Interaction
    Lee, Mina, Srivastava, Megha, Hardy, Amelia, Thickstun, John, Durmus, Esin, Paranjape, Ashwin, Gerard-Ursin, Ines, Li, Xiang Lisa, Ladhak, Faisal, Rong, Frieda, Wang, Rose E., Kwon, Minae, Park, Joon Sung, Cao, Hancheng, Lee, Tony, Bommasani, Rishi, Bernstein, Michael, and Liang, Percy
    Transactions on Machine Learning Research 2023

2022

  1. ISMIR
    Melody transcription via generative pre-training
    Donahue, Chris, Thickstun, John, and Liang, Percy
    In International Society for Music Information Retrieval 2022
  2. bioRxiv
    Reconstruction of visual images from mouse retinal ganglion cell spiking activity using convolutional neural networks
    Benster, Tyler, Babino, Darwin, Thickstun, John, Hunt, Matthew, Liu, Xiyang, Harchaoui, Zaid, Oh, Sewoong, and Gelder, Russell N Van
    2022
  3. Neurips Oral Presentation
    Diffusion-LM improves controllable text generation
    Li, Xiang Lisa, Thickstun, John, Gulrajani, Ishaan, Liang, Percy, and Hashimoto, Tatsunori B.
    In Advances in Neural Information Processing Systems 2022

2021

  1. Dissertation
    Leveraging generative models for music and signal processing
    Thickstun, John
    University of Washington 2021
  2. Neurips Outstanding Paper
    MAUVE: measuring the gap between neural text and human text using divergence frontiers
    Pillutla, Krishna, Swayamdipta, Swabha, Zellers, Rowan, Thickstun, John, Welleck, Sean, Choi, Yejin, and Harchaoui, Zaid
    In Advances in Neural Information Processing Systems 2021
  3. ICML
    Parallel and flexible sampling from autoregressive models via Langevin dynamics
    Jayaram, Vivek, and Thickstun, John
    In International Conference on Machine Learning 2021
  4. L4DC
    Faster policy learning with continuous-time gradients
    Ainsworth, Samuel K., Lowrey, Kendall, Thickstun, John, Harchaoui, Zaid, and Srinivasa, Siddhartha S.
    In Learning for Dynamics and Control 2021

2020

  1. arXiv
    Rethinking evaluation methodology for audio-to-score alignment
    Thickstun, John, Brennan, Jennifer, and Verma, Harsh
    arXiv preprint arXiv:2009.14374 2020
  2. EMNLP
    An information bottleneck approach for controlling conciseness in rationale extraction
    Paranjape, Bhargavi, Joshi, Mandar, Thickstun, John, Hajishirzi, Hannaneh, and Zettlemoyer, Luke
    In Conference on Empirical Methods in Natural Language Processing 2020
  3. ICML
    Source separation with deep generative priors
    Jayaram, Vivek, and Thickstun, John
    In International Conference on Machine Learning 2020

2019

  1. ISMIR
    Convolutional composer classification
    Verma, Harsh, and Thickstun, John
    In International Society for Music Information Retrieval 2019
  2. ISMIR
    Coupled recurrent models for polyphonic music composition
    Thickstun, John, Harchaoui, Zaid, Foster, Dean P, and Kakade, Sham M
    In International Society for Music Information Retrieval 2019

2018

  1. ICASSP Oral Presentation
    Invariances and data augmentation for supervised music transcription
    Thickstun, John, Harchaoui, Zaid, Foster, Dean P, and Kakade, Sham M
    In International Conference on Acoustics, Speech and Signal Processing 2018

2017

  1. MIREX
    Frequency domain convolutions for multiple F0 estimation
    Thickstun, John, Harchaoui, Zaid, Foster, Dean P, and Kakade, Sham M
    2017
  2. ICLR
    Learning features of music from scratch
    Thickstun, John, Harchaoui, Zaid, and Kakade, Sham M
    In International Conference on Learning Representations 2017