AI Research

  1. Muse Image and Muse Video

    Meta Superintelligence Labs · 2026 · Lead

    Muse Image generates and edits images agentically — it calls search and coding tools to ground facts and get details like QR codes and plots right, then refines its own output. Muse Video is built on the same foundation, with native audio. Muse Image ships across the Meta AI app, meta.ai, Instagram Stories and WhatsApp; Muse Video is in early preview.

  2. Sora

    OpenAI · 2024 · Co-lead

    A generative model for video. I co-led the project, which trained a diffusion transformer on video and images of varying durations, resolutions and aspect ratios.

  3. InstructPix2Pix

    Berkeley AI Research · 2023 · First author

    Learning to follow image editing instructions. Given a photograph and a written instruction — "make it winter", "turn him into a cyborg" — the model performs the edit directly, in a single forward pass, with no per-image fine-tuning, masks or extra descriptions. With Aleksander Holynski and Alyosha Efros.

    InstructPix2Pix — example output
  4. Long Video Generation

    NVIDIA · 2022 · First author

    Research on generative models for video, before video generation was a crowded field.

    Long Video Generation — example output
  5. Pixel Smartphone Camera

    Google · 2018

    Computational photography for the Pixel phone — the part of the camera that is a model rather than a lens.

    Pixel Smartphone Camera — example output