Events from January 20, 2017 – September 21, 2026 › Seminar › VASC Seminar › – Robotics Institute Carnegie Mellon University
2026-09-21T00:00:00-04:00
  • VASC Seminar
    Prof. Simon Lucey
    Director of AIML, Professor Adelaide University
    Adelaide University

    Cutting the Skip: Training Residual-Free Transformers

    Newell-Simon Hall 4305

    Abstract:   Transformers are ubiquitous. They influence nearly every aspect of modern AI. However, the mechanics of their training remain poorly understood. This poses a problem for the field due to the immense amounts of data, computational power, and energy being invested in the training of these networks. I highlight a recent intriguing empirical result from [...]

  • VASC Seminar
    Dhawal Sirikonda
    PhD Candidate
    Rendering and Imaging Science Lab (RISc), Department of Computer Science, Dartmouth College

    Using Sound to Steer Light and Light to Measure Sound

    3305 Newell-Simon Hall

    Abstract:   Light and sound are traditionally treated as distinct physical phenomena, yet their interaction provides a powerful mechanism for manipulating and sensing information across imaging, communication, and measurement. This talk explores computational acousto-optic systems that co-design acoustics, optics, and signal processing to enable programmable control of light and high-speed optical sensing without mechanical motion. [...]

    VASC Seminar
    Giljoo Nam
    Research Scientist
    Meta

    My Bitter Lesson with Computer Graphics

    3305 Newell-Simon Hall

    Abstract:  In his essay "The Bitter Lesson," Richard Sutton argued that general methods leveraging computation ultimately outperform hand-crafted ones. In this talk, I share my own version of this lesson, learned the hard way over a decade in computer graphics. Where does the bitter lesson apply to graphics, and where does it not? I argue [...]

  • VASC Seminar
    Postdoctoral Fellow
    Robotics Institute,
    Carnegie Mellon University

    Decision-Making in a World of Latent Particles

    3305 Newell-Simon Hall

    Abstract: Robots must often make decisions in scenes containing many objects: they need to identify what is present, understand where objects are, predict how they will interact, and choose actions accordingly. Learning these capabilities directly from pixels is challenging, especially when the number and arrangement of objects can change from one scene to another. In [...]

    VASC Seminar
    Sanjeev J. Koppal
    Associate Professor
    Electrical & Computer Engineering Department, University of Florida

    Adaptive Cameras: Bridging Novel Sensors and Robot Perception

    Newell-Simon Hall 4305

    Abstract:  Most cameras on robots today capture images without considering scene content. In contrast, animal eyes have fast mechanical movements that control how the scene is imaged in detail by the fovea, where visual acuity is highest. The prevalence of active vision during biological imaging, and the wide variety of it, makes it very clear [...]

    VASC Seminar
    Hanbyul Joo
    Associate Professor
    Computer Science and Engineering, Seoul National University

    From Capturing People to Teaching Robots

    Newell-Simon Hall 3305

    Abstract:   Equipping AI and robotic systems with the ability to understand human behavior is essential for enabling them to assist people across a wide range of everyday applications. This need is more pressing than ever: the heaviest consumers of such knowledge are no longer perception systems alone, but robot policies that must learn to act in [...]

  • VASC Seminar
    Gedas Bertasius
    Assistant Professor
    Department of Computer Science, University of North Carolina

    Video Intelligence for Understanding Behavior and Guiding Action

    3305 Newell-Simon Hall

    Abstract:   ​Modern video-language models have become remarkably good at describing what is happening in a video. Yet their ability to understand the behavior they observe, and to act on that understanding, remains limited. This talk presents our group's work on building these missing capabilities. I will begin with foundational methods for scalable video perception, from [...]