Skip to main content

Showing 1–2 of 2 results for author: Rädle, R

Searching in archive cs. Search in all archives.
.
  1. arXiv:2408.00714  [pdf, other

    cs.CV cs.AI cs.LG

    SAM 2: Segment Anything in Images and Videos

    Authors: Nikhila Ravi, Valentin Gabeur, Yuan-Ting Hu, Ronghang Hu, Chaitanya Ryali, Tengyu Ma, Haitham Khedr, Roman Rädle, Chloe Rolland, Laura Gustafson, Eric Mintun, Junting Pan, Kalyan Vasudev Alwala, Nicolas Carion, Chao-Yuan Wu, Ross Girshick, Piotr Dollár, Christoph Feichtenhofer

    Abstract: We present Segment Anything Model 2 (SAM 2), a foundation model towards solving promptable visual segmentation in images and videos. We build a data engine, which improves model and data via user interaction, to collect the largest video segmentation dataset to date. Our model is a simple transformer architecture with streaming memory for real-time video processing. SAM 2 trained on our data provi… ▽ More

    Submitted 28 October, 2024; v1 submitted 1 August, 2024; originally announced August 2024.

    Comments: Website: https://ai.meta.com/sam2

  2. AdaM: Adapting Multi-User Interfaces for Collaborative Environments in Real-Time

    Authors: Seonwook Park, Christoph Gebhardt, Roman Rädle, Anna Feit, Hana Vrzakova, Niraj Dayama, Hui-Shyong Yeo, Clemens Klokmose, Aaron Quigley, Antti Oulasvirta, Otmar Hilliges

    Abstract: Developing cross-device multi-user interfaces (UIs) is a challenging problem. There are numerous ways in which content and interactivity can be distributed. However, good solutions must consider multiple users, their roles, their preferences and access rights, as well as device capabilities. Manual and rule-based solutions are tedious to create and do not scale to larger problems nor do they adapt… ▽ More

    Submitted 29 March, 2018; v1 submitted 3 March, 2018; originally announced March 2018.

    Comments: formatting tweaks

    ACM Class: H.5.m