Fadime Sener

I am a Senior Research Scientist at Meta. Previously, I was at the National University of Singapore, and the University of Bonn, where I received my PhD in Computer Science with Angela Yao and Juergen Gall. I hold MSc and BSc degrees in Computer Engineering from Bilkent and Hacettepe University.

My research focuses on multimodal learning and model efficiency. During my PhD, I worked on unsupervised long video understanding through vision and language. At Meta, I worked on online action recognition and streaming video understanding with multimodal LLMs, and on multimedia generation and model distillation.

Email  /  Google Scholar  /  LinkedIn

profile photo

News

Publications

Decouple and Cache
Decouple and Cache: KV Cache Construction for Streaming Video Understanding
Zhanzhong Pang, Dibyadip Chatterjee, Fadime Sener, Angela Yao
ICML, 2026

Paper · Code

On Discriminative vs. Generative Classifiers
On Discriminative vs. Generative Classifiers: Rethinking MLLMs for Action Understanding
Zhanzhong Pang, Dibyadip Chatterjee, Fadime Sener, Angela Yao
ICLR, 2026

Paper · Code

Don't Pause! Every prediction matters in a streaming video
Don't Pause! Every prediction matters in a streaming video
Dibyadip Chatterjee, Zhanzhong Pang, Fadime Sener, Yale Song, Angela Yao
Under review

Paper · Project

OSMO
OSMO: Open-vocabulary Self-eMOtion Tracking
Mohamed Abdelfattah, Bugra Tekin, Fadime Sener, Necati Cihan Camgoz, Eric Sauser, Shugao Ma, Alexandre Alahi, Edoardo Remelli
CVPR, 2026
TrustCLIP
TrustCLIP: Learning Private Visual Features via Adversarial Reconstruction
Nikos Athanasiou, Ilya A. Petrov, Angela Yao, Shugao Ma, Eric Sauser, Edoardo Remelli, Shreyas Hampali, Johannes Schönberger, Fadime Sener, and Bugra Tekin
ECCV, 2026
SneakPeek
SneakPeek: Future-Guided Instructional Streaming Video Generation
Cheeun Hong, German Barquero, Fadime Sener, Markos Georgopoulos, Edgar Schönfeld, Stefan Popov, Yuming Du, Oscar Mañas, Albert Pumarola
Under review

Paper

PALM
PALM: A Dataset and Baseline for Learning Multi-subject Hand Prior
Zicong Fan, Edoardo Remelli, David Dimond, Fadime Sener, Liuhao Ge, Bugra Tekin, Cem Keskin, Shreyas Hampali
3DV, 2026

Paper · Code

2025
Memory-efficient Streaming VideoLLMs for Real-time Procedural Video Understanding
Memory-efficient Streaming VideoLLMs for Real-time Procedural Video Understanding
Dibyadip Chatterjee, Edoardo Remelli, Yale Song, Bugra Tekin, Abhay Mittal, Bharat Lal Bhatnagar, Necati Cihan Camgoz, Shreyas Hampali, Eric Sauser, Shugao Ma, Angela Yao, Fadime Sener
ICCV, 2025
2nd, EgoExo4D Fine-grained Keystep Recognition Challenge, CVPR 2025

Paper · Project · intern project

Context-Enhanced Memory-Refined Transformer for Online Action Detection
Context-Enhanced Memory-Refined Transformer for Online Action Detection
Zhanzhong Pang, Fadime Sener, Angela Yao
CVPR, 2025

Paper · Code

Spatial and temporal beliefs for mistake detection in assembly tasks
Spatial and temporal beliefs for mistake detection in assembly tasks
Guodong Ding, Fadime Sener, Shugao Ma, Angela Yao
CVIU, 2025

Paper

2024
Long-Tail Temporal Action Segmentation with Group-wise Temporal Logit Adjustment
Long-Tail Temporal Action Segmentation with Group-wise Temporal Logit Adjustment
Zhanzhong Pang, Fadime Sener, Shrinivas Ramasubramanian, Angela Yao
ECCV, 2024

Paper · Code

Cost-Sensitive Learning for Long-Tailed Temporal Action Segmentation
Cost-Sensitive Learning for Long-Tailed Temporal Action Segmentation
Zhanzhong Pang, Fadime Sener, Shrinivas Ramasubramanian, Angela Yao
BMVC, 2024

Paper · Code

On the Utility of 3D Hand Poses for Action Recognition
On the Utility of 3D Hand Poses for Action Recognition
Md Salman Shamil, Dibyadip Chatterjee, Fadime Sener, Shugao Ma, Angela Yao
ECCV, 2024

Paper · Code · Project

DiffH2O
DiffH2O: Diffusion-Based Synthesis of Hand-Object Interactions from Textual Descriptions
Sammy Christen, Shreyas Hampali, Fadime Sener, Edoardo Remelli, Tomas Hodan, Eric Sauser, Shugao Ma, Bugra Tekin
SIGGRAPH Asia, 2024

Paper · Project

X-MIC
X-MIC: Cross-Modal Instance Conditioning for Egocentric Action Generalization
Anna Kukleva, Fadime Sener, Edoardo Remelli, Bugra Tekin, Eric Sauser, Bernt Schiele, Shugao Ma
CVPR, 2024

Paper · Code · intern project

2023
Opening the Vocabulary of Egocentric Actions
Opening the Vocabulary of Egocentric Actions
Dibyadip Chatterjee, Fadime Sener, Shugao Ma, Angela Yao
NeurIPS, 2023

Paper · Code · Project

Temporal Action Segmentation
Temporal Action Segmentation: An Analysis of Modern Techniques
Guodong Ding, Fadime Sener, Angela Yao
TPAMI, 2023

Paper · Survey

AssemblyHands
AssemblyHands: Towards Egocentric Activity Understanding via 3D Hand Pose Estimation
Takehiko Ohkawa, Kun He, Fadime Sener, Tomas Hodan, Luan Tran, Cem Keskin
CVPR, 2023
EgoVis 2023/2024 Distinguished Paper Winner

Paper · Code · Project

2022
Transferring Knowledge from Text to Video
Transferring Knowledge from Text to Video: Zero-Shot Anticipation for Procedural Actions
Fadime Sener, Rishabh Saraf, Angela Yao
TPAMI, 2022

Paper

Transformed ROIs for Capturing Visual Transformations in Videos
Transformed ROIs for Capturing Visual Transformations in Videos
Abhinav Rai, Fadime Sener, Angela Yao
CVIU, 2022

Paper

Assembly101
Assembly101: A Large-Scale Multi-View Video Dataset for Understanding Procedural Activities
Fadime Sener, Dibyadip Chatterjee, Daniel Shelepov, Kun He, Dipika Singhania, Robert Wang, Angela Yao
CVPR, 2022
EgoVis 2022/2023 Distinguished Paper Winner

Paper · Code · Project · Video

2021
Temporal aggregate representations technical report
Technical report: Temporal aggregate representations
Fadime Sener, Dibyadip Chatterjee, Angela Yao
Arxiv, EPIC-KITCHENS-Challenges, 2021

Paper · Code

2020
Temporal aggregate representations for long-range video understanding
Temporal aggregate representations for long-range video understanding
Fadime Sener, Dipika Singhania, Angela Yao
ECCV, 2020
1st in action anticipation, 2nd in action recognition, EPIC-KITCHENS 2020 Challenge

Paper · Code

2019
Unsupervised Learning of Action Classes with Continuous Temporal Embedding
Unsupervised Learning of Action Classes with Continuous Temporal Embedding
Anna Kukleva, Hilde Kuehne, Fadime Sener, Jürgen Gall
CVPR, 2019

Paper · Code

Learning Style Compatibility for Furniture
Learning Style Compatibility for Furniture
Divyansh Aggarwal, Elchin Valiyev, Fadime Sener, Angela Yao
GCPR, 2019

Paper

Zero-Shot Anticipation for Instructional Activities
Zero-Shot Anticipation for Instructional Activities
Fadime Sener, Angela Yao
ICCV, 2019

Paper

2018
Unsupervised Learning and Segmentation of Complex Activities from Video
Unsupervised Learning and Segmentation of Complex Activities from Video
Fadime Sener, Angela Yao
CVPR, 2018
Spotlight

Paper

2017
DRAW: Deep networks for Recognizing styles of Artists
DRAW: Deep networks for Recognizing styles of Artists Who illustrate children's books
Samet Hicsonmez, Nermin Samet, Fadime Sener, Pinar Duygulu
ACM ICMR, 2017

Paper

Before 2017
Two-person interaction recognition via spatial multiple instance embedding
Two-person interaction recognition via spatial multiple instance embedding
Fadime Sener, Nazli Ikizler-Cinbis
JVCIR, 2015

Paper

Ensemble of multiple instance classifiers for image re-ranking
Ensemble of multiple instance classifiers for image re-ranking
Fadime Sener, Nazli Ikizler-Cinbis
IMAVIS, 2014

Paper

Identification of illustrators
Identification of illustrators
Fadime Sener, Nermin Samet, Pinar Duygulu
ECCV'W, 2012

Paper

On recognizing actions in still images via multiple features
On recognizing actions in still images via multiple features
Fadime Sener, Nazli Ikizler-Cinbis
ECCV'W, 2012
Best poster, ENS/INRIA Visual Recognition and Machine Learning Summer School, 2013

Paper