|
News
-
09 / 2026: 3 papers accepted to NeurIPS 2026 Workshops!
[NEW]
-
09 / 2026: 1 paper accepted to NeurIPS 2026!
[NEW]
-
09 / 2026: Serving as Associate Editor for ICRA 2027!
[NEW]
-
08 / 2026: 2 papers on LLM agents and robust reasoning are now available on arXiv!
-
01 / 2026: 1 paper accepted to ICRA 2026!
|
|
Watch, Infer, Coordinate: Inferring Robot Partner Constraints for Zero-Shot Coordination
Suyu Ye, Zheyuan Zhang, Vaishnav Tadiparthi, Hossein Nourkhiz Mahjoub, Ehsan Moradi Pari, Tianmin Shu, Homanga Bharadhwaj, Nakul Agarwal
Physical Understanding for Decision-Making NeurIPS, 2026
project   paper
|
|
JEPA-TTT: Persistent Test-Time Training of Latent World Models for Planning under Dynamics Shifts
Zheyuan Zhang, Suyu Ye, Nakul Agarwal, Hossein Nourkhiz Mahjoub, Ehsan Moradi Pari, Daniel Khashabi, Tianmin Shu, Vaishnav Tadiparthi
World Models in Physical AI NeurIPS, 2026
project   paper
|
|
From Internal Evidence to Behavioral Commitment in Embodied Skill Selection
Hanyu Wang, Vaishnav Tadiparthi, Nakul Agarwal, Hossein Nourkhiz Mahjoub, Ehsan Moradi Pari
Interpreting Agent Behavior NeurIPS, 2026
project   paper
|
|
Learning Robust Reasoning through Guided Adversarial Self-Play
Shuozhe Li, Vaishnav Tadiparthi, Kwonjoon Lee, Nakul Agarwal, Hossein Nourkhiz Mahjoub, Ehsan Moradi Pari, Lizhang Chen, Amy Zhang, Liu Leqi
NeurIPS, 2026
project   paper
|
|
Generative Skill Composition for LLM Agents
Xinyu Zhao, Zhen Tan, Vaishnav Tadiparthi, Nakul Agarwal, Kwonjoon Lee, Ehsan Moradi Pari, Hossein Nourkhiz Mahjoub, Tianlong Chen
arXiv, 2026
project   paper
|
|
SAGE: Synchronized Action-Gaze Recognition and Anticipation for Human Behavior Understanding
Chenyi Kuang, Nakul Agarwal
arXiv, 2026
project   paper
|
|
MERGE: Guided Vision-Language Models for Multi-Actor Event Reasoning and Grounding in Human-Robot Interaction
Joerg Deigmoeller*, Nakul Agarwal*, Stephan Hasler, Daniel Tanneberg, Anna Belardinelli, Reza Ghoddoosian, Chao Wang, Felix Ocker, Fan Zhang, Behzad Dariush, Michael Gienger
ICRA, 2026
project   paper
|
|
Towards Driver Behavior Understanding: Weakly-Supervised Risk Perception in Driving Scenes
Nakul Agarwal, Yi-Ting Chen, Behzad Dariush
IV, 2026
project   paper
|
|
Overcoming Multi-step Complexity in Multimodal Theory-of-Mind Reasoning: A Scalable Bayesian Planner
Chunhui Zhang, Sean Dae Houlihan, Kwonjoon Lee, Nakul Agarwal, Zhongyu Ouyang, Soroush Vosoughi, Shao-Yuan Lo
ICML, 2025 (Spotlight - Selection rate 2.6%)
project   paper
|
|
Pose-Aware Weakly-Supervised Action Segmentation
Zhihao Zhao, Reza Ghoddoosian, Isht Dwivedi, Nakul Agarwal, Behzad Dariush
Multimodal Learning and Applications CVPR, 2025
project   paper
|
|
ACE: Action Concept Enhancement of Video-Language Models in Procedural Videos
Reza Ghoddoosian, Nakul Agarwal, Isht Dwivedi, Behzad Dariush
WACV, 2025
project   paper
|
|
Vamos: Versatile Action Models for Video Understanding
Shijie Wang, Qi Zhao, Minh Quan Do, Nakul Agarwal, Kwonjoon Lee, Chen Sun
ECCV, 2024
project   paper   code
|
|
M2D2M: Multi-Motion Generation from Text with Discrete Diffusion Models
Seunggeun Chi*, Hyung-gun Chi*, Hengbo Ma, Nakul Agarwal, Faizan Siddiqui, Karthik Ramani, Kwonjoon Lee
ECCV, 2024
project   paper
|
|
Uncertainty-aware Action Decoupling Transformer for Action Anticipation
Hongji Guo, Nakul Agarwal, Shao-Yuan Lo, Kwonjoon Lee, Qiang Ji
CVPR, 2024 (Highlight - Selection rate 2.8%)
project   paper
|
|
Can’t make an Omelette without Breaking some Eggs: Plausible Action Anticipation using Large Video-Language Models
Himangi Mittal, Nakul Agarwal, Shao-Yuan Lo, Kwonjoon Lee
CVPR, 2024
project   paper
|
|
AntGPT: Can Large Language Models Help Long-term Action Anticipation from Videos?
Qi Zhao*, Ce Zhang*, Shijie Wang, Changcheng Fu, Minh Quan Do, Nakul Agarwal, Kwonjoon Lee, Chen Sun
ICLR, 2024
project   paper   code
|
|
Disentangled Neural Relational Inference for Interpretable Motion Prediction
Victoria Magdalena Dax, Jiachen Li, Enna Sachdeva, Nakul Agarwal, Mykel Kochenderfer
RA-L and ICRA, 2024
project   paper
|
|
Rank2Tell: A Multimodal Driving Dataset for Joint Importance Ranking and Reasoning
Enna Sachdeva*, Nakul Agarwal*, Suhas Chundi, Jiachen Li, Mykel Kochenderfer, Chiho Choi, Behzad Dariush,
WACV, 2024
project   paper
|
|
Object-centric Video Representation for Long-term Action Anticipation
Ce Zhang*, Changcheng Fu*, Shijie Wang, Nakul Agarwal, Kwonjoon Lee, Chiho Choi, Chen Sun
WACV, 2024
project   paper
|
|
Ordered Atomic Activity for Fine-Grained Interactive Traffic Scenario Understanding
Nakul Agarwal, Yi-Ting Chen
ICCV, 2023
project   paper
|
|
Weakly-Supervised Action Segmentation and Unseen Error Detection in Anomalous Instructional Videos
Reza Ghoddoosian, Isht Dwivedi, Nakul Agarwal, Behzad Dariush
ICCV, 2023
project   paper
|
|
Latency Matters: Real-Time Action Forecasting Transformer
Harshayu Girase, Nakul Agarwal, Chiho Choi, Karttikeya Mangalam
CVPR, 2023 (Highlight - Selection rate 2.6%)
project   paper
|
|
AdamsFormer for Spatial Action Localization in the Future
Hyung-gun Chi, Kwonjoon Lee, Nakul Agarwal, Yi Xu, Karthik Ramani, Chiho Choi
CVPR, 2023
project   paper
|
|
Risk Perception in Driving Scenes
Nakul Agarwal, Yi-Ting Chen
Machine Learning for Autonomous Driving NeurIPS, 2022
project   paper
|
|
Weakly-Supervised Online Action Segmentation in Multi-View Instructional Videos
Reza Ghoddoosian, Isht Dwivedi, Nakul Agarwal, Chiho Choi, Behzad Dariush
CVPR, 2022
project   paper   code
|
|
Unsupervised Domain Adaptation for Spatio-Temporal Action Localization
Nakul Agarwal, Yi-Ting Chen, Behzad Dariush, Ming-Hsuan Yang
BMVC, 2020
project   paper
Short Version: VL3 Workshop, CVPR, 2020
|
|
SegRPN: Scale Aware Joint Object Detection and Semantic Segmentation
Nakul Agarwal, Wei-Chih Hung, Yi-Hsuan Tsai, Ming-Hsuan Yang
Manuscript, 2019
paper
|
|
Improving Multiclass Classification by Deep Networks using DAGSVM and Triplet Loss
Nakul Agarwal, Vineeth N. Balasubramaniam, C. V. Jawahar
Pattern Recognition Letters, 2018
paper   bibtex
|
|
Connecting Visual Experiences using Max-flow Network with Application to Visual Localization
Nakul Agarwal*, A.H. Abdul Hafez*, C. V. Jawahar
arXiv, 2018
paper   bibtex
|
|
Exploring Action Recognition without using Deep Learning
Nakul Agarwal, 2019
report
|
|
Content Based Image Retrieval
Nakul Agarwal, 2018
report
|
|
Simultaneous Localization and Mapping using Extended Kalmann Filter
Nakul Agarwal, Aditya Ranganath, 2017
report
|
|
Software Engineering (CSE 120), UC Merced
Teaching Assistant (TA) with Chi Yan Leung
Fall 2017
|
|
Computer Architecture (CSE 140), UC Merced
Teaching Assistant (TA) with Chi Yan Leung
Spring 2018
|
|
Intro to Digital Image Processing (CSE 107), UC Merced
Teaching Assistant (TA) with Shawn Newsam
Fall 2018
|
|