University of Rochester · NIH-supported accessibility line · 2023-2024
OSCaR: object state captioning for egocentric video

Built the OSCaR codebase and host its five checkpoints and dataset on Hugging Face; object-state captioning and state-change representation for egocentric video (Findings of NAACL 2024), supported in part by an NIH R01 on accessible video description.
- Ownership 22 of 26 commits in the team's repository
- Released five checkpoints and the dataset on Hugging Face
- Who uses it downloaded by other groups every month
- Published Findings of NAACL 2024
Most commits in the team’s repository are mine (22 of 26); the checkpoints and dataset I host on Hugging Face are downloaded by other groups every month.