Skip to content

Author

Katia Sycara

4 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#natural language process... Preprint Sep 2026

Video2Skill: From Streaming Experience to Reusable Embodied Skills

This work introduces Video2Skill, a benchmark that covers robot tabletop manipulation and human kitchen activity and tests three core capabilities: locating manipulation events in time, grouping events of the same transformation, and deciding when to reuse an existing skill or create a new one.

Jian-Shu Zhang, Ce Zhang, Xi-Yuan Yang et al. · 0 citations
Jul 2026

LENS: Adaptive Spatio-Temporal Zooming for Keyframe Sampling in Long-Form Videos

LENS is a training-free keyframe sampling framework that dynamically decides when to zoom in for fine-grained details and when to zoom out for broader context based on the text query, enabling the model to reason across multiple granularities while capturing both high-fidelity details and long-range context.

Ce Zhang, Jinxi He, Katia Sycara et al. · 2 citations
Preprint Aug 2026

StreamScout: Learning When to Look Deeper for Streaming Video Understanding

Streaming video understanding requires answering questions that arrive at arbitrary moments over an unbounded video stream. Existing systems primarily focus on what to retain in a bounded memory, yet access that memory using the same fixed-cost procedure for every query, despite substantial variation in the evidence re...

Ce Zhang, Jing Bi, Jinxi He et al. · 2 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.