Sateesh Kumar
107
posts
I am presenting COLLAGE π¨ at
@corl_conftoday. Spotlight presentation: 3:30 pm Poster: 4:30 - 6:00 pm. Poster #41. COLLAGE π¨ is a data curation approach that automatically combines data subsets selected using different metrics, by weighting each subset based on its relevance
Excited to share our work MimicDroid. Humanoid learns by watching human videos!
Shivin and team show how to turn data curation for behavior cloning into a supervised learning problem using DataModels. Clean formulation and strong results!
Super impressive work on fine-tuning IL policies with RL using sparse rewards!
π Despite efforts to scale up Behavior Cloning for Robots, large-scale BC has yet to live up to its promise. How can we break through the performance plateau? Introducing π₯FLaRe: fine-tuning large-scale robot policies with Reinforcement Learning. robot-flare.github.io π§΅




