Sateesh Kumar Β· X (formerly Twitter)

Sateesh Kumar

107

posts

user avatar

@sateeshk21

San Diego, California

Joined August 2019

  • Pinned

    user avatar

    Which data is best for training few-shot imitation policies for robot manipulation? Some think it’s the data that looks similar, or has similar motion, or comes with related language labels. They are all right AND wrong: depending on the task, sometimes this similarity helps but

    00:00

  • user avatar

    I am presenting COLLAGE 🎨 at

    @corl_conf

    today. Spotlight presentation: 3:30 pm Poster: 4:30 - 6:00 pm. Poster #41. COLLAGE 🎨 is a data curation approach that automatically combines data subsets selected using different metrics, by weighting each subset based on its relevance

    user avatar

    Which data is best for training few-shot imitation policies for robot manipulation? Some think it’s the data that looks similar, or has similar motion, or comes with related language labels. They are all right AND wrong: depending on the task, sometimes this similarity helps but

    00:00

  • user avatar

    Excited to share our work MimicDroid. Humanoid learns by watching human videos!

    user avatar

    Intelligent humanoids should have the ability to quickly adapt to new tasks by observing humans Why is such adaptability important? 🌍 Real-world diversity is hard to fully capture in advance 🧠 Adaptability is central to natural intelligence We present MimicDroid πŸ‘‡ 🌐

    00:00

  • user avatar

    Shivin and team show how to turn data curation for behavior cloning into a supervised learning problem using DataModels. Clean formulation and strong results!

    user avatar

    Ever wondered which data from large datasets (like OXE) actually helps when training/tuning a policy for specific tasks? We present DataMIL, a framework for measuring how each training sample influences policy performance, hence enabling effective data selection 🧡

    00:00

  • user avatar

    Super impressive work on fine-tuning IL policies with RL using sparse rewards!

    user avatar

    πŸš€ Despite efforts to scale up Behavior Cloning for Robots, large-scale BC has yet to live up to its promise. How can we break through the performance plateau? Introducing πŸ”₯FLaRe: fine-tuning large-scale robot policies with Reinforcement Learning. robot-flare.github.io 🧡

Read the original on x.com β†—