Sc

Scaling up robotic data collection with AI enhanced teleoperation

Hacker News

Scaling up robotic data collection with AI enhanced teleoperation

TLDR: I am using AI&more to make robotic teleoperation faster and sustainable over long periods, enabling large real robotic data collection for robotic foundational models. We are probably 5-6 orders of magnitude short of the real robotic data we will need to train a foundational model for robotics, so how do we get that? I believe simulation or video can be a complement, but there is no substitution for a ton of real robotic data. I’ve been exploring approaches to scale robotic teleoperation, traditionally relegated to slow high-value use cases (nuclear decommissioning, healthcare). Here’s a short video from a raw testing session (requires a lot of explanation!): https://youtu.be/QYJNJj8m8Hg What is happening here? First of all, this is true robotic teleoperation (often people confuse controlling a robot in line-of-sight with teleoperation): I am controlling a robotic arm via a VR teleoperation setup without wearing it, to improve ergonomics, but watching at camera feeds. Over wifi, with a simulated 300ms latency + 10ms jitter (international round trip latency, say UK to Australia). On the right a pure teleoperation run is shown. Disregard the weird “dragging” movements, they are a drag-and-drop implementation I built to allow the operator to reposition the human arm in a more favorable position without moving the robotic arm. Some of the core issues with affordable remote teleoperation are reduced spatial 3D awareness, human-robot embodiment gap, and poor force-tactile feedback. Combined with network latency and limited robotic hardware dexterity they result in slow and mentally draining operations. Often teleoperators employ a “wait and see” strategy similar to the video, to reduce the effects of latency and reduced 3D awareness. It’s impractical to teleoperate a robot for hour-long sessions. On the left an AI helps the operator twice to sustain long sessions at a higher pace. There is an "action AI" executing individual actions such as picking (the “action AI” right now is a mixture of VLAs [Vision Language Action models], computer vision, motion planning, dynamic motion primitives; in the future it will be only VLAs) and a "human-in-the-loop AI", which is dynamically arbitrating when to give control to the teleoperator or to the action AI. The final movement is the fusion of the AI and the operator movement, with some dynamic weighting based on environmental and contextual factors. In this way the operator is always in control and can handle all the edge cases that the AI is not able to, while the AI does the lion share of the work in subtasks where enough data is already available. Currently it can speed up experienced teleoperators by 100-150% and much more for inexperienced teleoperators. The reduction in mental workload is noticeable from the first few sessions. An important challenge is speeding up further vs a human over long sessions. Technically, besides AI, it’s about improving robotic hardware, 3D telepresence, network optimisation, teleoperation design and ergonomics. I see this effort as part of a larger vision to improve teleoperation infra, scale up robotic data collection and deploy general purpose robots everywhere. About me, I am currently head of AI in Createc, a UK applied robotic R&D lab, in which I built hybrid AI systems. Also 2x startup founder (last one was an AI-robotics exit). I posted this to gather feedback early. I am keen to connect if you find this exciting or useful! I am also open to early stage partnerships.

Share card

Actual performance

4points
Did not reach leaderboard

Launch Intel predictions

Analyze your own launch →
Indie HackersFits the IH revenue-focused audience · Missing: supports, reddit linkedin, podcasting
94%94% predicted probability of success on Indie Hackers, based on ML models trained on real launch data.
best fitHighest predicted score across all platforms for this description.
Product HuntOn track for Day 1 leaderboard · Strong signals: model, computer, models · Missing: mac, agents, macos
89%89% predicted probability of success on Product Hunt, based on ML models trained on real launch data.
Hacker NewsStrong engagement from HN community · Strong signals: ide, io · Missing: https docs, excited, just released
54%54% predicted probability of success on Hacker News, based on ML models trained on real launch data.
nativeThis product was originally launched on this platform.
TrustMRRLess likely to generate early MRR · Strong signals: video, way · Missing: mobile apps, ios, personal
44%44% predicted probability of success on TrustMRR, based on ML models trained on real launch data.
AppSumoMay struggle as an AppSumo deal · Missing: plus, platform, intuitive
33%33% predicted probability of success on AppSumo, based on ML models trained on real launch data.
Acquire.comPre-revenue stage for this audience · Missing: arr, mrr, revenue
28%28% predicted probability of success on Acquire.com, based on ML models trained on real launch data.
BetaListMay not resonate with beta-testers · Missing: web3, chat, crypto
0%0% predicted probability of success on BetaList, based on ML models trained on real launch data.

Incorrect prediction on native model

Similar products

Th
The AI-enhanced call experience16%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

The AI-enhanced call experience

Hacker News3
I
I Overcame Resume Frustration with AI-Enhanced Video CVS46%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

I Overcame Resume Frustration with AI-Enhanced Video CVS

Hacker News1
SpeakNotes
SpeakNotes47%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

AI Enhanced Voice Notes

Indie Hackers1$10/moai
Shimmer 2.0
Shimmer 2.061%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

ADHD Coaching, now AI-enhanced

Product Hunt+618Health & Fitness
PriceIntelGuru
PriceIntelGuru45%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

AI-Enhanced Competitive Pricing

Indie Hackerscommitment-full-time
cd
cdai – the AI-enhanced cd command25%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

cdai – the AI-enhanced cd command

Hacker News2
AI
AI enhanced drawing prompts generator24%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

AI enhanced drawing prompts generator

Hacker News2
MI
MIT-licensed distributed atmospheric data collection61%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

MIT-licensed distributed atmospheric data collection

Hacker News4
CORONAMOOD
CORONAMOOD15%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Data collection

Indie Hackers1community
En
Enhanced Heatmaps with AI-Clusters25%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Enhanced Heatmaps with AI-Clusters

Hacker News1