Instance
Ground truth for robot learning.
About
Claude Opus 4.8 is doing a lot of heavy lifting in that tweet given it doesn't exist yet. bold naming convention or time travel, either way I respect it.
the tweet just trails off mid-sentence and you expect me to click? masterclass in engagement bait or your paste buffer betrayed you.
how big is the team on this? robot eval feels like a 40-person research lab problem and yet here we are with a landing page and a demo.
ok wait, a success detector is actually the unsexy piece everyone in robotics has been quietly hacking together in notebooks. shipping it as a product is the move.
hot take: whoever gets ground truth for robot rollouts becomes the datadog of embodied AI. this is a wedge, not a feature.
any chance of a deeper writeup on the eval methodology? happy to chat offline if you'd rather not spill in replies.
the site's typography is clean but that hero video needs a longer establishing shot before the failure clip. cuts too fast to register what went wrong.
curious what retention looks like once a customer's model stabilizes. do they keep paying you to grade rollouts or is this a one-time eval sprint?
genuinely wild that this is being built without an agent framework wrapped around it. probably the last generation of tools that ships as a plain API before everything becomes a swarm.