8 ms·
Flux 3 X Mimic: The Next Generation of Video-Action Models
- rlupi 2mo agoIt's nice to see partnerships between European startups.
- DarkNova6 2mo agoWasn’t Flux purchased by Meta?
- deleted 2mo ago[deleted]
- deleted 2mo ago[deleted]
- deton3991 2mo ago[flagged]
- johnbarron 2mo agoIf Amazon does not buy them, the CEO should be fired...
- kensai 2mo agoAre you American? We would prefer it to stay a European company. I personally hope Mistral buys them or somehow they unite to make a stand.
- windexh8er 2mo agoI think you meant to type "promoted".
- Yokohiii 2mo agoBFL resides in the same region as BMW and Audi. Quite wealthy and pretty fierce to protect their car industry. If Audi or any other german car maker sees BFL as the future of automation, it will be rather hard to buy them out.
- vessenes 2mo agoReally interesting. Upshot: a well trained multimodal video generation model has a world representation model trained inside it. They’ve done some work lifting this world model out and deploying it to robots, where it seems to work well. On the one hand, this isn’t a new idea, and the quality video models certainly have understanding of materials, light, the world (at least in an Occam’s razor sense of understanding). I’m not aware of a video lab that’s turned itself into a robot lab yet, though; perhaps this would be a first, or a new sort of obvious-in-retrospect business path: train video model, sell video generation, scale, use scale to train robot things: profit. I found their hands very interesting - looks like a bunch of stuff hidden in gloves - Xiami’s Robotics-1 foundation model just released videos of users training on some pretty standard looking grippers; to the point that there are demo videos of people putting on gripper type gloves to make video to train that model. The BFL model looks like it doesn’t need that at all. Given the difficulty of the hardware side, I’ll be curious to see what they do with this.
- mike_hearn 2mo agoI think this tactic is standard, no? Nvidia has demoed robots being trained using video generation models, and Waymo has been doing that for a while too. https://waymo.com/blog/2026/02/the-waymo-world-model-a-new-frontier-for-autonomous-driving-simulation/ https://waymo.com/blog/2026/02/the-waymo-world-model-a-new-f...
- lairv 2mo agoFYI several other video labs are getting into robotics Luma Labs https://lumalabs.ai/news/luma-open-physical-ai-lab https://lumalabs.ai/news/luma-open-physical-ai-lab Runway https://runway.com/product/robotics https://runway.com/product/robotics
- davidguetta 2mo agoRhoda Ai also does vidéo to robot models
- GiffertonThe3rd 2mo agoThe video at around 3.30min, where the robot arm took 3 attempts to reseat the window trim, was quite unnerving - I have not seen such resolving before. Is it new or am I way out of the loop?
- ra 2mo agoYes that was impressive.
- dinfinity 2mo agoYou are indeed out of the loop. Google did something arguably more impressive more than a year ago, with a VLA based bot replacing a tensioned timing belt: https://www.youtube.com/watch?v=2AAFiuEP7iE https://www.youtube.com/watch?v=2AAFiuEP7iE
- quadrature 2mo agoif you want to catch up, i would say Generalist is the state of the art in dexterous manipulation tasks, you can learn more here https://generalistai.com/blog/gen-1 https://generalistai.com/blog/gen-1
- takd 2mo agoAmazing work. u.i. Zugló robot mikor?
- flufluflufluffy 2mo agoOk the phrasing here, it’s, it’s just - > However, compared to more specialized approaches for representation learning they produce less disentangled representations, which puts a ceiling on their usefulness for tasks that require world understanding. Only an LLM would use a less disentangled representation of the concept “more entangled” when trying to explain to people in the real world why less disentangled representations are not as useful for modeling the real world.
- pennomi 2mo agoBad news, humans write silly things all the time. If anything, LLMs are less likely to make awkward phrasings than people, because they aren’t found in the training data very often.
- phoghed 2mo agoWe’re just pulling signs of LLM touched writing out of our ass now. Might be time to move on from the accusations, assume all writing is at least LLM assisted and judge it purely on the quality.
- stronglikedan 2mo ago> Ok the phrasing here, it’s, it’s just - That dash definitely means an LLM wrote this comment. /s
- doubleorseven 2mo agothis is the future. but also the present. if you know some one out of the software development/engineering worlds, please forward this to them
- htrp 2mo agoOpenAi cries in sora
- idiotsecant 2mo agoWe're barreling head-on into a crisis where humans have even less value unless they own the means of production. Dark days ahead, friends.
- altern8 2mo agoThis is awesome. What's sad to me is that we have all this awesome technology, but movies are worse than ever. I usually watch movies from decades ago just to find something decent and it's amazing how with goofy-looking puppets the storytelling was 1000 times better. This is probably an unrelated rant, sorry
- phoghed 2mo agoSurvivorship bias, possible that you just aren’t watching all the shitty ones that nobody remembers.
- acdanger 2mo agotime is the great filter
- altern8 2mo agoWhat's a good recent movie I could watch? Today is Friday
- phoghed 2mo agoWas going to pull out my default recommendation of Oldboy but then I realized it’s 23 years old, which just aged me greatly
- altern8 2mo agoWelcome to the club! :-)
- thomastjeffery 2mo agoNo Other Choice (by the same director) is incredible, and even relevant to this post.
- Rebelgecko 2mo agoOdyssey is supposed to be pretty good. Main complaint I've seen is that it's not as horny as the source material.
- deleted 2mo ago[deleted]
- tancoai_dev 2mo ago[dead]
- reindeer2 2mo ago[flagged]