For anyone who doesn’t know, theirs a video going around of a Timelapse of OpenAI’s newest model GPT-6, painting someone’s picture on Canva. I feel like this isn’t really hard for an LLM, at a high level it understands the screenshot (as it’s still screenshot based, doesn’t get video feed of PC) as color values of course it doesn’t really see in the same way humans do. And so since it understands the color incredibly precisely it basically just has to do a color match on Canvas color pallete from the portraits original pixel/token color data. And then like I guess scale the actual position of the original color to the portrait cords. Idk how to explain it well but basically saying since it knows the exact colors it’s basically color matching and then copying that in Canva which isn’t that impressive. What I find way more impressive is the long term planning and backtracking it’s able to do more than being able to “paint a portrait”. And also the rest of using a browser and stuff has been possible for a while now so yeah. I might be wrong though! submitted by /u/meh_coder
Originally posted by u/meh_coder on r/ArtificialInteligence
