AI is a giant, transactional if/then statement about language leading to an algorithm choice.
Let's say it's 1995 and you want to erase pimples from a photo. You fire up Photoshop, select the 'tools' menu and then select a filter - there were several that would do the job in different ways. Professional photo editors would understand those ways based on trial and error and select the tool they felt would get them the results they desire without turning your high school senior (or whatever) photos into a nightmare.
Fast forward 15 years later, and those tools have become incredibly effective - professional photo editors have given feed back and iterated the filters until Adobe could create "Adobe Photoshop Elements" where the hard choices were taken away from you and the average hobo with an iPhone can make his Gutter Princess look like an unblemished Gutter Queen.
Everyone took this and went "I am a professional photo editor!"
Ten years on, someone took dozens of these tools and now, instead of a point and click or touch and tap menu selection of "Tools -> Remove this pimple" you have a language model on the front. Now you go "Hey, Adobe AI, remove this pimple". Now the AI processes your words and tries to correct for the fact that we all speak differently and guess your meaning from other context (e.g. having a pimple on the screen in an image you are looking at) and then selects "Tools -> Remove this Pimple" for you.
The problem is that, it guesses wrong as often as it guesses right. I was working with a video generator toy and told it to take an image seed and "spread the arms". It "spread the arms" at the elbow - basically actuating the joint that was detected front and center and extending the arms outward like the guy was reaching for something. Rephrasing to "spread the arms left and right" got it to actually separate the arms and move them to the sides.....and then the guy looked like someone turned him into Stretch Armstrong.
A professional video editor (which I am not) would have known how to turn that image into a 3d model, map joints and apply skeletons and all of that extra stuff and then made him believably move his arms around. AI guesses at every decision along that tree based on the snippet of language you provide and the random crap it was trained on before the language model was packed up.
This. I have posted this over and over.
AI is a giant, transactional if/then statement about language leading to an algorithm choice.
Let's say it's 1995 and you want to erase pimples from a photo. You fire up Photoshop, select the 'tools' menu and then select a filter - there were several that would do the job in different ways. Professional photo editors would understand those ways based on trial and error and select the tool they felt would get them the results they desire without turning your high school senior (or whatever) photos into a nightmare.
Fast forward 15 years later, and those tools have become incredibly effective - professional photo editors have given feed back and iterated the filters until Adobe could create "Adobe Photoshop Elements" where the hard choices were taken away from you and the average hobo with an iPhone can make his Gutter Princess look like an unblemished Gutter Queen.
Everyone took this and went "I am a professional photo editor!"
Ten years on, someone took dozens of these tools and now, instead of a point and click or touch and tap menu selection of "Tools -> Remove this pimple" you have a language model on the front. Now you go "Hey, Adobe AI, remove this pimple". Now the AI processes your words and tries to correct for the fact that we all speak differently and guess your meaning from other context (e.g. having a pimple on the screen in an image you are looking at) and then selects "Tools -> Remove this Pimple" for you.
The problem is that, it guesses wrong as often as it guesses right. I was working with a video generator toy and told it to take an image seed and "spread the arms". It "spread the arms" at the elbow - basically actuating the joint that was detected front and center and extending the arms outward like the guy was reaching for something. Rephrasing to "spread the arms left and right" got it to actually separate the arms and move them to the sides.....and then the guy looked like someone turned him into Stretch Armstrong.
A professional video editor (which I am not) would have known how to turn that image into a 3d model, map joints and apply skeletons and all of that extra stuff and then made him believably move his arms around. AI guesses at every decision along that tree based on the snippet of language you provide and the random crap it was trained on before the language model was packed up.
Artificial intelligence isn't.
User deleted by comment