Dude… Why do you think the whole point is to have properly tagged data? Or why there were thousands of people working at Amazon Turk categorizing images and files for cents per document?
No. You still have to give it a starting point and the starting point is manually configured tables basically.
And a lot of how it “learns” is by adding more of these into the tables. Every time someone make a video that “brokes” a LLM, these companies will put put those trick questions and screw in the correct answer to train the next generation manually.
Which is why the messing up counting stuffs stayed broken for so long, there’s infinite amount of variations of things that can be counted and validated by a human extremely easily.
Dude… Why do you think the whole point is to have properly tagged data? Or why there were thousands of people working at Amazon Turk categorizing images and files for cents per document?
No. You still have to give it a starting point and the starting point is manually configured tables basically.
And a lot of how it “learns” is by adding more of these into the tables. Every time someone make a video that “brokes” a LLM, these companies will put put those trick questions and screw in the correct answer to train the next generation manually.
Which is why the messing up counting stuffs stayed broken for so long, there’s infinite amount of variations of things that can be counted and validated by a human extremely easily.