Posts by u0421793@toot.pikopublish.ing
(DIR) Post #B5quFb1zJ0n2FeISlU by u0421793@toot.pikopublish.ing
0 likes, 0 repeats
@futurebird@sauropods.win this is a simplistic view – that it’s all about token prediction of similar vectors using gradient descent to arrive at the more likely next token to place in line with what is already there. There’s also RLHF – reinforcement learning through human feedback. This involves the human dressing up as a wizard and sitting behind the curtain ensuring that all the responses are the sort of responses that the wizard I mean human would actually prefer and approve of. Technically, this is achieved using a lot of smoke and some carefully placed bidirectional mirrors which act as beam splitters to construct a hologram which fools people into thinking that the machine did it all, instead of poorly paid workers.
(DIR) Post #B6Hci9cfjfzXjkqEMK by u0421793@toot.pikopublish.ing
0 likes, 0 repeats
@gutenberg_org@mastodon.social Galileo,