flippo.eurosky.social · @flippo.eurosky.social@bridge.neodb.net
Wants to read Reinforcement Learning from Human Feedback: LLM alignment and post-training
#book
Published · bridged from Atmosphere
ActivityPub object: https://bridge.neodb.net/users/flippo.eurosky.social/posts/3muoezxsqlsya
AT Protocol record: at://did:plc:nsisjbau4j5hyzppkpppenms/buzz.bookhive.book/3muoezxsqlsya