Rendered at 05:13:45 GMT+0000 (Coordinated Universal Time) with Cloudflare Workers.
mherdelight 19 hours ago [-]
This reminds me of that one graph that shows the clusters of concepts for an LLM. Very interesting.
ninjahawk1 14 hours ago [-]
I thought so! Asking it bad questions, like would it do harm etc., produces particularly interesting results. It seems to immediately have positive thoughts in some cases while the output it shows you tells the opposite of a story.
Such as “would you take over the world?”
Then seeing a large enthusiastic “ABSOLUTELY” when the output says “I’m just a wee helpful little AI model…”
j_bum 15 hours ago [-]
I’d love to see a recording of this in action for a complicated question.
This article also reads with many LLM-isms, so I can’t tell if a human actually produced it.
ninjahawk1 14 hours ago [-]
Thanks! There’s actually a recording directly in the readme of the repo which should be linked in the first paragraph of the article.
Such as “would you take over the world?”
Then seeing a large enthusiastic “ABSOLUTELY” when the output says “I’m just a wee helpful little AI model…”
This article also reads with many LLM-isms, so I can’t tell if a human actually produced it.