


Like how the brain was understood by its clinical failures, in LLM hallucinations we see bare the true causal mechanisms that lead to its answers. What we mistook for human-like cognition is revealed as something more alien:

By
–



Like how the brain was understood by its clinical failures, in LLM hallucinations we see bare the true causal mechanisms that lead to its answers. What we mistook for human-like cognition is revealed as something more alien:
By
–
Starting a campaign to rebrand automation bias as “the LGTM effect”. The LGTM effect: The unshakable tendency of humans to read the first 15% of LLM-generated prose or code, think “lgtm!”, and commit/publish/tweet something that is wrong/batshit-insane/unfunny, respectively.
By
–
For instruct-tuned GPT-3: “You” is automatic transmission, “The following is…” is stick-shift. Second-person imperative makes it emit a single answer and stop. If you pretend you’re writing a document, completions ramble on and you need to plan stop sequences.

By
–
‘Learning to see and learning to read’: #ArtificialIntelligence enters a new era https://
princeton.edu/news/2023/01/0
3/learning-see-and-learning-read-artificial-intelligence-enters-new-era
… @Princeton #AI #MachineLearning #DataScience #IoT #100DaysofCode #womenwhocode #serverless @enilev @FmFrancoise @Shi4Tech @anijov #NLP #BigData #Analytics #DeepLearning
By
–
Really appreciate that Nathan. Having the time of my life discovering all this stuff about ChatGPT.
By
–
The poem one is kind of amazing, actually. It gets very creative. Definitely others should ask ChatGPT to write random poems about stuff.
By
–
Sorry, I mis-spoke in that tweet. Not hundreds of hours with ChatGPT specifically. 750 hours at least with GPT-2, GPT-3, plus a bunch of lesser known large language models and their various modules. Another 250 hours at least with AI art tools as well.
By
–
I've been working with AI tools since GPT-2 over 4 years ago at this point I think
By
–
The addition of “very” is the prompt engineering, not the logit bias stuff. If you’re saying use of logit bias disqualifies “very” from counting as prompt engineering I don’t see how.