New from us: Given they are trained on human data, can you use psychological techniques that work on humans to persuade AI? Yes! Applying Cialdini's principles for human influence more than doubles the chance of GPT-4o-mini agrees to objectionable requests compared to controls
@emollick
-

EU FLOP Limits Force Major AI Models Into Compliance
By
–
So every major model is already exceeding or will soon exceed the EU's systemic risk FLOP limit when it comes into effect next year.
-

LLMs Struggle With Coherent Adventure Generation and Self-Learning
By
–
Historically, LLMs have trouble building a coherent adventure because there are many pieces that fit together. I also thought it was interesting that it researched how to write an adventure first, and followed the rules it discovered.
-

ChatGPT Agent Generates 19-Page D&D Adventure PDF
By
–
ChatGPT agent: "create a PDF of a novel D&D adventure, add illustrations, make it super interesting and deep, add tables, etc" "Fix the formatting, build it out more" Got a 19 page PDF. Agent doesn't do layouts well, but pulls off building a coherent adventure, hard for LLMs.
-
Bad dataset quality impacts AI model performance outcomes
By
–
The dataset I gave the AI was bad, it worked with that data.
-
Co-Intelligence: How AI Amplifies Human Expertise and Productivity
By
–
I saved what would have taken a tremendous amount of time, acting as a manager rather than an analyst, but my expertise was really important in guiding the AI in the right direction. It was, dare I say, a good example of co-intelligence.
-

ChatGPT Agent Limitations: Data Analysis Power and Human AI Collaboration
By
–
An example of the power & limitations of ChatGPT agent I asked it to analyze a dataset from Kaggle, and turn it into a PPT and Excel. It made no errors, but I thought some of the data was odd. I gave that feedback & the AI figured out the data was bad and why. Human + AI needed
-
Microsoft Missed Opportunity in AI Agents for Knowledge Workers
By
–
One implication from ChatGPT agent (not a creative name, but a descriptive one – a rare naming win!) is the labs are learning that many knowledge workers live in Excel & PowerPoint. Surprised that Microsoft did not do more to push past Copilots when they had this to themselves.
-
Manus vs o3: Comparing AI Agent Capabilities
By
–
It feels like a cross between Manus and o3. Manus is capable of somewhat more complex tasks, but agents is better at integrating research and doing a wider range of tasks well (this is early impressions, and I know they are still working on the tool)
-
AI Paradigm Shift: From Prompting to Task Delegation
By
–
It feels much more like working with an actual human intern capable of a wider range of analytical and computer tasks, and, like an intern, you want to give it feedback and work back & forth. Not all the way there yet, but the paradigm is shifting from prompting to delegating.
