I want to know exactly what happens to documents I upload to it – how they are split and chunked, how text is extracted from PDFs, what gets embedded and stored, and what gets retrieved and included in prompts (Ideally I'd like a debug interface that lets me see it for myself)
ETHICS
-
The Ambiguity Problem in LLM Labeling and Training Data
By
–
Consider being a labeler for an LLM. The prompt is “give me a random number between 1 and 10”. What SFT & RM labels do you contribute? What does this do the network when trained on? In subtle way this problem is present in every prompt that does not have a single unique answer.
-
Privacy and Free Expression in AI: Critical Concerns
By
–
For anyone who cares about privacy and free expression, this is really important
-
Scheduling Computational Workloads to Run on Humans
By
–
# scheduling workloads to run on humans Some computational workloads in human organizations are best "run on a CPU": take one single, highly competent person and assign them a task to complete in a single-threaded fashion, without synchronization. Usually the best fit when
-
A/B Testing Risks on Developer Platforms Raise Reliability Concerns
By
–
I'm suspicious enough of A/B testing as an end-user, the idea that a developer playform I'm trying to build software on might be running A/B tests that affect my applications is horrifying
-
Disturbing Technology Raises Ethical and Safety Concerns
By
–
This is giving "Soma" level of creepiness 😬 pic.twitter.com/cyD8esi8Wl
— Ryan Browne (@Ryan_Browne_) 17 avril 2024This is giving "Soma" level of creepiness
-
Hacker Terminology: Distinguishing Ethical Hackers from Criminals
By
–
You pretty much stepped in it. Try this naming taxonomy: All hackers are good. Criminals are not hackers, they are criminals. Don't call criminals black hat hackers, call them criminals.
-
Claude Model Degradation Claims: Perception vs Reality
By
–
I'm suspicious that many cases of models "degrading" are just people imagining things – see the recent Claude incident where people complained it had degraded when the model was entirely static since launch
-

Naming Phenomena in Academia: Entropy and Conceptual Clarity
By
–
Highly recommend that academics spend the time to come up with good names for phenomena they study. And this is great: “In the second place, and more important, nobody knows what entropy really is, so in a debate you will always have the advantage."”
-
AI-Powered Robot Improves Safety for 1.25 Million US Recycling Workers
By
–
Thanks to this AI-powered #robot, 1.25 million recycling workers in the U.S. can have a safer working environment.
— Harold Sinnott #MWC26 (@HaroldSinnott) 16 avril 2024
via @gigadgets_ #recycling #Sustainability #robotics #AI #IoT #5G #FutureOfWork @gvalan @Hal_Good
pic.twitter.com/xd0TX4B6kVThanks to this AI-powered #robot, 1.25 million recycling workers in the U.S. can have a safer working environment. via @gigadgets_ #recycling #Sustainability #robotics #AI #IoT #5G #FutureOfWork @gvalan @Hal_Good