LLMs are quite capable coders. Now, people say they are limited by programming languages designed for humans. I think it is the opposite: they are enabled by programming languages designed for humans. You cannot just design a new programming language for LLMs because where
@rasbt
-
LLM Comparison Issues Apply Equally Between Companies
By
–
Yeah, but if company x compares their LLM to that of company Y the same concept (issue) applies.
-
The Challenge of Evaluating LLMs: Balancing Marketing Claims and Independent Testing
By
–
It’s kind of a dilemma. You want to check independent evals because the original ones might be inflated for marketing purposes.
At the same time, independent evals may also undersell the LLM because of accidental bad prompting, bad batching, bad optimization etc -
Black Forest Labs FLUX model licensing uncertainty clarified
By
–
Black forest labs engineer says its their FLUX model. Maybe they licensed different ones. Not sure.
-
Heavy Lifting with Subword Splits and Max Length Settings
By
–
Literally heavy lifting with subword splits and max_length sets and all
-
Open Weight Models and On-Device AI Privacy Benefits
By
–
Releasing the open weight models to the research community and developers in general is not a bad thing though. On the topic of on-device. It can be interesting for a lot contexts though, which is why Apple does/tried it. One is privacy (I don’t think most people want third
-
AWS passive income strategy hosting open source LLMs
By
–
He can sit back while startups develop and host the best open source LLMs on AWS. Solid passive income strategy.
-
Nvidia Chips as a Subscription Service Model
By
–
Actually I think that’s smart. This way you don’t have to deal with reselling and recycling old chips every few years. Nvidia chips should be offered as a subscription 😀
-
Training Materials Using Project Gutenberg Public Domain Corpus
By
–
Yes! The bonus materials include training on the Project Gutenberg public domain book corpus. I don’t want to go beyond that though and curate other datasets because of copyright concerns. However, you could eg use the FineWeb dataset which is available from hugging face.
-
Reinforcement Pre-training Already Explored in Earlier Research
By
–
Have to look into this one more, but it’s already been a thing; eg “Reinforcement Pre-training” from earlier this summer