An interesting historical note is that neural language models have actually been around for a very long time but noone really cared anywhere near today's extent. LMs were thought of as specific applications, not as mainline research unlocking new general AI paths and capabilities
LLMS
-

Argonne’s GPT Model Wins Gordon Bell Prize for COVID Research
By
–
In partnership with @Argonne
, we’ve been awarded the 2022 Gordon Bell Special Prize for HPC Based COVID-19 Research! We trained ANL’s GPT-style models on full Covid-19 genome, holding the entire sequence length (10240 tokens) on a single device. Read more: https://
hubs.li/Q01sCD600 -
GALA Model Code Released on GitHub for Researchers
By
–
The code is available. https://
github.com/paperswithcode
/galai
… So, we can try, but the general public can't. -

LangChain 0.0.15 Release: Dark Mode and SQL Improvements
By
–
LangChain version 0.0.15 – "the @nlarusstone release" improve color highlighting so it looks good in dark mode @nlarusstone (see below) add tables to ignore/include in SQL DB chain @nlarusstone add concept of document metadata add `apply` method to all chains
-

Groq Showcases Compiler, Developer Tools, and RealScale at Supercomputing
By
–
At @Supercomputing
? Drop by booth 3047 to:
– Discuss your model! Did you know 500+ models run on our #compiler?
– See dev tools like GroqFlow & GroqAPI, made for fine-grained control
– Discuss RealScale™, the technology that extends performance and #lowlatency from #chip to rack -
Galactica Model Repetition Issues Decoding Parameters Analysis
By
–
I’m a bit wondering why we see so much repetitions in the outputs of Galactica people share. If it comes from the model itself or the demo decoding parameters
-
Reliable Knowledge from Parametric Memory in AI Remains an Open Problem
By
–
Yeah — getting reliable, truthful knowledge entirely from parametric memory seems like an open problem. The near-term solutions all involve grounding with queries to external resources.
-
AI Prompting Techniques and Hallucination Challenges
By
–
I spent some time trying to solve this problem through CoT/scratchpads but there were often secondary hallucinations, e.g. over-skepticism of non-trick questions. Nothing seemed to work as well as the zero-shot above.
-
LLM LLCs Focus on AI Power, Not Decentralization Like DAOs
By
–
haha, I'm high level familiar with DAOs and I don't think so. LLM LLCs are about AI Power, not about decentralization, transparency, or governance. Actually in many ways opposite of DAOs in a basic execution of the idea.
-
Language Models Continue Sequences from Prompts, Not Maximize Rewards
By
–
they don't maximize rewards, they are given a prompt (a kind of inception) and continue the sequence