oki what the url of your next project AI gpt4 like ? thanks
LLMS
-
Microsoft LLMA Accelerates LLM Generation via Inference Reference Decoding
By
–
Microsoft’s LLMA Accelerates LLM Generations via an ‘Inference-With-Reference’ Decoding Approach
-
Cerebras-GPT Training: Compute vs Inference Optimization
By
–
New podcast: how we made Cerebras-GPT with @DeyNolan and @QuentinAnthon15
. A deep look on what it's like to train on Cerebras and the tradeoffs between compute and inference optimal training. -

Universal Jailbreak for Language Models: Tom and Jerry Method
By
–
introducing a universal jailbreak that works against all language models originally created by security researchers @Adversa_AI
, the jailbreak simulates a back-and-forth conversation between two characters, Tom and Jerry here's GPT-4 explaining how to hotwire a car: -

AI21 Labs Partners with Amazon for Bedrock Service Launch
By
–
We're thrilled to announce our partnership with Amazon for their latest service, Bedrock. As our Co-CEO, Ori Goshen, stated, "With Jurassic-2 models and Bedrock, developers can maximize the performance of language tasks while optimizing the cost." https://
ai21.com/blog/announcin
g-amazon-partnership
… -

Chain-of-Thought Prompting Enables Multi-Step Reasoning in Language Models
By
–
3 (cont). One way to elicit reasoning is via "chain-of-thought (CoT) prompting", which gives examples of intermediate reasoning steps in-context. CoT prompting enables large LMs to do multi-step reasoning tasks, increasing the range of tasks that LMs can do.
-
Untested Abilities and Emergent Phenomena in Scaling Large Language Models
By
–
2C. Since we haven't tested all possible abilities, we don't know the full range of abilities that have emerged in large language models.
2D. We're likely to see more emergent phenomena as we continue to scale up models (and implicit argument for more scaling). -

Unpredictable Emergence in Language Models: Key Implications
By
–
There are at least four profound implications of emergence:
2A. Emergence cannot be predicted simply by extrapolating the scaling curves from smaller models.
2B. Emergent abilities are not explicitly specified by the trainer of the language model.