Whoever said that is wrong. Sure, some use-cases are possible with 8k, but the interesting ones require more.
@mattshumer_
-
Community Long-Context LLaMA 3 Variants: What Exists?
By
–
Are there any long-context LLaMA 3 variants that the community has trained yet?
-

Open-Source Model Beats Claude 3 Opus at 300 Tokens Per Second
By
–
We now have an open-source model that is beating Claude 3 Opus…
— Matt Shumer (@mattshumer_) 19 avril 2024
being served at nearly **300 tokens per second** on @GroqInc.
The applications built off of this tech will be nothing short of revolutionary. pic.twitter.com/v934g0rwU5We now have an open-source model that is beating Claude 3 Opus… being served at nearly **300 tokens per second** on @GroqInc
. The applications built off of this tech will be nothing short of revolutionary. -
Groq Serves LLaMA 3 at Record 800 Tokens Per Second
By
–
My mind is blown.@GroqInc is serving LLaMA 3 at over 800 tokens per second!
— Matt Shumer (@mattshumer_) 19 avril 2024
800. Tokens. Per. Second.
This unlocks so many incredible use-cases.
It's one thing to see my demo — it's another thing entirely to experience it for yourself.
Do yourself a favor and try it asap. pic.twitter.com/Rd5NW5SDlWMy mind is blown. @GroqInc is serving LLaMA 3 at over 800 tokens per second! 800. Tokens. Per. Second. This unlocks so many incredible use-cases. It's one thing to see my demo — it's another thing entirely to experience it for yourself. Do yourself a favor and try it asap.
-

HyperWrite Launches Fine-Tuned LLaMA 3 Model
By
–
An initial HyperWrite fine-tune of LLaMA 3 is working and we're evaluating it in the HyperWrite platform. Will iterate, then ship to users soon!
-
LLaMA 3 Long-Context Training Timeline Awaited
By
–
Great! Any timeline there? Waiting to really push hard on training LLaMA 3 till I can use long-context.
-

Texting girlfriend about llama day AI model
By
–
Texting my gf, who is as outside of tech as can be, about llama day
-

LLaMA 3 Sequence Length Limitations Community Extensions
By
–
The main issue with the LLaMA 3 models is the sequence length… currently 8k. I'm confident the community will extend this pretty quickly.
-

Open-Source GPT-4-Level Models Now Freely Accessible Globally
By
–
We're entering a new world where GPT-4-level models are open-source and freely accessible. Absolutely massive.
-
Training 400B+ Parameter Model to Outperform GPT-4
By
–
They're training a 400B+ parameter version that should outperform GPT-4.