– data rephrasing https://
arxiv.org/abs/2401.16380 – multiple epochs https://
latent.space/p/jan-2024 – give @BlancheMinerva more money and grad students
@swyx
-
Data Rephrasing, Multiple Epochs, and Research Funding Priorities
By
–
-

Meeting at NeurIPS 2023 and Latent Space Papers
By
–
confirmed when i met them at neurips https://
latent.space/p/neurips-2023
-papers
… -

Infini-attention: Google’s approach to infinite context in LLMs
By
–
Can LLMs have infinite context? Researchers from Google say yes. A new paper proposed Infini-attention, which lets LLMs have infinite context. How Infini-attention works:
• has local attention like any transformer
• has global attention via compression
• combines local and -
GPT-4 Model Versions and Naming Clarification in 2024
By
–
in the past year: gpt 4.0: gpt-4-0314
gpt 3.9: gpt-4-0613
gpt 4.3: gpt-4-1106-preview
gpt 4.4: gpt-4-1106-vision-preview
gpt 4.2: gpt-4-0125-preview
gpt 4.5: gpt-4-turbo-2024-04-09 lets be clear what we are talking about when saying another model is "gpt4 level", because this -

Summary from Young Phlo shared via GitHub Gist
By
–
summary from @YoungPhlo_ https://
gist.github.com/swyxio/3b39927
36879e2c2931b91cc7127894f
… very accurate! -
Sweagent vs Devin: Speed Trade-offs and Feature Gaps
By
–
i did end up trying sweagent. its definitely faster but also very very workflow specific to making pr’s for issues. needs a lot more work to have the iterative feedback flow of devin. opendevin seems to be closer, but doesnt have browsing yet https://t.co/NXXSggMdyi
— swyx 🐣 (@swyx) 14 avril 2024i did end up trying sweagent. its definitely faster but also very very workflow specific to making pr’s for issues. needs a lot more work to have the iterative feedback flow of devin. opendevin seems to be closer, but doesnt have browsing yet
-
Devin Livestream Challenge: Comparing Devin, OpenDevin, and SWE Agent
By
–
lots of comments from 1 viral youtube from another guy who also hasnt tried devin at all haha. @GergelyOrosz we can livestream to try it out if u want, i got the go ahead to do it, theyre not scared just watch my devin vs opendevin vs swe agent livestream to understand how
-
End to End Latency vs Opus Weaviate Embeddings
By
–
whats end to end latency? vs just opus + weaviate + embeddings
-

DBRX Model Assessment: Early Judgment and Performance Catchup
By
–
judging dbrx too quickly? its big but we’ll catch up to it rather than other way around