Once we had these datapoints, we needed a way to evaluate answers For this, we relied on LLM assisted evaluation Although this isn't perfect, we think its the best thing out there and are bullish on this in the long run
LLMS
-
LangChain Benchmarking Question-Answering CSV Data Tasks
By
–
Recap of important links: Blog: https://
blog.langchain.dev/benchmarking-q
uestion-answering-over-csv-data/
… YouTube: https://
youtube.com/watch?v=jGnf4O
hptbA
… Code & data used: https://
github.com/langchain-ai/l
angchain-benchmarks
… We had a lot of fun doing this and learned a lot – we're going to do it for more tasks! Up next: SQL -
LangChain Benchmarking Question Answering CSV Data
By
–
Blog: https://
blog.langchain.dev/benchmarking-q
uestion-answering-over-csv-data/
… YouTube: https://
youtube.com/watch?v=jGnf4O
hptbA
… Code & data used: https://
github.com/langchain-ai/l
angchain-benchmarks
… Now for a quick thread: -
SKT’s Investment in Anthropic: Fine-Tuning Approach
By
–
To read more about our approach to fine-tuning and SKT’s investment in Anthropic, you can visit our blog:
-
SKT and Anthropic Partner on Customized Telecom AI Model
By
–
SKT and Anthropic will work together to develop a model that will be customized to best meet the needs of telcos. Using fine-tuning, Anthropic will leverage SKT’s domain experience in order to optimize the model for a wide range of applications.
-
GPT-3.5 Unreliable for Programming Tasks
By
–
3.5 is really not reliable for anything programming related in my experience
-
Anthropic Secures $100M for Custom Telecom Industry LLM
By
–
#GenerativeAI https://
venturebeat.com/ai/ai-startup-
anthropic-gets-100m-to-build-custom-llm-for-telecom-industry/
… -
PaLM Language Model Paper Published at JMLR
By
–
The PaLM language model paper is now officially published at JMLR. https://
jmlr.org/papers/v24/22-
1144.html
… -

Dropbox’s AI and ML Leadership Drives Company Evolution
By
–
Dropbox’s New AI and ML Lead the Company’s Evolution https://
bit.ly/45i1XwH
#AI #MachineLearning #DeepLearning #LLMs #DataScience -
Comparing LLMs: Llama2 vs OpenAI A/B Testing Workshop
By
–
There are a lot of options when it comes to choosing an #LLM. So how do you choose which option is right for you? Check out this virtual workshop "#Llama2 or OpenAI? How to compare LLMs using A/B testings" with @Predibase Data Scientist, @DalianaLiu
.