The best places to start learning about the Hugging Face ecosystem are our courses: https://
huggingface.co/learn and the tasks page: https://
huggingface.co/tasks
OPEN SOURCE
-
Hugging Face Learning Resources: Start Your AI Journey
By
–
-
LangChain Twitter Fine-tuning Repository Released
By
–
In the github repo! https://
github.com/langchain-ai/t
witter-finetune
… -
Key Hugging Face Team Member Departure Announcement
By
–
It’s a bittersweet moment for sure for HF Still, excited to see what’s next for you and confident it will be great, impactful work no matter what. On behalf of the whole team, and the @huggingface community, I am grateful to have been able to work with you those past 3 years
-
Apify enables Twitter data integration with LangChain for AI
By
–
.
@apify can do this pretty well! Example here: https://
python.langchain.com/docs/integrati
ons/chat_loaders/twitter
… -
Code Llama API Available on Replicate with Multiple Models
By
–
Run Code Llama with an API on Replicate. Code Llama can generate and discuss code. It's the best open-source model for things like debugging and code completion. • 7b https://
replicate.com/replicate/code
llama-7b
…
• 7b Python https://
replicate.com/replicate/code
llama-7b-python
…
• 13b -
RAG Closes Gap Between Open Source and API Providers
By
–
And using RAG could close the gap even further between the performance of open source alongside the API providers, particularly for companies that don’t want to hand over control to a provider and are looking for a cheaper, faster, and perhaps more importantly, predictable tool.
-
Open Source AI Models Accessible Without Fine-Tuning Resources
By
–
It's a tantalizing prospect for companies that are exploring the use of open source models, but don't have the resources (personnel or financial) to fine-tune or pre-train a model. It works right out of the box without any significantly advanced technical requirements.
-
Pretraining Llama 2 with New Data: Tutorial and Approach
By
–
Instead of training from scratch, you could take the existing Llama 2 base model and pretrain it for a few more epochs on new data and see how it performs. I set up a tutorial here the other day (you may want to swap the dataset depending on your usecase): https://
github.com/Lightning-AI/l
it-gpt/blob/main/tutorials/pretrain_redpajama.md
… -
Summer LLM Developments: Llama 2, CodeLlama, and GPT-4
By
–
Llama 2, CodeLlama, and leaked GPT-4 details. Here's my new write-up on the noteworthy developments around LLMs of this summer so far:
-
Implementing AI: accessibility and RLHF resources
By
–
Still a great list. Today I would add to dive as soon as possible in implementing something yourself since recent AI developments have become so accessible. We still need more good book/ressources on RLHF, maybe @natolambert or @_lewtun will fill this gap soon 🙂