Yeah, that's pretty much what I figured – at this point I'm wondering if there are ANY examples of fine-tuned models to serve additional knowledge out there
LLMS
-
Fine-tuning models for knowledge addition and RAG applications
By
–
That's not quite what I'm looking for – I believe that fine-tuning can work for things like adjusting the tone of the model's responses, the thing I'm seeking evidence for is fine-tuning in order to add additional knowledge to a model such that it can be used for RAG-style Q&A
-
Fine-tuning models for knowledge acquisition beyond RAG
By
–
Phind is running RAG – you can tell from the sources it displays in the side-bar They may have fine-tuned a model for being more effective at RAG, but that's not what I'm looking for here – I want an example of a model that was fine-tuned in order to bake additional knowledge
-
RAG vs Fine-tuned Models: Understanding the Key Differences
By
–
That looks like RAG or Q&A against a prompt context, not a custom fine-tuned model
-
Cursor’s OpenAI Technology: RAG or GPT-3.5 Fine-tuning?
By
–
The smallprint says "Cursor uses artificial intelligence technology provided by OpenAI, LLC" – so presumably it's RAG? Or did they do a fine-tune of GPT-3.5?
-
Targeted AI Systems Answering Documentation Questions Efficiently
By
–
By "targeted" I mean something like "answer questions about this documentation" – the kind of thing you might alternatively use RAG for
-
Fine-tuned Chatbot Demo with Targeted Knowledge Addition
By
–
Does anyone have a link to a chatbot demo I can try out right now that's serving a model that has been fine-tuned to add additional targetted knowledge (not behavior – actual information that it can answer questions about) to that model?
-
Cohere Build Day NYC showcases Command R AI models
By
–
The debut of Cohere Build Day in New York City was a testament to the power of collaboration and continuous learning in the AI industry. Attendees from Microsoft, Salesforce, Github, and more gathered to experience Command R and R+ hands-on.
— Cohere (@cohere) 2 mai 2024
Now we’re gearing up for Toronto,… pic.twitter.com/0xnMygSXPuThe debut of Cohere Build Day in New York City was a testament to the power of collaboration and continuous learning in the AI industry. Attendees from Microsoft, Salesforce, Github, and more gathered to experience Command R and R+ hands-on. Now we’re gearing up for Toronto,
-
Llama-3 Memory Requirements: 5GB to 40GB Across Model Sizes
By
–
Nope. From 5 GB for llama-3 8B to 40GB for llama-3 70B. Check it out here at @ollama
: