These capabilities allow models to access knowledge beyond their training data. OSS models have made progress with releases from @AiEleuther @Meta @BigscienceW @StabilityAI @TIIuae @salesforce @BigCodeProject @databricks however their ability to use software APIs is unclear(3/10)
@sambanovaai
-
Open Source LLMs Become Effective Tool Manipulators
By
–
TECHNICAL UPDATE: We are excited to share our work on how to enable OSS to become effective tool manipulators. https://
sambanova.ai/blog/enabling-
open-source-llms-to-become-effective-tool-manipulators/
… The benchmarks, leaderboard, and tuned models are available on @huggingface
. (1/10) -
Teaching LLMs to Use Tools Over Direct Computation
By
–
For example, it's better to teach a LLM to use a calculator rather than teaching it to do complex math. (2/10)
-
First Deep Learning LLM Purpose-Built for Financial Services
By
–
Live in London: it was an honor for our team to present the first Deep Learning large language model purpose-built and pre-trained specifically for the financial services industry. #ai #artificialintelligence #financialservices #nlp #deeplearning #deeplearningai #aifs23
-

SambaNovaAI CEO Rodrigo Liang Hosts NYSE Event
By
–
Much appreciation to @NYSE for hosting @SambaNovaAI CEO and Co-Founder @RodrigoLiang today in NYC.
-

SambaNova CEO discusses enterprise generative AI in oil gas
By
–
"You better better know how to go with the flow." PODCAST:
@SambaNovaAI CEO and Co-founder
@RodrigoLiang joins the Digital Innovations in Oil and Gas with @geoffreycann Podcast. Listen to the full episode at https://
digitaloilgas.libsyn.com/rodrigo-liang-
on-using-enterprise-generative-ai-tools-in-oil-and-gas
… -
Open License Model Released for Long Sequence Understanding
By
–
It is available with an open license and created to understand long sequence capabilities. It is not meant to be a drop-in replacement for chat models. We are excited to see how people use the checkpoint. (9/10)
-
Open Source AI Checkpoint Released on Hugging Face
By
–
When we do a LikeRT human evaluation of our checkpoint and compare with other models, we achieve comparable results. We are open sourcing the checkpoint on @huggingface for the community to try. (8/10)
-
SambaNova Releases SN-13B-8k-Instruct Language Model
By
–
@huggingface | SN-13B-8k-Instruct, a 13 billion parameter model https://
huggingface.co/sambanovasyste
ms/SN-13B-8k-Instruct
… Join our Discord to ask questions and discuss. https://
discord.gg/8z2Pe7cpRv Read the full blog and technical details at https://
sambanova.ai/blog/training-
long-sequence-size-models-on-sambanova/
… (10/10) -
13B Parameter Long Sequence Model Achieves Competitive Accuracy
By
–
Using this recipe, we are able to train a competitive long sequence model at 13B parameter scale. We achieve 2-12 points better accuracy across a wide variety of long sequence tasks from Scrolls and ZeroScrolls. (7/10)