decent results with 8b q4 on a 15 Pro, might be able to squeeze it into a 14 if you’re willing to use most of the phone ram
LLMS
-
Inference Performance Not Bad For Extended Usage
By
–
it’s not bad, you won’t be constantly inferencing for hours
-
ChatGPT as Our Main Ally in Projects
By
–
Quand #chatgpt est notre meilleur allié 😅 pic.twitter.com/fBRfAZZb4T
— Alexandre Tsicopoulos (@Alex_Tsico) 1 janvier 2025When #chatgpt is our best ally
-

Abacus AI LLM Fine-Tunes API Offers Affordable Open-Source Performance
By
–
Check out the @AbacusAI "LLM Fine-Tunes" inference API = Unlocks the power of affordable, Open-Source #LLMs — exceeds #GPT4o performance and is cheaper! Start here: https://
abacus.ai/llmapi
——
#LLMOps #AI #GenerativeAI #MachineLearning #MLOps #ML #GenAI #DeepLearning #DataScience -
AI Research Highlights 2024: Mixture-of-Experts and LLM Scaling Laws
By
–
Happy New Year! To kick off the year, I've finally been able to format and upload the draft of my AI Research Highlights of 2024 article. It covers a variety of topics, from mixture-of-experts models to new LLM scaling laws for precision:
-

New LLM Developer Course: 48 Notebooks and Hands-On Project
By
–
Our new course, "From Beginner to Advanced LLM Developer," contains 48 notebooks and a template repository for the final course project (because you can't learn everything and be a complete LLM developer through notebooks only). Notebooks are really powerful to teach the core
-
LLMs in 2024: Year in Review and Key Developments
By
–
link https://
simonwillison.net/2024/Dec/31/ll
ms-in-2024/
… -

LLM World Themes 2024: A Comprehensive Overview
By
–
a really nice buffet of themes from the llm world in 2024 from @simonw it’s worth a read
-
Small reasoning models improving offline phone AI capabilities
By
–
one of the best you can do offline on a phone for now, but small reasoning models are going to be much better