today we are announcing reinforcement finetuning, which makes it really easy to create expert models in specific domains with very little training data. livestream going now: https://
openai.com/12-days/ alpha program starting now, launching publicly in q1
LLMS
-
OpenAI Announces Reinforcement Finetuning for Expert Domain Models
By
–
-
OpenAI adds o1-mini to fine-tuning
By
–
Teasing o1-mini in the fine-tuning dropdown on OpenAI Platform
-

Databricks Partners with Meta on Llama 3.3 Model
By
–
Thrilled to partner with @AIatMeta on the new Llama 3.3 model on Databricks. Starting next week, our customers will be able to serve the updated model to build agent systems. Together, we're making AI more accessible & cost-effective for developers & enterprises.
-

Llama 3.3 70B Instruct Now Available on Hugging Chat
By
–
Llama 3.3 70B Instruct now on Hugging Chat! Try it out here: https://
huggingface.co/chat/models/me
ta-llama/Llama-3.3-70B-Instruct
… -

Meta Defies Scaling Laws with Llama 3.3 70B Model
By
–
Meta is challenging “death of scaling law” rumors with Llama 3.3 70B.
They’re defying traditional scaling limits, improving models without increasing parameters or changing the fundamental model architecture. Quality matters, not just quantity. https://
hubs.la/Q02-LBLK0 -
A legal agent fine-tuned with o1-mini?
By
–
Will we see an o1-mini fine-tuned legal agent? The model will learn how to reason in a custom domain
-
Meta’s Llama 3.3 70B Challenges Death of Scaling Law
By
–
Done – https://
groq.com/a-new-scaling-
paradigm-metas-llama-3-3-70b-challenges-death-of-scaling-law/
… -

Llama 3.3 70B Challenges Death of Scaling Law
By
–
Your wish is our command – https://
groq.com/a-new-scaling-
paradigm-metas-llama-3-3-70b-challenges-death-of-scaling-law/
… -
Developers can fine-tune o1 for custom LLMs
By
–
Fine-tuning will be possible for o1. Devs will be able to build o1 level custom use LLMs