Judge just asked OpenAI's lawyer if GPT models have NYT content in its database. Lawyer said not doesn't, but then clarified it doesn’t have a database but instead relies on weights. I wonder how NYT's lawyers will respond to that.
LLMS
-
OpenAI Lawyer Explains LLM Use Case in NYT Trial
By
–
One of OpenAI's lawyers in this NYT v OpenAI trial is explaining LLMs to the judge & said someone can use ChatGPT to make a letter to a landlord "sound less angry." haha I've literally did that when our heat went out last year! (Full circle with out heat broken again today.)
-
Using agentic LLM setups to improve knowledge reliability
By
–
But using LLMs for knowledge is a thing most people already do with ChatGPT (although it's misled that they do this), so using an agentic setup to increase reliability for knowledge makes more sense IMO!
-

Agentic AI Setups Boost LLM Benchmark Performance by 40+ Points
By
–
Cliché: "Agents are just hype"
Reality: Agentic setups can easily bring >40 percentage point increase compared to vanilla LLMs on some benchmarks This crazy score increase makes sense: if I had to answer a SimpleQA question like "Which Dutch player scored an open-play goal -
Comparing o1, 4o, and Claude: Model Performance Analysis
By
–
That’s what’s been working for me lately. I'm not a huge fan of o1 for now, and I actually prefer working with 4o with Canvas over Claude, even if many prefer Claude.
-

Five Steps to Improve Writing with AI
By
–
My five steps to improve writing with AI I was working on a new video script this morning and realized I’d leaned on AI for at least half my process. This ratio just keeps growing as LLMs improve. In case it can help in your writing, here’s how it usually goes for me:
-

Perplexity AI: 10 prompts for entrepreneurs to launch a business
By
–
Perplexity AI is a powerful research assistant. It can help you research, plan, and launch a business effortlessly. Here are 10 prompts that make it the ultimate tool for entrepreneurs:
-
Classic LLMs Like ChatGPT Explanation Capabilities
By
–
Classic LLMs like ChatGPT will probably happily explain
-
Rejection Sampling and SFT with 17k Examples Discussion
By
–
I mean, rejection sampling and SFT on just 17k examples, what's not to like ?
-

Sky-T1-32B: Open-Source Reasoning Model Rivals o1-Preview Performance
By
–
"Sky-T1-32B-Preview, our reasoning model that performs on par with o1-preview on popular reasoning and coding benchmarks."
That was quick! Is this already the Alpaca moment for reasoning models? Source: https://
novasky-ai.github.io/posts/sky-t1/