That's a good question. I'd say more background is always helpful, but probably not essential. I have a ~40-page fast-track to deep learning with PyTorch in Appendix A that might get you up to speed (might be a steeper learning curve, but might be more respectful of your time).
@rasbt
-
Skepticism about Apple’s AI capabilities and Siri improvements
By
–
I am a bit skeptical. Maybe it's because Siri was neglected for so long. I think too that they have great talent but they also operate under more constrains, so I don't expect their LLMs to be quite as good as the ones from the current main LLM developers.
-
Claude 3.5 Sonnet naming confusion clarification
By
–
yes, also known as "new Claude 3.5 Sonnet", which I always found super confusing.
-
Apple’s Limited AI Capabilities and LLM Competitive Disadvantages
By
–
oh yes, of course they are limited. They were a bit late to the party. They are also not known for AI/ML work (see Siri being notoriously bad). That being said, I doubt they will have competitive large-scale LLMs. They don't have the server capacities for that at that user
-

Major LLMs Show Similar Safety Measures Despite Company Marketing Differences
By
–
Yeah. To be fair, every major company does the "safer" marketing. Yet, as a end-user the flag ship LLMs feel little different on the "safety" spectrum. Just an observation.
-
Contradictory Safety Marketing Claims Among AI Chatbots
By
–
Maybe to add a bit more context. What I mean is, based on the marketing, Claude is safer than ChatGPT because it has more content moderation guardrails, and Grok is safer than ChatGPT because it has fewer guardrails? Please make it make sense
-
Apple’s Own Foundation Models for Machine Learning
By
–
but they do have their own LLMs? https://
machinelearning.apple.com/research/intro
ducing-apple-foundation-models
… -
Claude et Grok sont-ils vraiment plus sûrs que ChatGPT?
By
–
Maybe this is a good opportunity to ask to ELI5: how are Claude and Grok safer than ChatGPT? Or taking a step back: has "safer than" (not "safety" per so) become just a marketing term?
-
Claude 3.5 Sonnet release week announcement
By
–
It's going to be an eventful week! And I am glad that it's not called "new 'new Claude 3.5 Sonnet' "
-
Training Data Contamination: Why AI Models Misidentify Themselves
By
–
Ouch. But to be fair, this could be similar to DeepSeek mistakenly identifying itself as a ChatGPT model. Likely a result of "poisoned" training data from the internet. In other words, it might just mean it's due to insufficient data filtering & system prompt design.