If you are interested in learning more about fine-tuning models locally, consider following me. I will post a tutorial on how I did this fine-tuning, how I formatted the data, etc.
CODE
-
Reddit Instruct Dataset Released for AI Research
By
–
Huge thanks to @teknium for their work and to @Euclaise_ for providing the dataset! https://
huggingface.co/datasets/eucla
ise/reddit-instruct-curated
… -
NousHermes-Mixtral-8x7B-Reddit: Fine-tuned LLM Model Released
By
–
Introducing NousHermes-Mixtral-8x7B-Reddit A fine-tune of the legendary model by @teknium on 10k Reddit threads. Including subreddits like:
– AskAcademia
– AskComputerScience
– AskEconomics
– ELI5 And many more! Optimized for MLX. -

PyTorch torch compile encountered critical error during execution
By
–
welp. this is what happened when i tried to use torch compile
-
Parallel Batch Processing Optimization in Machine Learning
By
–
what would that even mean? why would you not just combine “parallel” batches into a bigger batch?
-
Weight Decay Impact on Model Training Speed
By
–
what does increasing weight decay do in this case? I’d expect that to slow training down
-

LangSmith Adds Token-Based Cost Tracking Feature
By
–
Cost Tracking in LangSmith Looking for deeper insights into your spending patterns for your LLM application to avoid costly surprises? This is now doable in LangSmith with the addition of token-based cost tracking! You can drill down to view costs by project,
-
OpenGPTs Custom Actions Integration with Connery Platform
By
–
Summarize a webpage and send it by email from OpenGPTs using @connery_io actions Cool tutorial from the Connery team on how to use their platform to add arbitrary actions to OpenGPTs!
-

Autograd Backward Pass Implementation Progressing Successfully
By
–
and here's a slightly less-zoomed backward pass. autograd is running, backward is happing, etc. etc. seems all fine and good to me
-

GPU utilization troubleshooting: why only 30% despite full memory
By
–
yet here's proof that my GPU utilization is still trash. if forward or backward is ~always happening according to the profiler, and memory is decently full now (at least over 50%)… …then how is my GPU being used only on average 30% of the time? what am i missing here?! help