VSL – Utilizing variable sequence lengths to achieve longer sequences at lower cost – https://
hubs.li/Q01Z9jxc0
COMPUTING
-

VSL: Variable Sequence Lengths for Cost-Effective Long Context
By
–
-

SparseGPT: Sparsifying LLMs for efficient inference
By
–
SparseGPT – Sparsifying LLMs for efficient inference – https://
hubs.li/Q01Z9ysx0 -

Sparse models match dense accuracy with fewer flops
By
–
SPDF – Matching the downstream accuracy of a dense model with a sparse model using fewer flops – https://
hubs.li/Q01Z9v0G0 -
Tenstorrent VP Explains Why Open Source Is One-Way Door
By
–
"We realized that going open source was a one-way door,” he added. “Nobody goes to #opensource and then back to proprietary solutions. The benefits of open source are just too strong.” – Stan Sokorac, VP of Software Engineering at @tenstorrent in @eetimes
-

SAS Explore Vegas: Level Up Your Data Science Skills
By
–
Is it time to double down to level up your data science and analytics skills? Then we want to see you in Vegas! SVP of R&D Jared Peterson will be your host at SAS Explore, and he's got all the details on why you need to be there! http://
2.sas.com/6014Pn2SG #ExploreSAS -

AWS Launches NVIDIA H100 Instances to General Availability
By
–
in case you missed it @awscloud just shipped their nvidia H100 instances as GA yesterday only the 8xH100 flavor is available for now, so it's a bit hefty https://
aws.amazon.com/ec2/instance-t
ypes/p5/
… -

Tenstorrent and RISC-V Graduate Event Cambridge Next Week
By
–
It’s not too late to join us next week in Cambridge to talk all things #tenstorrent and #riscv. Sign up here —> https://
bit.ly/tt_graduate -
Deploying Dedicated Inference Endpoints for Large-Scale AI
By
–
(For large-scale deployments, you can deploy a dedicated Inference Endpoint)
