ACL Outstanding Paper
Language model acceptability judgements are not always robust to context
LLMS
-
Language Model Acceptability Judgements Lack Contextual Robustness
By
–
-

Groq Enhances LLM Inference Strategy for Federal AI ROI
By
–
Following a sound #inference strategy will be the difference between success and failure when it comes to deploying #LLM workloads. We're thrilled to have Marc Wilson on our team to help Federal customers achieve a generational leap in the ROI of #AI solutions. pic.twitter.com/ZjRShiNMwo
— Groq Inc (@GroqInc) 12 juillet 2023Following a sound #inference strategy will be the difference between success and failure when it comes to deploying #LLM workloads. We're thrilled to have Marc Wilson on our team to help Federal customers achieve a generational leap in the ROI of #AI solutions.
-
Poe AI Platform Now Available for Mac Download
By
–
You can download it from http://
poe.com/download . As usual, all bots from across Poe are available, including ChatGPT, GPT-4, Claude 2, Claude Instant, PaLM, and all user-created bots. All your conversations from Poe on all other platforms will sync to Poe on Mac. -
Transform Unstructured Documents into Structured Tables with LLMs
By
–
Ready to turn your #unstructured docs into #structured tables with #LLMs? Don't miss our upcoming talk with @_odsc
. Register for free! https://
app.aiplus.training/courses/From-D
ocs-to-Tables-Generating-Structured-Data-with-LLMs
… -

Transformer Models: Introduction and Comprehensive Catalog 2023
By
–
Transformer models: an introduction and catalog — 2023 Edition https://
bit.ly/3PbEJnk #AI #MachineLearning #DeepLearning #LLMs #DataScience -

Self-Supervised Learning for LLM Pretraining and Evaluation Methods
By
–
We use self-supervised learning to pretrain LLMs (e.g., next-word prediction). Here's an interesting take using self-supervised learning for evaluating LLMs: https://
arxiv.org/abs//2306.13651
Turns out, there's correlation between self-supervised evaluations & human evaluations. -

Microsoft DeepSpeed ZeRO++ Optimizes Large Model Training Efficiency
By
–
Microsoft's DeepSpeed ZeRO++ is a system of communication optimization strategies built on top of ZeRO to offer unmatched efficiency for large model training, regardless of batch size limitations or cross-device bandwidth constraints. https://
bit.ly/46D9VSA? -

Richard Socher joins Hugging Face Hub platform
By
–
Eminent Deep learning for NLP educator, founder of @YouSearchEngine, friend of the huggingface, and recent recipient of ACL Test-of-time award, @RichardSocher is now on the huggingface Hub! https://
huggingface.co/RichardSocher Welcome -

Stability AI Launches StableLM Suite Language Models
By
–
In April, we also launched the first of our StableLM Suite of language models. Newer models are on their way!
-
Transformer Encoder Decoder Architecture for Edge Device Inference
By
–
In our latest blog article, our Director of System Architecture, @brt_mns
, introduces the transformer encoder and decoder architecture and the challenges of inferring them on edge devices. Read more https://
axelera.ai/decoding-trans
formers-on-edge-devices/
… #ArtificialIntelligence #NLP #GenerativeAI