Outside of the mathematical setting, large language models can be prone to making logical mistakes. Today we present an evaluation benchmark for mistake identification across settings and examine how LLMs might learn to correct their own logical errors. →
https://
goo.gle/48Ox58T
GENERATIVE AI
-

LLMs Logical Error Identification and Self-Correction Benchmark
By
–
-
DesignerGPT Video: AI-Powered Design Tool Showcase
By
–
Great video from @showprogress on DesignerGPT!
-

Midjourney: Encouragement and Support for AI Image Generation
By
–
Come on Midjourney. You can do it. I believe in you.
-

NVIDIA RTX SUPER GPUs Launch for Generative AI Performance
By
–
Announced at #CES2024: @NVIDIAGeForce RTX SUPER desktop GPUs for supercharged generative AI performance, new AI laptops from every top manufacturer, and new NVIDIA RTX-accelerated AI software and tools for both developers and consumers. https://
nvda.ws/48OuWtX #AIonRTX -
NVIDIA NeMo Custom Generative AI Models on Amazon EKS
By
–
Watch this YouTube live stream to hear NVIDIA experts discuss building and using custom generative AI models with NVIDIA NeMo on Amazon EKS. Learn how businesses can benefit from generative AI in this special episode of Containers from the Couch.
-
Researchers Jailbreak ChatGPT API by Accessing Hidden Token Probabilities
By
–
fun research story about how we jailbroke the the chatGPT API: so every time you run inference with a language model like GPT-whatever, the model outputs a full probabilities over its entire vocabulary (~50,000 tokens) but when you use their API, OpenAI hides all this info from
-
OpenChat 3.5-0106: Open-Source AI Model Now Available
By
–
HuggingFace: https://huggingface.co/openchat/openchat-3.5-0106 Live Demo: https://openchat.team GitHub: https://github.com/imoneoi/openchat @openchatdev
-
Direct Preference Optimization Paper Receives Standing Ovation
By
–
It is only rarely that, after reading a research paper, I feel like giving the authors a standing ovation. But I felt that way after finishing Direct Preference Optimization (DPO) by @rm_rafailov @archit_sharma97 @ericmitchellai @StefanoErmon @chrmanning and @chelseabfinn
. This -

Together Launches langchain-together Python Package for Embeddings
By
–
@togethercompute Embeddings Today we’re excited to announce the new langchain-together Python package, which has connections to Together’s LLM and Embeddings endpoints. This is the easiest way to access the next-gen embedding models their team has been developing. Learn
-
Policy Brief on AI by Leading Researchers
By
–
This policy brief was authored by @PeterHndrsn
, @lxuechen
, @jurafsky
, @tatsu_hashimoto
, Mark A. Lemley, and @percyliang
: