Super smart. Who are some of the best creators of this sort that own the full stack, so to speak?
AI
-
Beta AI Product Legal Arguments and Commercial Licensing Strategy
By
–
The current version is for non-commercial use because it's "beta" and they're still testing the strength of their legal argument, hehe.
-
Decision importance based on human and societal impact
By
–
The importance of a decision is significantly based on how it affects you and other people.
-

Self-Retrieval Architecture for Long-Range Language Modeling
By
–
9/ Long-range Language Modeling with Self-Retrieval – an architecture and training procedure for jointly training a retrieval-augmented language model from scratch for long-range language modeling tasks.
-

Scaling MLPs: How Inductive Bias Compensates at Scale
By
–
10/ Scaling MLPs: A Tale of Inductive Bias – shows that the performance of MLPs improves with scale and highlights that lack of inductive bias can be compensated.
-

Theory-of-Mind Evaluation Framework for Large Language Models
By
–
7/ Understanding Theory-of-Mind in LLMs with LLMs – a framework for procedurally generating evaluations with LLMs; proposes a benchmark to study the social reasoning capabilities of LLMs with LLMs.
-

Self-Supervised LLM Evaluation Framework for Deployment Monitoring
By
–
8/ Evaluations with No Labels – a framework for self-supervised evaluation of LLMs by analyzing their sensitivity or invariance to transformations on input text; can be used to monitor LLM behavior on datasets streamed during live model deployment.
-

DragDiffusion: Precise Spatial Control in Diffusion-Based Image Editing
By
–
6/ DragDiffusion – extends interactive point-based image editing using diffusion models; it optimizes the diffusion latent to achieve precise spatial control and complete high-quality editing efficiently.
-

Computer Vision Through Natural Language: LLM-Based Modular Approach
By
–
3/ Computer Vision Through the Lens of Natural Language – a modular approach for solving computer vision problems by leveraging LLMs; the LLM is used to reason over outputs from independent and descriptive modules that provide information about an image.
-

Visual Navigation Transformer: Pretrained Model for Robotic Navigation
By
–
4/ Visual Navigation Transformer – a foundational model that leverages the power of pretrained models to vision-based robotic navigation; built on a flexible Transformer-based architecture that can tackle various navigational tasks.