9/ ChatDev – a virtual chat-powered software development company mirroring the waterfall model; shows efficacy in software generation, even completing the entire software development process in less than seven minutes for less than one dollar.
@dair_ai
-

LLMs Self-Align Without Finetuning via Self-Boosting
By
–
4/ LLMs Can Align Themselves without Finetuning? – discovers that by integrating self-evaluation and rewind mechanisms, unaligned LLMs can directly produce responses consistent with human preferences via self-boosting.
-

Survey of Hallucination Phenomena in Large Language Models
By
–
6/ A Survey of Hallucination in LLMs – classifies different types of hallucination phenomena and provides evaluation criteria for assessing hallucination along with mitigation strategies.
-
Quadrupedal Robot Learns Parkour via Vision-Based Policy Transfer
By
–
5/ Robot Parkour Learning – presents a system for learning end-to-end vision-based parkour policy which is transferred to a quadrupedal robot using its ecocentric depth camera.https://t.co/6720eLUt7w
— DAIR.AI (@dair_ai) 17 septembre 20235/ Robot Parkour Learning – presents a system for learning end-to-end vision-based parkour policy which is transferred to a quadrupedal robot using its ecocentric depth camera.
-
EvoDiff: Diffusion Models for Controllable Protein Generation
By
–
3/ EvoDiff – combines evolutionary-scale data with diffusion models for controllable protein generation in sequence space; it can generate proteins inaccessible to structure-based models.https://t.co/8XhjIfeAOT
— DAIR.AI (@dair_ai) 17 septembre 20233/ EvoDiff – combines evolutionary-scale data with diffusion models for controllable protein generation in sequence space; it can generate proteins inaccessible to structure-based models.
-

The Rise and Potential of LLM Based Agents
By
–
2/ The Rise and Potential of LLM Based Agents – a comprehensive overview of LLM based agents; covers from how to construct these agents to how to harness them for good.
-

Qwen-VL: Advanced Vision-Language Model for Multiple Tasks
By
–
10/ Qwen-VL – introduces a set of large-scale vision-language models demonstrating strong performance in tasks like image captioning, question answering, visual localization, and flexible interaction.
-
FaceChain: Personalized Portrait Generation Framework Using AI
By
–
9/ FaceChain – a personalized portrait generation framework combining customized image-generation models and face-related perceptual understanding models to generate truthful personalized portraits; it works with a handful of portrait images as input.
-
Nougat: Neural Optical Document Understanding for Academic PDFs
By
–
6/ Nougat – proposes an approach for neural optical understanding of academic documents; it supports the ability to extract text, equations, and tables from academic PDFs, i.e., convert PDFs into LaTeX/markdown.
-

FacTool: Detecting Factual Errors in LLM Generated Text
By
–
7/ Factuality Detection in LLMs – proposes a tool called FacTool to detect factual errors in texts generated by LLMs; shows the necessary components needed and the types of tools to integrate with LLMs for better detecting factual errors.
