Volume 212 Proceedings of The Cell Segmentation Challenge in Multi-modality High-Resolution Microscopy Images https://
proceedings.mlr.press/v212/ Is now available on PMLR.
MULTIMODAL AI
-
Cell Segmentation Challenge in Multi-modality High-Resolution Microscopy
By
–
-

Vocabulary-free Image Classification Using Vision-Language Models
By
–
Vocabulary-free Image Classification paper page: https://
huggingface.co/papers/2306.00
917
…
demo: https://
altndrr-vic.hf.space Recent advances in large vision-language models have revolutionized the image classification paradigm. Despite showing impressive zero-shot capabilities, a pre-defined set of -
VideoComposer: Controllable Video Synthesis with Temporal Conditions
By
–
VideoComposer: Compositional Video Synthesis
— AK (@_akhaliq) 5 juin 2023
with Motion Controllability
project page: https://t.co/n5DgXaCRgH
presents VideoCompoer that allows users to flexibly compose a video with textual conditions, spatial conditions, and more importantly temporal conditions.… pic.twitter.com/IEDi9wHo5pComposer: Compositional Synthesis
with Motion Controllability project page: https://
videocomposer.github.io presents VideoCompoer that allows users to flexibly compose a video with textual conditions, spatial conditions, and more importantly temporal conditions. -

StableRep: Synthetic Images Improve Visual Representation Learning
By
–
StableRep: Synthetic Images from Text-to-Image Models Make Strong Visual Representation Learners Another paper demonstrating the power of training ML models using *only* synthetically generated content leading to good improvements in learning efficiency! https://
arxiv.org/abs/2306.00984 -
ChatGPT Constructs Building: Generative AI in Construction
By
–
@Insolentiae merci pour la mention 🙂 https://
insolentiae.com/vas-y-chat-gpt
-construit-cet-immeuble-on-te-regarde/
… -
HQ-SAM: High Quality Segment Anything Model
By
–
Segment Anything in High Quality
— AK (@_akhaliq) 5 juin 2023
paper page: https://t.co/IG4IAN4a4V
propose HQ-SAM, equipping SAM with the ability to accurately segment any object, while maintaining SAM's original promptable design, efficiency, and zero-shot generalizability. Our careful design reuses and… pic.twitter.com/6xG0FES7hvSegment Anything in High Quality paper page: https://
huggingface.co/papers/2306.01
567
… propose HQ-SAM, equipping SAM with the ability to accurately segment any object, while maintaining SAM's original promptable design, efficiency, and zero-shot generalizability. Our careful design reuses and -

MERT: Self-Supervised Acoustic Music Understanding Model
By
–
6/ MERT – an acoustic music understanding model with large-scale self-supervised training; it incorporates a superior combination of teacher models to outperform conventional speech and audio approaches.
-
MineCLIP Embeddings Open-Sourced on MineDojo
By
–
Our MineCLIP embeddings are open-sourced on the Minedojo website. Fire away! https://
github.com/MineDojo -

Minecraft Voyager: Embodied AI and Interactive Curriculum Learning
By
–
Thank you @erikbryn Minecraft Voyager is an excited step towards building embodied #AI and getting #GPT4 to interactively build a curriculum and skill library
-
MineDojo Advances Self-Supervised Behavioral Cloning Research
By
–
Great to see more work building on our MineDojo! It combines self-supervised behavioral cloning and best practices from ext-conditioned image generation. https://t.co/z4LDplrBd6
— Prof. Anima Anandkumar (@AnimaAnandkumar) 3 juin 2023Great to see more work building on our MineDojo! It combines self-supervised behavioral cloning and best practices from ext-conditioned image generation.