PDF2Audio
— AK (@_akhaliq) 23 septembre 2024
Convert PDFs into an audio podcast, lecture, summary and others pic.twitter.com/zSPeewxvDE
PDF2Audio Convert PDFs into an audio podcast, lecture, summary and others

By
–
PDF2Audio
— AK (@_akhaliq) 23 septembre 2024
Convert PDFs into an audio podcast, lecture, summary and others pic.twitter.com/zSPeewxvDE
PDF2Audio Convert PDFs into an audio podcast, lecture, summary and others

By
–
🎶 OpenMusic
— AK (@_akhaliq) 23 septembre 2024
Diffusion That Plays Music 🎧 🎹
github: https://t.co/l3QK4VLqvh
demo: https://t.co/ZnDaiI20Ow
model: https://t.co/DvTxqc49CU pic.twitter.com/s6C9LRy0L6
OpenMusic Diffusion That Plays Music github: https://
github.com/ivcylc/qa-mdt
demo: https://
huggingface.co/spaces/jadecho
ghari/OpenMusic
…
model: https://
huggingface.co/jadechoghari/o
penmusic
…

By
–
NASA and IBM present Prithvi WxC Foundation Model for Weather and Climate discuss: https://
huggingface.co/papers/2409.13
598
… Triggered by the realization that AI emulators can rival the performance of traditional numerical weather prediction models running on HPC systems, there is now an

By
–
Meta presents Imagine yourself Tuning-Free Personalized Image Generation paper page: https://
huggingface.co/papers/2409.13
346
… Diffusion models have demonstrated remarkable efficacy across various image-to-image tasks. In this research, we introduce Imagine yourself, a state-of-the-art model

By
–
Colorful Diffuse Intrinsic Image Decomposition in the Wild discuss: https://
huggingface.co/papers/2409.13
690
… Intrinsic image decomposition aims to separate the surface reflectance and the effects from the illumination given a single photograph. Due to the complexity of the problem, most prior

By
–
V^3: Viewing Volumetric Videos on Mobiles via Streamable 2D Dynamic Gaussians
— AK (@_akhaliq) 23 septembre 2024
discuss: https://t.co/uSKBKDqleF
Experiencing high-fidelity volumetric video as seamlessly as 2D videos is a long-held dream. However, current dynamic 3DGS methods, despite their high rendering… pic.twitter.com/C84AesHSwP
V^3: Viewing Volumetric Videos on Mobiles via Streamable 2D Dynamic Gaussians discuss: https://
huggingface.co/papers/2409.13
648
… Experiencing high-fidelity volumetric video as seamlessly as 2D videos is a long-held dream. However, current dynamic 3DGS methods, despite their high rendering

By
–
Portrait Video Editing Empowered by Multimodal Generative Priors
— AK (@_akhaliq) 23 septembre 2024
discuss: https://t.co/bTsaTK0THF
We introduce PortraitGen, a powerful portrait video editing method that achieves consistent and expressive stylization with multimodal prompts. Traditional portrait video editing… pic.twitter.com/fsliAEnwoL
Portrait Editing Empowered by Multimodal Generative Priors discuss: https://
huggingface.co/papers/2409.13
591
… We introduce PortraitGen, a powerful portrait video editing method that achieves consistent and expressive stylization with multimodal prompts. Traditional portrait video editing

By
–
MuCodec Ultra Low-Bitrate Music Codec discuss: https://
huggingface.co/papers/2409.13
216
… Music codecs are a vital aspect of audio codec research, and ultra low-bitrate compression holds significant importance for music transmission and generation. Due to the complexity of music backgrounds and

By
–
Hacker News for AI research papers is up on @github built with o1-mini and sonnet 3.5 in @cursor_ai github: https://
github.com/AK391/dailypap
ersHN
…
app: https://
huggingface.co/spaces/akhaliq
/dailypapershackernews
…
By
–
there is an api: https://
huggingface.co/api/daily_pape
rs
… more info: