2nd Edition, 746 pages, massive! Modern Computer Vision with #PyTorch #DeepLearning — from practical fundamentals to advanced applications and Generative AI: http://
amzn.to/3xAkB7X v/ @PacktDataML ——
#DataScience #MachineLearning #ML #GenAI #DataScientist
——
𝓚𝓮𝔂
MULTIMODAL AI
-

Modern Computer Vision with PyTorch — 2nd Edition Book
By
–
-
Seedance 2.0 also enables regular video editing
By
–
Not to mention that ”just” regular video editing is also possible with seedance 2.0https://t.co/3TwnCZntcZ
— Linus ✦ Ekenstam (@LinusEkenstam) 15 février 2026Not to mention that ”just” regular video editing is also possible with seedance 2.0
-
Seedance 2.0 generates clips and removes cameraman seamlessly
By
–
Seedance 2.0 does not stop to amaze.
— Linus ✦ Ekenstam (@LinusEkenstam) 15 février 2026
Beyond generating clips of anyone doing anything, it also single handily becomes your most capable video editor.
Here it removes the camera man, but look closely, can you see it?
pic.twitter.com/KOxNlf6HSQSeedance 2.0 does not stop to amaze. Beyond generating clips of anyone doing anything, it also single handily becomes your most capable video editor. Here it removes the camera man, but look closely, can you see it?
-

Build a Text-to-Image Generator from Scratch
By
–
Hot New Release >> Build a Text-to-Image Generator (from Scratch), with transformers and diffusions: http://
amzn.to/3MFbyK4 by @mark_h_liu v/ @ManningBooks -

Video AI and World Models: HKUST Research on Physical Understanding
By
–
Can video generation AI actually understand the physical world, or is it just a digital illusion? Researchers from HKUST (GZ), Tongji University, and Kuaishou Technology present a new mechanistic framework to bridge the gap between video generation and true world models! They
-
Four Major AI Trends for 2026: 3D, Robotics, Agentic Orchestration
By
–
There are four major trends that I see for 2026 in AI: – 3D generation – Robotics, particularly data-to-actions systems (VLM, LAM, world models) – Agentic management and orchestration (and I mean real orchestrations, not bullshit stuff from PDF sellers with n8n) – The
-
Download or clone repo for open source video and audio model
By
–
You can download now, or clone the repo, explore the bleeding edge of Open Source and Audio model from @ltx_model
-
Workflow: one photo to multiple angles cinematic video
By
–
Workflows that you might be used to with closed source models, like nanobanana pro etc can now be achieved through mixing different providers.
— Linus ✦ Ekenstam (@LinusEkenstam) 14 février 2026
Here us an example of:
One photo → multiple camera angles → cinematic video
video from: @FloyoAI pic.twitter.com/0pbcZIyDP4Workflows that you might be used to with closed source models, like nanobanana pro etc can now be achieved through mixing different providers. Here us an example of: One photo → multiple camera angles → cinematic video video from: @FloyoAI
-
Combining Seedance 2.0 with LTX-2 dubbing for borderless content
By
–
You can also combine things like Seedance 2.0 with LTX-2 dubbing.
— Linus ✦ Ekenstam (@LinusEkenstam) 14 février 2026
Then you get something like this, content knows no borders.
from @antonch44379293 pic.twitter.com/l0Aojr0zZEYou can also combine things like Seedance 2.0 with LTX-2 dubbing. Then you get something like this, content knows no borders. from @antonch44379293
-
Google Launches Gemini 3 Deep Think for Complex Science
By
–
We launches! Here’s a recap of what went out this week: — An upgrade to Gemini 3 Deep Think that solves complex modern science and engineering challenges. Scientists, researchers, and enterprises can express interest in Deep Think via our early access program and Google AI