Looks like Apple is very interested in JEPA! What if your AI could “read” an image’s caption to solve visual puzzles? Apple researchers present TC-JEPA: a new self-supervised method that uses image captions to guide masked patch predictions. By conditioning on text, the model
Apple Researchers Introduce TC-JEPA for Vision-Language Learning
By
–
