AI with Papers - Artificial Intelligence & Deep Learning
17K subscribers
162 photos
286 videos
14 files
1.49K links
All the AI with papers. Every day fresh updates about #DeepLearning #MachineLearning #LLM & #ComputerVision

Curated by Alessandro Ferrari | https://www.linkedin.com/in/visionarynet/

#AI #chatGPT
Download Telegram
This media is not supported in your browser
VIEW IN TELEGRAM
πŸ‘»Emerging Objs from MotionπŸ‘»

πŸ‘‰Motion boundaries provide a strong signal for object-level grouping and can be used to derive pseudo-instance supervision. Suitable for: mono-depth, 3D object detection, 3D occupancy, and end-to-end planning. Repo under Apache 2.0πŸ’™

πŸ‘‰Review https://lnkd.in/p/eezZrSJE
πŸ‘‰Paper https://arxiv.org/pdf/2609.04348
πŸ‘‰Project https://tj12342.github.io/object-concepts-from-motion/
πŸ‘‰Repo https://github.com/TJ12342/object-concepts-from-motion/tree/main
❀2πŸ‘2😍1
This media is not supported in your browser
VIEW IN TELEGRAM
πŸ”₯#AIwithPapers: we are 17,000+πŸ”₯

πŸ‘‰ Even though 100+ bots are trying to join the discussion chats every day, there are 17,000 of us! Almost all of us are still humans 🧟

😈 Invite -> https://t.me/AI_DeepLearning
❀25🍾14πŸ‘4
This media is not supported in your browser
VIEW IN TELEGRAM
πŸ€McByte++ tracking-by-detectionπŸ€

πŸ‘‰McByte++ is the newer extension of McByte that advances training-free sports MOT toward long-term ID tracking, while simultaneously improving efficiency and runtime performance. Repo under Apache 2.0πŸ’™

πŸ‘‰Review https://lnkd.in/p/e4-diVJS
πŸ‘‰Paper https://lnkd.in/e_Vxky-b
πŸ‘‰Repo https://lnkd.in/e8SeCYmk
❀8πŸ‘2πŸ”₯1πŸ’©1🍾1
This media is not supported in your browser
VIEW IN TELEGRAM
πŸ”₯πŸ”₯ Marigold V2 is out πŸ”₯πŸ”₯

πŸ‘‰Marigold V2 is out: depth, (impressive) see-through depth, surface normals, albedo, and other dense modalities. SOTA results. Repo under Apache 2.0πŸ’™

#AI #deeplearning #AIwithPapers

πŸ‘‰Review https://lnkd.in/p/eKM44yDQ
πŸ‘‰Paper https://arxiv.org/pdf/2609.08084
πŸ‘‰Repo https://github.com/huawei-bayerlab/marigold-v2
πŸ‘‰Project https://huggingface.co/spaces/huawei-bayerlab/marigold-v2-web
πŸ”₯9❀3πŸ‘3πŸ‘1
This media is not supported in your browser
VIEW IN TELEGRAM
🦺Efficient/Scalable Video Pretraining🦺

πŸ‘‰LeVJEPA1 (Yann Lecun) is the first video encoder trained under LeJEPA’s collapse-free objective, and evaluate it under frozen probing against video and image pretraining baselines retrained on identical data, in both epoch-matched and FLOP-matched regimes. Repo under MITπŸ’™

πŸ‘‰Review https://lnkd.in/p/eJQAm3AN
πŸ‘‰Paper https://lnkd.in/eCzzTiNH
πŸ‘‰Project https://levjepa.github.io/
πŸ‘‰Repo https://lnkd.in/etiF5CDj
❀8πŸ”₯5πŸ‘1
This media is not supported in your browser
VIEW IN TELEGRAM
πŸ”₯ RelateAnything is gold! πŸ”₯

πŸ‘‰RelateAnything is a 53M-parameter relation model that takes an image and a set of regions from any source and returns scored relations over a predicate vocabulary supplied at inference as a list of strings. Impressive results. Repo under Apache 2.0πŸ’™

πŸ‘‰Review https://lnkd.in/p/etAcdFM3
πŸ‘‰Paper https://arxiv.org/pdf/2609.12552
πŸ‘‰Repo https://github.com/Maelic/RelateAnything
πŸ‘‰Project https://maelic.github.io/RelateAnythingProject/
❀12πŸ”₯4πŸ‘3πŸ‘1🀯1
This media is not supported in your browser
VIEW IN TELEGRAM
πŸ‘‹ EventEgoHands++ is out! πŸ‘‹

πŸ‘‰EventEgoHands++ is a novel framework for event-based 3D hand mesh reconstruction from an egocentric viewpoint. 1M+ samples dataset! Code/Data releasedπŸ’™

πŸ‘‰Review https://lnkd.in/p/eTbPvXbW
πŸ‘‰Paper https://arxiv.org/pdf/2609.17189
πŸ‘‰Repo https://github.com/ryhara/EventEgoHandsV2
πŸ‘‰Project https://ryhara.github.io/EventEgoHandsV2/
❀3πŸ”₯2πŸ‘1
This media is not supported in your browser
VIEW IN TELEGRAM
πŸ’¦SOTA Splashing LiquidsπŸ’¦

πŸ‘‰SplashSplat reconstructs splashing liquids from real multi-view vide. Impose physical structure only where the observations can constrain it. Impressive results, SOTA. Code TBR under MITπŸ’™

πŸ‘‰Review https://lnkd.in/p/ejMTHcp7
πŸ‘‰Paper https://arxiv.org/pdf/2609.20818
πŸ‘‰Project niko-creater.github.io/splashsplat-web/
πŸ‘‰Repo https://github.com/Niko-creater/Splashsplat
πŸ‘4❀2πŸ”₯2πŸ‘1
+++ Breaking +++
❀3πŸ‘2😒2🀣1
This media is not supported in your browser
VIEW IN TELEGRAM
πŸ”₯Agentic Image-to-SceneπŸ”₯

πŸ‘‰HARMONY by UPenn is a hierarchical chain-of-thought framework that leverages both agentic reasoning and visual geometry foundation. Impressive 3D scenes. Repo TBAπŸ’™

πŸ‘‰Review https://lnkd.in/p/ep2hmRSp
πŸ‘‰Paper https://arxiv.org/pdf/2609.26793
πŸ‘‰Project https://cwchenwang.github.io/harmony/
πŸ‘‰Data https://huggingface.co/datasets/ShufanSun/harmony
πŸ”₯8❀2πŸ‘1
This media is not supported in your browser
VIEW IN TELEGRAM
🍿PanoSeg3R: SOTA 3D Segmentation🍿

πŸ‘‰PanoSeg3R is a novel feed-forward framework for 3D panoramic semantic segmentation. New SOTA. Code comingπŸ’™

πŸ‘‰Review https://lnkd.in/p/eKCKWv3g
πŸ‘‰Paper https://arxiv.org/pdf/2609.22687
πŸ‘‰Project https://harryyoon777.github.io/PanoSeg3R/#
πŸ‘‰Repo TBA
❀4πŸ‘1πŸ”₯1πŸ‘1
This media is not supported in your browser
VIEW IN TELEGRAM
🩻Universal X-ray Segmentation🩻

πŸ‘‰FleXray: universal anatomical segmentation across the entire body in clinical X-rays. Built on a scalable, physics-based generative X-ray data engine. Repo under MITπŸ’™

πŸ‘‰Review https://lnkd.in/p/e9MUk_eq
πŸ‘‰Paper https://arxiv.org/pdf/2609.26756
πŸ‘‰Project https://flexray.csail.mit.edu/
πŸ‘‰Repo https://github.com/VictorButoi/FleXray
πŸ‘4❀3πŸ”₯2πŸ‘2
This media is not supported in your browser
VIEW IN TELEGRAM
🦴3D Foundational Radiology🦴

πŸ‘‰nnFoundation: 3D radiological foundation models designed for transferable representation learning across heterogeneous tasks/datasets. Models releasedπŸ’™

πŸ‘‰Review https://lnkd.in/p/eNajRGBi
πŸ‘‰Paper https://arxiv.org/pdf/2609.26924
πŸ‘‰Models https://huggingface.co/collections/MIC-DKFZ/nnfoundation
❀4πŸ‘2πŸ‘1🀩1