16 GB RAM. No cloud subscription. Which local AI model actually fits?
How AI Helps built a free Telegram model picker. Choose your task, RAM or VRAM, language, runtime, and commercial-use requirement.
Then compare a shortlist by memory, license, sources, download options, and launch commands when available.
Join How AI Helps and open the pinned model-picker guide
How AI Helps built a free Telegram model picker. Choose your task, RAM or VRAM, language, runtime, and commercial-use requirement.
Then compare a shortlist by memory, license, sources, download options, and launch commands when available.
Join How AI Helps and open the pinned model-picker guide
โค10๐2๐ฏ2
Top YouTube Channels to Master Tech Skills ๐
1. SQL ๐ป
๐ youtube.com/@joeyblue1
2. Excel ๐
๐ youtube.com/@excelisfun
3. Statistics ๐
๐ youtube.com/@statquest
4. Math ๐งฎ
๐ youtube.com/results?searchโฆ
5. Python ๐
๐ youtube.com/@BroCodez
6. Data Analysis ๐
๐ youtube.com/@AlexTheAnalyst
7. Machine Learning ๐ค
๐ youtube.com/@campusx-officโฆ
8. Deep Learning ๐ง
๐ youtube.com/@deeplizard
9. Java โ
๐ youtube.com/@Telusko
10. Big Data ๐ฆ
๐ youtube.com/@thedatatech
11. Data Engineering โ๏ธ
๐ youtube.com/@dataengineeriโฆ
12. NLP (Natural Language Processing) ๐ฃ๏ธ
๐ youtube.com/@codebasics
13. Computer Vision & AI ๐๏ธ
๐ youtube.com/@murtazasworksโฆ
14. Generative AI โจ
๐ youtube.com/@sunnysavita10
15. University-Level Courses ๐
๐ youtube.com/@stanfordonline
๐ youtube.com/@mitocw
16. All-in-One Learning ๐
๐ youtube.com/@freecodecamp
#TechSkills #YouTube #DataScience #Programming #MachineLearning #LearnTech
โจ Join Best TG Channels https://t.me/addlist/0f6vfFbEMdAwODBk
โญ๏ธ Join Our WhatsApp Channel https://whatsapp.com/channel/0029VaC7Weq29753hpcggW2A
1. SQL ๐ป
๐ youtube.com/@joeyblue1
2. Excel ๐
๐ youtube.com/@excelisfun
3. Statistics ๐
๐ youtube.com/@statquest
4. Math ๐งฎ
๐ youtube.com/results?searchโฆ
5. Python ๐
๐ youtube.com/@BroCodez
6. Data Analysis ๐
๐ youtube.com/@AlexTheAnalyst
7. Machine Learning ๐ค
๐ youtube.com/@campusx-officโฆ
8. Deep Learning ๐ง
๐ youtube.com/@deeplizard
9. Java โ
๐ youtube.com/@Telusko
10. Big Data ๐ฆ
๐ youtube.com/@thedatatech
11. Data Engineering โ๏ธ
๐ youtube.com/@dataengineeriโฆ
12. NLP (Natural Language Processing) ๐ฃ๏ธ
๐ youtube.com/@codebasics
13. Computer Vision & AI ๐๏ธ
๐ youtube.com/@murtazasworksโฆ
14. Generative AI โจ
๐ youtube.com/@sunnysavita10
15. University-Level Courses ๐
๐ youtube.com/@stanfordonline
๐ youtube.com/@mitocw
16. All-in-One Learning ๐
๐ youtube.com/@freecodecamp
#TechSkills #YouTube #DataScience #Programming #MachineLearning #LearnTech
โจ Join Best TG Channels https://t.me/addlist/0f6vfFbEMdAwODBk
โญ๏ธ Join Our WhatsApp Channel https://whatsapp.com/channel/0029VaC7Weq29753hpcggW2A
โค13๐ฅ1
Forwarded from Python Courses & Resources
Free Generative AI Courses
Generative AI Full Course: Gemini Pro, OpenAI, Llama, Langchain, Pinecone, Vector Databases & More
๐ Free Video Course
โฐ Duration: 30 hrs
๐โโ๏ธ Self Paced
๐ Difficulty: Beginner to Intermediate
๐จโ๐ซ Instructors: Krish Naik, Sunny Savita & Boktiar Ahmed Bappy via freeCodeCamp
๐ Course Link
5-Day Gen AI Intensive Course with Google
๐ Free Video + Hands-On Codelabs
โฐ Duration: 5-day structure
๐โโ๏ธ Self Paced
๐ Difficulty: Beginner to Intermediate
๐จโ๐ซ Created by: Google & Kaggle
๐ Course Link
Free GenAI 65-Hour Bootcamp
๐ Free Video Course
โฐ Duration: 65 hrs
๐โโ๏ธ Self Paced
๐ Difficulty: Beginner to Intermediate
๐จโ๐ซ Instructor: Andrew Brown (ExamPro) via freeCodeCamp
๐ Course Link
Generative AI for Beginners
๐ Free Video Course
โฐ Duration: Multi-hour
๐โโ๏ธ Self Paced
๐ Difficulty: Beginner
๐จโ๐ซ Created by: Great Learning Academy
๐ Course Link
Introduction to Generative AI
๐ Free Video Course
โฐ Duration: 45 min
๐โโ๏ธ Self Paced
๐ Difficulty: Beginner
๐จโ๐ซ Created by: Google Skills
๐ Course Link
Generative AI for Beginners
๐ Text Course
โฐ Duration: 21 lessons
๐โโ๏ธ Self Paced
๐ Difficulty: Beginner
๐จโ๐ซ Created by: Microsoft Cloud Advocates
๐ Course Link
AI Capabilities and Limitations
๐ Free Video Course
โฐ Duration: Self-paced
๐โโ๏ธ Self Paced
๐ Difficulty: Beginner
๐จโ๐ซ Created by: Anthropic Academy
๐ Course Link
Generative AI for Beginners
๐ Free Video Course
โฐ Duration: 4 hrs
๐โโ๏ธ Self Paced
๐ Difficulty: Beginner
๐จโ๐ซ Created by: Simplilearn
๐ Course Link
Reading Materials
๐ Prompt Engineering Guide
๐ Awesome Generative AI (Curated Resource List)
๐ Generative AI: A Beginner's Guide
๐ Understanding Generative AI Capabilities
๐Stanford HAI: 2025 AI Index Report
Generative AI Full Course: Gemini Pro, OpenAI, Llama, Langchain, Pinecone, Vector Databases & More
๐ Free Video Course
โฐ Duration: 30 hrs
๐โโ๏ธ Self Paced
๐ Difficulty: Beginner to Intermediate
๐จโ๐ซ Instructors: Krish Naik, Sunny Savita & Boktiar Ahmed Bappy via freeCodeCamp
๐ Course Link
5-Day Gen AI Intensive Course with Google
๐ Free Video + Hands-On Codelabs
โฐ Duration: 5-day structure
๐โโ๏ธ Self Paced
๐ Difficulty: Beginner to Intermediate
๐จโ๐ซ Created by: Google & Kaggle
๐ Course Link
Free GenAI 65-Hour Bootcamp
๐ Free Video Course
โฐ Duration: 65 hrs
๐โโ๏ธ Self Paced
๐ Difficulty: Beginner to Intermediate
๐จโ๐ซ Instructor: Andrew Brown (ExamPro) via freeCodeCamp
๐ Course Link
Generative AI for Beginners
๐ Free Video Course
โฐ Duration: Multi-hour
๐โโ๏ธ Self Paced
๐ Difficulty: Beginner
๐จโ๐ซ Created by: Great Learning Academy
๐ Course Link
Introduction to Generative AI
๐ Free Video Course
โฐ Duration: 45 min
๐โโ๏ธ Self Paced
๐ Difficulty: Beginner
๐จโ๐ซ Created by: Google Skills
๐ Course Link
Generative AI for Beginners
๐ Text Course
โฐ Duration: 21 lessons
๐โโ๏ธ Self Paced
๐ Difficulty: Beginner
๐จโ๐ซ Created by: Microsoft Cloud Advocates
๐ Course Link
AI Capabilities and Limitations
๐ Free Video Course
โฐ Duration: Self-paced
๐โโ๏ธ Self Paced
๐ Difficulty: Beginner
๐จโ๐ซ Created by: Anthropic Academy
๐ Course Link
Generative AI for Beginners
๐ Free Video Course
โฐ Duration: 4 hrs
๐โโ๏ธ Self Paced
๐ Difficulty: Beginner
๐จโ๐ซ Created by: Simplilearn
๐ Course Link
Reading Materials
๐ Prompt Engineering Guide
๐ Awesome Generative AI (Curated Resource List)
๐ Generative AI: A Beginner's Guide
๐ Understanding Generative AI Capabilities
๐Stanford HAI: 2025 AI Index Report
YouTube
Generative AI Full Course โ Gemini Pro, OpenAI, Llama, Langchain, Pinecone, Vector Databases & More
Learn about generative models and different frameworks, investigating the production of text and visual material produced by artificial intelligence. This course was originally recorded live.
Instructors: Krish Naik, Sunny Savita, and Boktiar Ahmed Bappy.โฆ
Instructors: Krish Naik, Sunny Savita, and Boktiar Ahmed Bappy.โฆ
โค9
๐ Stop Maintaining Scrapers. Start Shipping Products.
Build AI products, not scraping infrastructure.
CoreClaw provides ready-to-use Workers & APIs for 1000+ websites โ including Google Maps, Instagram, Facebook, YouTube, Amazon, Tiktok and Google Search Scraper.
โ๏ธ No infrastructure
โ๏ธ No proxy management
โ๏ธ No scraper maintenance
โ๏ธ JSON / CSV / REST API
๐ Create a free account. Get free credits. Explore every Worker.
๐ https://coreclaw.com
Build AI products, not scraping infrastructure.
CoreClaw provides ready-to-use Workers & APIs for 1000+ websites โ including Google Maps, Instagram, Facebook, YouTube, Amazon, Tiktok and Google Search Scraper.
โ๏ธ No infrastructure
โ๏ธ No proxy management
โ๏ธ No scraper maintenance
โ๏ธ JSON / CSV / REST API
๐ Create a free account. Get free credits. Explore every Worker.
๐ https://coreclaw.com
โค4
Your AI helper right in your messenger โ in 5 minutes, free
Amplify (UK) plugs an AI agent straight into your Telegram, WhatsApp, Slack, WeChat, or Discord. Not just a GPT chat โ an assistant that reaches into the real world.
Handles it all: emails, reminders, spreadsheets, Telegram-channel digests, image and video generation, PDFs, Google Drive, Notion. Send it voice notes on the go โ it gets everything.
Pricing: $10/mo + pay-as-you-go for the AI model, all costs transparent and tracked. Already have OpenAI subscription? Link it and skip paying for the model.
๐ Promo code
https://getamplify.team/
Amplify (UK) plugs an AI agent straight into your Telegram, WhatsApp, Slack, WeChat, or Discord. Not just a GPT chat โ an assistant that reaches into the real world.
Handles it all: emails, reminders, spreadsheets, Telegram-channel digests, image and video generation, PDFs, Google Drive, Notion. Send it voice notes on the go โ it gets everything.
Pricing: $10/mo + pay-as-you-go for the AI model, all costs transparent and tracked. Already have OpenAI subscription? Link it and skip paying for the model.
๐ Promo code
CODEPROGRAMMER2 โ 2 months free + $10 credit. Bring someone in โ another month free.https://getamplify.team/
โค3๐1๐ฅ1
This media is not supported in your browser
VIEW IN TELEGRAM
๐ A useful training tool for Data Scientists ๐
๐ซก Real-world tasks from IT companies;
๐ซก SQL practice;
๐ซก Python tasks;
๐ซก Preparation for Data Science interviews.
โ Link to the training tool
https://www.stratascratch.com/
๐ท #DataScience #SQL #Python #InterviewPrep #TechTraining #DataAnalyst
โจ Join Best TG Channels https://t.me/addlist/0f6vfFbEMdAwODBk
โญ๏ธ Join Our WhatsApp Channel https://whatsapp.com/channel/0029VaC7Weq29753hpcggW2A
๐ซก Real-world tasks from IT companies;
๐ซก SQL practice;
๐ซก Python tasks;
๐ซก Preparation for Data Science interviews.
โ Link to the training tool
https://www.stratascratch.com/
๐ท #DataScience #SQL #Python #InterviewPrep #TechTraining #DataAnalyst
โจ Join Best TG Channels https://t.me/addlist/0f6vfFbEMdAwODBk
โญ๏ธ Join Our WhatsApp Channel https://whatsapp.com/channel/0029VaC7Weq29753hpcggW2A
โค6
Forwarded from Machine Learning
This media is not supported in your browser
VIEW IN TELEGRAM
A Powerful Alternative to Pandas ๐
This is an optimized replacement for Pandas that can significantly speed up data processing without requiring major changes to your code. โ๏ธ
To get started, simply replace a single import:
Performance Benchmarks demonstrate speed improvements in various use cases. ๐
More: https://colab.research.google.com/drive/1UIokuJ4cytoiVSabRDqcziDXOan8bVua?usp=sharing
#Pandas #Python #DataScience #Performance #Fireducks #BigData
โจ Join Best TG Channels https://t.me/addlist/0f6vfFbEMdAwODBk
โญ๏ธ Join Our WhatsApp Channel https://whatsapp.com/channel/0029VaC7Weq29753hpcggW2A
This is an optimized replacement for Pandas that can significantly speed up data processing without requiring major changes to your code. โ๏ธ
To get started, simply replace a single import:
import fireducks.pandas as pd
Performance Benchmarks demonstrate speed improvements in various use cases. ๐
More: https://colab.research.google.com/drive/1UIokuJ4cytoiVSabRDqcziDXOan8bVua?usp=sharing
#Pandas #Python #DataScience #Performance #Fireducks #BigData
โจ Join Best TG Channels https://t.me/addlist/0f6vfFbEMdAwODBk
โญ๏ธ Join Our WhatsApp Channel https://whatsapp.com/channel/0029VaC7Weq29753hpcggW2A
โค8๐1
Please open Telegram to view this post
VIEW IN TELEGRAM
โค3
Forwarded from Machine Learning with Python
This channels is for Programmers, Coders, Software Engineers.
0๏ธโฃ Python
1๏ธโฃ Data Science
2๏ธโฃ Machine Learning
3๏ธโฃ Data Visualization
4๏ธโฃ Artificial Intelligence
5๏ธโฃ Data Analysis
6๏ธโฃ Statistics
7๏ธโฃ Deep Learning
8๏ธโฃ programming Languages
โ
https://t.me/addlist/8_rRW2scgfRhOTc0
โ
https://t.me/Codeprogrammer
Please open Telegram to view this post
VIEW IN TELEGRAM
โค3
A collection of resources on MLOps for those who want to understand how machine learning systems are brought to production. ๐๐ค
https://github.com/visenger/awesome-mlops
#MLOps #MachineLearning #DevOps #AI #DataScience #TechResources
โจ Join Best TG Channels https://t.me/addlist/0f6vfFbEMdAwODBk
โญ๏ธ Join Our WhatsApp Channel https://whatsapp.com/channel/0029VaC7Weq29753hpcggW2A
https://github.com/visenger/awesome-mlops
#MLOps #MachineLearning #DevOps #AI #DataScience #TechResources
โจ Join Best TG Channels https://t.me/addlist/0f6vfFbEMdAwODBk
โญ๏ธ Join Our WhatsApp Channel https://whatsapp.com/channel/0029VaC7Weq29753hpcggW2A
โค6
This media is not supported in your browser
VIEW IN TELEGRAM
U-Net by hand โ๏ธ ~ 17 steps walkthrough below
I consider U-Net as a key milestone in deep learning, the first image-to-image model that really worked!
It came out of medical imaging, an unusual place, not from NeurIPS or CVPR or ACL.
Now it is the backbone of diffusion models, which you see in almost all modern image generation models.
I drew the network as a C so the matrix multiplication flows naturally down.
Tilt your head to the right and it is a U again. ๐คฃ
Goal: push a 3 x 16 image down to a 2 x 4 bottleneck and back out again, filling in every cell yourself.
= 1. Given =
An image of three channels, R, G and B, sixteen pixels wide, and every kernel the network will use.
= 2. Convolution 1 =
Let us slide the first kernel over the image. Each output is one multiply-and-add over a 2 x 3 window, and the result is the green feature map.
= 3. Find the maxima =
We circle the largest value in each 1 x 2 window. Circling first is worth the extra step: it is the pooling decision, made before anything is written down.
= 4. Max pool 1 =
Let us copy those maxima down. Sixteen columns become eight, and half the detail is gone for good.
= 5. Convolution 2 =
We convolve again with the second kernel, deeper into the contracting path. The feature map is blue now.
= 6. Find the maxima again =
Same move as step 3, on the blue map.
= 7. Max pool 2 =
Eight columns become four.
= 8. The bottleneck =
Let us convolve once more. This is the bottom of the U, a 2 x 4 block that is everything the network kept.
= 9. Spread it out =
We start back up. The transposed convolution writes each bottleneck value into a wider grid, leaving gaps between them.
= 10. Transposed convolution 1 =
Let us fill those gaps by convolving over the spread-out grid. Four columns become eight.
= 11. The first skip =
We copy the encoder's matching row straight across. This is the skip connection, and it is the whole reason a U-Net can recover detail that pooling threw away.
= 12. Convolution with the skip =
Let us convolve the upsampled features together with the copied ones.
= 13. Spread it out again =
Same as step 9, one level up.
= 14. Transposed convolution 2 =
Eight columns become sixteen, back to the width we started at.
= 15. The second skip =
The encoder's first feature map comes across, the one made before any pooling happened.
= 16. Convolution and ReLU =
We convolve, then cross out every negative and set it to zero.
= 17. Output convolution =
Let us apply the last kernel. Out comes R', G' and B', an image the same size as the one we started with.
The outputs:
Congrats! You just calculated a U-Net by hand.
๐พ Save this post!
#UNet #DeepLearning #AI #NeuralNetworks #ComputerVision #MachineLearning
โจ Join Best TG Channels https://t.me/addlist/0f6vfFbEMdAwODBk
โญ๏ธ Join Our WhatsApp Channel https://whatsapp.com/channel/0029VaC7Weq29753hpcggW2A
I consider U-Net as a key milestone in deep learning, the first image-to-image model that really worked!
It came out of medical imaging, an unusual place, not from NeurIPS or CVPR or ACL.
Now it is the backbone of diffusion models, which you see in almost all modern image generation models.
I drew the network as a C so the matrix multiplication flows naturally down.
Tilt your head to the right and it is a U again. ๐คฃ
Goal: push a 3 x 16 image down to a 2 x 4 bottleneck and back out again, filling in every cell yourself.
= 1. Given =
An image of three channels, R, G and B, sixteen pixels wide, and every kernel the network will use.
= 2. Convolution 1 =
Let us slide the first kernel over the image. Each output is one multiply-and-add over a 2 x 3 window, and the result is the green feature map.
= 3. Find the maxima =
We circle the largest value in each 1 x 2 window. Circling first is worth the extra step: it is the pooling decision, made before anything is written down.
= 4. Max pool 1 =
Let us copy those maxima down. Sixteen columns become eight, and half the detail is gone for good.
= 5. Convolution 2 =
We convolve again with the second kernel, deeper into the contracting path. The feature map is blue now.
= 6. Find the maxima again =
Same move as step 3, on the blue map.
= 7. Max pool 2 =
Eight columns become four.
= 8. The bottleneck =
Let us convolve once more. This is the bottom of the U, a 2 x 4 block that is everything the network kept.
= 9. Spread it out =
We start back up. The transposed convolution writes each bottleneck value into a wider grid, leaving gaps between them.
= 10. Transposed convolution 1 =
Let us fill those gaps by convolving over the spread-out grid. Four columns become eight.
= 11. The first skip =
We copy the encoder's matching row straight across. This is the skip connection, and it is the whole reason a U-Net can recover detail that pooling threw away.
= 12. Convolution with the skip =
Let us convolve the upsampled features together with the copied ones.
= 13. Spread it out again =
Same as step 9, one level up.
= 14. Transposed convolution 2 =
Eight columns become sixteen, back to the width we started at.
= 15. The second skip =
The encoder's first feature map comes across, the one made before any pooling happened.
= 16. Convolution and ReLU =
We convolve, then cross out every negative and set it to zero.
= 17. Output convolution =
Let us apply the last kernel. Out comes R', G' and B', an image the same size as the one we started with.
The outputs:
R' = [3, 0, 7, 0, 7, 0, 17, 0, 3, 0, 9, 0, 2, 0, 6, 0]
G' = [1, 20, 1, 10, 1, 12, 1, 19, 2, 5, 1, 11, 1, 3, 1, 7]
B' = [4, 20, 8, 10, 8, 12, 18, 19, 5, 5, 10, 11, 3, 3, 7, 7]
Congrats! You just calculated a U-Net by hand.
๐พ Save this post!
#UNet #DeepLearning #AI #NeuralNetworks #ComputerVision #MachineLearning
โจ Join Best TG Channels https://t.me/addlist/0f6vfFbEMdAwODBk
โญ๏ธ Join Our WhatsApp Channel https://whatsapp.com/channel/0029VaC7Weq29753hpcggW2A
โค8๐2
Forwarded from Machine Learning
๐ Over 300 real-world case studies of ML systems from top companies. ๐ค
We found a repository that collects genuine ML engineering experience โ not theory from textbooks, but real stories of implementing models in production. ๐
Inside, you'll find case studies from Uber, Netflix, Google, and other companies: how they built the architecture, what problems arose, where the systems failed, and what solutions helped them recover. ๐๏ธ
โ Link to GitHub
https://github.com/Engineer1999/A-Curated-List-of-ML-System-Design-Case-Studies
#MachineLearning #MLCaseStudies #DataScience #Engineering #Uber #Netflix
โจ Join Best TG Channels https://t.me/addlist/0f6vfFbEMdAwODBk
โญ๏ธ Join Our WhatsApp Channel https://whatsapp.com/channel/0029VaC7Weq29753hpcggW2A
We found a repository that collects genuine ML engineering experience โ not theory from textbooks, but real stories of implementing models in production. ๐
Inside, you'll find case studies from Uber, Netflix, Google, and other companies: how they built the architecture, what problems arose, where the systems failed, and what solutions helped them recover. ๐๏ธ
โ Link to GitHub
https://github.com/Engineer1999/A-Curated-List-of-ML-System-Design-Case-Studies
#MachineLearning #MLCaseStudies #DataScience #Engineering #Uber #Netflix
โจ Join Best TG Channels https://t.me/addlist/0f6vfFbEMdAwODBk
โญ๏ธ Join Our WhatsApp Channel https://whatsapp.com/channel/0029VaC7Weq29753hpcggW2A
โค6
Media is too big
VIEW IN TELEGRAM
Google just released a free 2-hour course on full Graph engineering: 1 prompt โ 100 agents โ loops โ graphs from 0% to 100%: ๐คโ๏ธ
10% โ 17:44 - build your first agent ๐
30% โ 39:30 - Loop engineering: iterate, check, break ๐
60% โ 1:12:38 - Graph engineering ๐ธ๏ธ
75% โ 1:34:26 - agents that throttle themselves โก
100% โ 1:55:05 - full graph for multi-agentic systems ๐๏ธ
everyone builds one agent and calls it done - this is the full system where agents wire themselves into a graph.
watch the course, build the graph - then read the full architecture below โ
More: https://telegra.ph/Graph-Engineering-build-1000-agent-loops-in-one-window-from-one-prompt-full-5-step-course-08-02
#GraphEngineering #AI #Agents #Graphs #Tech #Coding
โจ Join Best TG Channels https://t.me/addlist/0f6vfFbEMdAwODBk
โญ๏ธ Join Our WhatsApp Channel https://whatsapp.com/channel/0029VaC7Weq29753hpcggW2A
10% โ 17:44 - build your first agent ๐
30% โ 39:30 - Loop engineering: iterate, check, break ๐
60% โ 1:12:38 - Graph engineering ๐ธ๏ธ
75% โ 1:34:26 - agents that throttle themselves โก
100% โ 1:55:05 - full graph for multi-agentic systems ๐๏ธ
everyone builds one agent and calls it done - this is the full system where agents wire themselves into a graph.
watch the course, build the graph - then read the full architecture below โ
More: https://telegra.ph/Graph-Engineering-build-1000-agent-loops-in-one-window-from-one-prompt-full-5-step-course-08-02
#GraphEngineering #AI #Agents #Graphs #Tech #Coding
โจ Join Best TG Channels https://t.me/addlist/0f6vfFbEMdAwODBk
โญ๏ธ Join Our WhatsApp Channel https://whatsapp.com/channel/0029VaC7Weq29753hpcggW2A
โค4๐1
Forwarded from Machine Learning
This media is not supported in your browser
VIEW IN TELEGRAM
Attention Heatmap vs Token Pruning ๐โ๏ธ
๐ More: https://www.overshoot.ai/blogs/an-introduction-to-token-pruning-for-vlms
#AI #MachineLearning #TokenPruning #DeepLearning #TechNews #VLM
โจ Join Best TG Channels https://t.me/addlist/0f6vfFbEMdAwODBk
โญ๏ธ Join Our WhatsApp Channel https://whatsapp.com/channel/0029VaC7Weq29753hpcggW2A
๐ More: https://www.overshoot.ai/blogs/an-introduction-to-token-pruning-for-vlms
#AI #MachineLearning #TokenPruning #DeepLearning #TechNews #VLM
โจ Join Best TG Channels https://t.me/addlist/0f6vfFbEMdAwODBk
โญ๏ธ Join Our WhatsApp Channel https://whatsapp.com/channel/0029VaC7Weq29753hpcggW2A
โค4๐4
This media is not supported in your browser
VIEW IN TELEGRAM
Generative Adversarial Network (GAN) by hand โ๏ธ ~ 9 steps walkthrough below
The Gen in GenAI came from this landmark paper by Ian Goodfellow et al., 12 years ago.
The paper showed that a neural network can not only classify but also turn upside down to generate realistic looking images.
The secret? We pit two of them against each other: a Generator turns noise into fake data, and a Discriminator learns to tell fake from real, pushing the Generator to keep doing better.
One runs upside down, the other right way up.
I drew and calculated one entirely by hand.
Goal: generate realistic 4D data out of 2D noise, filling in every cell yourself.
= 1. Given =
Four noise vectors in 2D, and four real data vectors in 4D.
= 2. Generator, first layer =
Let us multiply the noise by weights and biases to get new features.
= 3. ReLU =
We apply the activation, and -1 and -2 are crossed out and set to 0.
= 4. Generator, second layer =
Let us multiply again. ReLU applies here too, but every value is already positive, so nothing changes. What comes out is the fake data F, made by a two-layer generator out of nothing but noise.
= 5. Discriminator, first layer =
We feed it both, the four fakes and the four real vectors, through the same weights. It never learns which is which from the layout, only from the numbers.
= 6. Discriminator, second layer =
Let us reduce each data vector to a single feature Z. Eight vectors in, eight numbers out.
= 7. Sigmoid =
We turn each Z into a probability Y. A 1 means the discriminator is certain the data is real, a 0 means certain it is fake.
= 8. Training the Discriminator =
Let us take the gradients as Y minus YD, where YD is what the discriminator should have said: 0 for the four fakes, 1 for the four real. Why so simple? Because pairing sigmoid with binary cross entropy loss makes the math collapse to exactly this subtraction. Its loss uses both halves of the page.
= 9. Training the Generator =
We do it again, as Y minus YG, and YG is [1, 1, 1, 1]: the generator wants the discriminator to call every fake real. Same predictions, different target, opposite goal. Its loss uses only the fakes.
The outputs:
Fake data F = [1, 2, 3, 1], [1, 1, 2, 1], [2, 2, 4, 2], [1, 0, 1, 1]
Predictions on fakes = [.7, .5, .9, .3]
Predictions on real = [.7, .9, .9, 1]
Discriminator gradients = [.7, .5, .9, .3] and [-.3, -.1, -.1, 0]
Generator gradients = [-.3, -.5, -.1, -.7]
The takeaway: the adversarial part is one subtraction done twice. The same eight predictions, scored against two opposite targets, send one set of gradients back through the blue weights and another back through the green ones.
The Gen in GenAI came from this landmark paper by Ian Goodfellow et al., 12 years ago.
The paper showed that a neural network can not only classify but also turn upside down to generate realistic looking images.
The secret? We pit two of them against each other: a Generator turns noise into fake data, and a Discriminator learns to tell fake from real, pushing the Generator to keep doing better.
One runs upside down, the other right way up.
I drew and calculated one entirely by hand.
Goal: generate realistic 4D data out of 2D noise, filling in every cell yourself.
= 1. Given =
Four noise vectors in 2D, and four real data vectors in 4D.
= 2. Generator, first layer =
Let us multiply the noise by weights and biases to get new features.
= 3. ReLU =
We apply the activation, and -1 and -2 are crossed out and set to 0.
= 4. Generator, second layer =
Let us multiply again. ReLU applies here too, but every value is already positive, so nothing changes. What comes out is the fake data F, made by a two-layer generator out of nothing but noise.
= 5. Discriminator, first layer =
We feed it both, the four fakes and the four real vectors, through the same weights. It never learns which is which from the layout, only from the numbers.
= 6. Discriminator, second layer =
Let us reduce each data vector to a single feature Z. Eight vectors in, eight numbers out.
= 7. Sigmoid =
We turn each Z into a probability Y. A 1 means the discriminator is certain the data is real, a 0 means certain it is fake.
= 8. Training the Discriminator =
Let us take the gradients as Y minus YD, where YD is what the discriminator should have said: 0 for the four fakes, 1 for the four real. Why so simple? Because pairing sigmoid with binary cross entropy loss makes the math collapse to exactly this subtraction. Its loss uses both halves of the page.
= 9. Training the Generator =
We do it again, as Y minus YG, and YG is [1, 1, 1, 1]: the generator wants the discriminator to call every fake real. Same predictions, different target, opposite goal. Its loss uses only the fakes.
The outputs:
Fake data F = [1, 2, 3, 1], [1, 1, 2, 1], [2, 2, 4, 2], [1, 0, 1, 1]
Predictions on fakes = [.7, .5, .9, .3]
Predictions on real = [.7, .9, .9, 1]
Discriminator gradients = [.7, .5, .9, .3] and [-.3, -.1, -.1, 0]
Generator gradients = [-.3, -.5, -.1, -.7]
The takeaway: the adversarial part is one subtraction done twice. The same eight predictions, scored against two opposite targets, send one set of gradients back through the blue weights and another back through the green ones.
โค4