• bitcoinBitcoin(BTC)$76,782.00-1.86%
  • ethereumEthereum(ETH)$2,449.90-0.73%
  • tetherTether(USDT)$1.000.00%
  • binancecoinBNB(BNB)$711.60-1.44%
  • rippleXRP(XRP)$1.34-3.83%
  • usd-coinUSDC(USDC)$1.000.01%
  • solanaSolana(SOL)$98.85-2.42%
  • tronTRON(TRX)$0.3404600.39%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.040.98%
  • zcashZcash(ZEC)$1,088.35-11.99%
  • HyperliquidHyperliquid(HYPE)$79.09-4.98%
  • dogecoinDogecoin(DOGE)$0.083266-3.11%
  • RainRain(RAIN)$0.015742-1.94%
  • USDSUSDS(USDS)$1.00-0.01%
  • moneroMonero(XMR)$511.010.14%
  • whitebitWhiteBIT Coin(WBT)$79.50-1.58%
  • chainlinkChainlink(LINK)$11.48-2.24%
  • leo-tokenLEO Token(LEO)$9.190.01%
  • cardanoCardano(ADA)$0.205925-2.75%
  • stellarStellar(XLM)$0.175573-2.68%
  • daiDai(DAI)$1.000.00%
  • bitcoin-cashBitcoin Cash(BCH)$225.56-10.27%
  • Ethena USDeEthena USDe(USDE)$1.00-0.01%
  • USD1USD1(USD1)$1.00-0.01%
  • litecoinLitecoin(LTC)$52.47-1.37%
  • CantonCanton(CC)$0.098257-5.13%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.36-0.74%
  • uniswapUniswap(UNI)$6.03-1.93%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • hedera-hashgraphHedera(HBAR)$0.075153-1.88%
  • nearNEAR Protocol(NEAR)$2.47-0.17%
  • avalanche-2Avalanche(AVAX)$7.46-4.02%
  • suiSui(SUI)$0.73-4.60%
  • shiba-inuShiba Inu(SHIB)$0.000005-3.41%
  • paypal-usdPayPal USD(PYUSD)$1.000.01%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • crypto-com-chainCronos(CRO)$0.056285-3.23%
  • tether-goldTether Gold(XAUT)$4,323.15-1.61%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • MemeCoreMemeCore(M)$1.15-5.48%
  • Ripple USDRipple USD(RLUSD)$1.00-0.01%
  • okbOKB(OKB)$109.05-4.16%
  • BittensorBittensor(TAO)$235.78-6.76%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.140.05%
  • polkadotPolkadot(DOT)$1.12-0.15%
  • AsterAster(ASTER)$0.70-3.51%
  • mantleMantle(MNT)$0.57-5.11%
  • aaveAave(AAVE)$121.48-2.51%
  • pax-goldPAX Gold(PAXG)$4,328.20-1.54%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.0561400.24%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Meet MC-JEPA: A Joint-Embedding Predictive Architecture for Self-Supervised Learning of Motion and Content Features

July 31, 2023
in AI & Technology
Reading Time: 4 mins read
A A
Meet MC-JEPA: A Joint-Embedding Predictive Architecture for Self-Supervised Learning of Motion and Content Features
ShareShareShareShareShare

Recently, techniques focusing on learning content features—specifically, features holding the information that lets us identify and discriminate objects—have dominated self-supervised learning in vision. Most techniques concentrate on identifying broad characteristics that perform well in tasks like item categorization or activity detection in films. Learning localized features that excel at regional tasks like segmentation and detection is a relatively recent concept. However, these techniques concentrate on comprehending the content of pictures and videos rather than being able to learn characteristics about pixels, such as motion in films or textures. 

In this research, authors from Meta AI, PSL Research University, and New York University concentrate on simultaneously learning content characteristics with generic self-supervised learning and motion features utilizing self-supervised optical flow estimates from movies as a pretext problem. When two pictures—for example, successive frames in a movie or images from a stereo pair—move or have a dense pixel connection, it is captured by optical flow. In computer vision, estimating is a basic problem whose resolution is essential to operations like visual odometry, depth estimation, or object tracking. According to traditional methods, estimating optical flow is an optimization issue that aims to match pixels with a smoothness requirement. 

The challenge of categorizing real-world data instead of synthetic data limits approaches based on neural networks and supervised learning. Self-supervised techniques now compete with supervised techniques by allowing learning from substantial amounts of real-world video data. The majority of current approaches, however, only pay attention to motion rather than the (semantic) content of the video. This issue is resolved by simultaneously learning motion and content elements in pictures using a multi-task approach. Recent methods identify spatial relationships between video frames. The objective is to follow the movement of objects to collect content data that optical flow estimates cannot. 

These methods are object-level motion estimation methods. With relatively weak generalization to other visual downstream tasks, they acquire highly specialized characteristics for the tracking job. The low quality of the visual characteristics learned is reinforced by the fact that they are frequently trained on tiny video datasets that need more diversity than larger picture datasets like ImageNet. Learning several activities simultaneously is a more reliable technique for developing visual representations. To solve this problem, they offer MC-JEPA (Motion-Content Joint-Embedding Predictive Architecture). Using a common encoder, this joint-embedding-predictive architecture-based system learns optical flow estimates and content characteristics in a multi-task environment. 

The following is a summary of their contributions: 

• They offer a technique based on PWC-Net that is augmented with numerous extra elements, such as a backward consistency loss and a variance-covariance regularisation term, for learning self-supervised optical flow from synthetic and real video data. 

• They use M-JEPA with VICReg, a self-supervised learning technique trained on ImageNet, in a multi-task configuration to optimize their estimated flow and provide content characteristics that transfer well to several downstream tasks. The name of their ultimate approach is MC-JEPA. 

• They tested MC-JEPA on a variety of optical flow benchmarks, including KITTI 2015 and Sintel, as well as image and video segmentation tasks on Cityscapes or DAVIS, and they found that a single encoder performed well on each of these tasks. They anticipate that MC-JEPA will be a precursor to self-supervised learning methodologies based on joint embedding and multi-task learning that can be trained on any visual data, including images and videos, and perform well across various tasks, from motion prediction to content understanding.


Check out the Paper. All Credit For This Research Goes To the Researchers on This Project. Also, don’t forget to join our 27k+ ML SubReddit, 40k+ Facebook Community, Discord Channel, and Email Newsletter, where we share the latest AI research news, cool AI projects, and more.


YOU MAY ALSO LIKE

How These XL Phones Compete

Meet Redis LangCache: A Managed Semantic Cache That Cuts LLM API Costs by Up to 90% and Returns Cache Hits Up to 15x Faster

Aneesh Tickoo is a consulting intern at MarktechPost. He is currently pursuing his undergraduate degree in Data Science and Artificial Intelligence from the Indian Institute of Technology(IIT), Bhilai. He spends most of his time working on projects aimed at harnessing the power of machine learning. His research interest is image processing and is passionate about building solutions around it. He loves to connect with people and collaborate on interesting projects.


🔥 Gain a competitive
edge with data: Actionable market intelligence for global brands, retailers, analysts, and investors. (Sponsored)

Credit: Source link

ShareTweetSendSharePin

Related Posts

How These XL Phones Compete
AI & Technology

How These XL Phones Compete

September 10, 2026
Meet Redis LangCache: A Managed Semantic Cache That Cuts LLM API Costs by Up to 90% and Returns Cache Hits Up to 15x Faster
AI & Technology

Meet Redis LangCache: A Managed Semantic Cache That Cuts LLM API Costs by Up to 90% and Returns Cache Hits Up to 15x Faster

September 10, 2026
Meta Is Testing Community Notes In Latin America. Fact Checkers Are Worried.
AI & Technology

Meta Is Testing Community Notes In Latin America. Fact Checkers Are Worried.

September 10, 2026
IDScan Is Offering Free Credit Monitoring And ID Protection After Leaking Driver’s Licenses
AI & Technology

IDScan Is Offering Free Credit Monitoring And ID Protection After Leaking Driver’s Licenses

September 10, 2026
Next Post
Athleticwear Maker Under Armour Now Has the Largest Online Fitness Community

Athleticwear Maker Under Armour Now Has the Largest Online Fitness Community

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Deadly flooding hits Eastern U.S., Bertha impacts Texas

Deadly flooding hits Eastern U.S., Bertha impacts Texas

September 6, 2026
NVIDIA Announces CUDA Rust with cuda-oxide (SIMT) and cutile-rs (Tile) for Compile-Time-Safe GPU Kernels

NVIDIA Announces CUDA Rust with cuda-oxide (SIMT) and cutile-rs (Tile) for Compile-Time-Safe GPU Kernels

September 8, 2026
Friday Facts: Are Leveraged ETFs A Real Risk For Financial Markets?

Friday Facts: Are Leveraged ETFs A Real Risk For Financial Markets?

September 5, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!