HyperAIHyperAI

Command Palette

Search for a command to run...

3 months ago

$α$ DARTS Once More: Enhancing Differentiable Architecture Search by Masked Image Modeling

Bicheng Guo Shuxuan Guo Miaojing Shi Peng Chen Shibo He Jiming Chen Kaicheng Yu

$α$ DARTS Once More: Enhancing Differentiable Architecture Search by Masked Image Modeling

Abstract

Differentiable architecture search (DARTS) has been a mainstream direction in automatic machine learning. Since the discovery that original DARTS will inevitably converge to poor architectures, recent works alleviate this by either designing rule-based architecture selection techniques or incorporating complex regularization techniques, abandoning the simplicity of the original DARTS that selects architectures based on the largest parametric value, namely $α$. Moreover, we find that all the previous attempts only rely on classification labels, hence learning only single modal information and limiting the representation power of the shared network. To this end, we propose to additionally inject semantic information by formulating a patch recovery approach. Specifically, we exploit the recent trending masked image modeling and do not abandon the guidance from the downstream tasks during the search phase. Our method surpasses all previous DARTS variants and achieves state-of-the-art results on CIFAR-10, CIFAR-100, and ImageNet without complex manual-designed strategies.

Benchmarks

BenchmarkMethodologyMetrics
neural-architecture-search-on-nas-bench-201α-DARTS
Accuracy (Test): 46.34
Accuracy (Val): 46.17
neural-architecture-search-on-nas-bench-201-1α-DARTS
Accuracy (Test): 94.3
Accuracy (Val): 91.49
neural-architecture-search-on-nas-bench-201-2α-DARTS
Accuracy (Test): 73.16
Accuracy (Val): 73.21

Build AI with AI

From idea to launch — accelerate your AI development with free AI co-coding, out-of-the-box environment and best price of GPUs.

AI Co-coding
Ready-to-use GPUs
Best Pricing
Get Started

Hyper Newsletters

Subscribe to our latest updates
We will deliver the latest updates of the week to your inbox at nine o'clock every Monday morning
Powered by MailChimp
$α$ DARTS Once More: Enhancing Differentiable Architecture Search by Masked Image Modeling | Papers | HyperAI