neural-architecture-search

Tag

Cards List
#neural-architecture-search

Neural Architecture Search for Traffic Prediction: A Survey of Methods, Challenges, and Future Directions

arXiv cs.LG · 2026-07-30 Cached

This survey reviews neural architecture search (NAS) methods applied to traffic prediction, organizing them by search strategy (gradient-based, evolutionary, one-shot weight-sharing) and discussing challenges such as computational scalability, cross-city generalization, and future directions.

0 favorites 0 likes
#neural-architecture-search

OrchNAS: Orchestrated Neural Architecture Search Service for Personalised Federated Edge Intelligence

arXiv cs.LG · 2026-07-28 Cached

The paper proposes OrchNAS, an energy-aware personalized federated edge intelligence framework that uses a Neural Architecture Search service to automatically design service-adaptive models for heterogeneous edge environments, addressing energy constraints and statistical heterogeneity.

0 favorites 0 likes
#neural-architecture-search

Scaling Closed-Loop Feature Channel Configuration with LLMs

arXiv cs.LG · 2026-07-24 Cached

This paper scales a closed-loop LLM-based channel configuration search to 250 candidates per cycle, showing positive accuracy trends and improved parameter efficiency on CIFAR-100, and revealing architectural regularities in LLM-generated channel priors.

0 favorites 0 likes
#neural-architecture-search

BearingNAS: Obtaining In-Sensor Intelligent Fault Diagnosis Systems for Bearings Using a Laptop

arXiv cs.LG · 2026-07-22 Cached

The paper presents BearingNAS, a hardware-aware neural architecture search framework that designs intelligent fault diagnosis systems for bearings, targeting microcontrollers and sensor processing units with extremely limited memory (4-8 KiB RAM, 16-32 KiB Flash) while running entirely on a laptop CPU and achieving 99.50% diagnostic accuracy.

0 favorites 0 likes
#neural-architecture-search

LLM-Driven AutoML for Cross-Lingual Handwritten OCR: Closed-Loop Neural Architecture Search with GPT-5, GPT-4o, and Claude Sonnet 4

arXiv cs.AI · 2026-07-20 Cached

This paper presents an LLM-driven pipeline using GPT-5, GPT-4o, and Claude Sonnet 4 to automatically design neural network architectures for cross-lingual handwritten OCR, achieving over 93% accuracy across Arabic, English, and Persian scripts without human intervention.

0 favorites 0 likes
#neural-architecture-search

@cwolferesearch: Open tech reports / artifacts are so valuable. I'm currently reading through all of the nemotron tech reports and they …

X AI KOLs Timeline · 2026-07-12 Cached

NVIDIA introduces Llama-Nemotron, an open family of reasoning models (Nano 8B, Super 49B, Ultra 253B) that rival DeepSeek-R1 with superior inference efficiency, dynamic reasoning toggle, and open post-training datasets.

0 favorites 0 likes
#neural-architecture-search

Agentic Neural Architecture Search

arXiv cs.AI · 2026-07-10 Cached

Introduces AgentNAS, a mechanism that uses an LLM to generate a seed architecture and decompose it into a slotted architecture, defining a task-specific search space for conventional NAS to explore, achieving state-of-the-art on 11 of 17 tasks.

0 favorites 0 likes
#neural-architecture-search

LEMUR 2: Unlocking Neural Network Diversity for AI

arXiv cs.LG · 2026-07-09 Cached

LEMUR 2 introduces a large-scale dataset of over 14,000 neural network architectures and 750,000 training records across multimodal tasks, supporting NAS, AutoML, and deployment analysis.

0 favorites 0 likes
#neural-architecture-search

LLM-Driven Neural Network Generation with Same-Family Architecture Guidance: Disentangling Transfer and Adaptation

arXiv cs.LG · 2026-07-08 Cached

This paper proposes a source-guided protocol where an LLM generates candidate modifications for a weak target model using a stronger same-family source model, showing substantial accuracy improvements on CIFAR-10 and SVHN benchmarks while disentangling transfer from adaptation effects.

0 favorites 0 likes
#neural-architecture-search

CamoNAS: Neural Architecture Search for Enhanced Camouflaged Object Detection

arXiv cs.AI · 2026-07-03 Cached

Introduces CamoNAS, a frequency-aware multi-resolution Neural Architecture Search framework for camouflaged object detection, achieving state-of-the-art results on four benchmarks.

0 favorites 0 likes
#neural-architecture-search

EVOTS: Evolutionary Transformer Search for Time Series Forecasting

arXiv cs.LG · 2026-07-02 Cached

Introduces an evolutionary neural architecture search framework (EvoTS) for discovering task-adaptive Transformer-like models for multivariate time-series forecasting. The approach uses a modular genome representation and achieves competitive performance on ETT benchmark datasets.

0 favorites 0 likes
#neural-architecture-search

EVOM: Agentic Meta-Evolution of Actor-Critic Architectures for Reinforcement Learning

arXiv cs.LG · 2026-06-26 Cached

Introduces EVOM, an agentic meta-evolution framework using an LLM-based design agent to automatically discover high-performance actor-critic architectures for reinforcement learning, outperforming manual baselines and prior methods on continuous control tasks.

0 favorites 0 likes
#neural-architecture-search

Neural Architecture Search for Generative Adversarial Networks: A Comprehensive Review and Critical Analysis

arXiv cs.LG · 2026-06-26 Cached

This paper provides a comprehensive review of Neural Architecture Search (NAS) methods applied to Generative Adversarial Networks (GANs), categorizing approaches and highlighting benefits and limitations.

0 favorites 0 likes
#neural-architecture-search

On-Device Neural Architecture Search

arXiv cs.LG · 2026-06-25 Cached

Proposes a lightweight neural architecture search performed directly on the deployment device for near-sensor computing, validated on sEMG sign language and fault diagnosis datasets, achieving improved accuracy and reduced RAM occupancy.

0 favorites 0 likes
#neural-architecture-search

Systematic Exploration of 4-Expert Heterogeneous Mixture-of-Experts via Automated Pipeline Search

arXiv cs.LG · 2026-06-24 Cached

This paper presents an automated pipeline for searching heterogeneous 4-expert Mixture-of-Experts architectures, exploring 4.8% of the theoretical combination space and identifying high- and low-yield expert families. The work releases analysis artifacts and a corrected generator as part of the open-source NNGPT project.

0 favorites 0 likes
#neural-architecture-search

LLM Compression with Jointly Optimizing Architectural and Quantization choices

arXiv cs.LG · 2026-06-04 Cached

Researchers from UiT and University of Oslo propose a differentiable NAS framework that jointly optimizes architectural configurations and mixed-precision quantization for LLM compression, achieving up to 1.4× faster inference or 6% higher accuracy across seven reasoning tasks compared to sequential NAS-then-quantization baselines.

0 favorites 0 likes
#neural-architecture-search

@dair_ai: NEW paper from Meta: Agentic Discovery of Neural Architectures. This is a hot new area of research! Keep an eye on it.

X AI KOLs Following · 2026-05-18 Cached

Meta's new paper presents an agentic system that autonomously discovers neural architectures outperforming Llama 3.2 at 350M, 1B, and 3B scales within a 24-hour compute budget.

0 favorites 0 likes
#neural-architecture-search

Agentic Discovery of Neural Architectures: AIRA-Compose and AIRA-Design

Hugging Face Daily Papers · 2026-05-15 Cached

This paper introduces AIRA-Compose and AIRA-Design, dual frameworks using AI agents to autonomously discover neural architectures that outperform standard Transformers and scale efficiently.

0 favorites 0 likes
#neural-architecture-search

Learngene Search Across Multiple Datasets for Building Variable-Sized Models

arXiv cs.LG · 2026-05-12 Cached

This paper introduces LSAMD, a method for extracting 'learngenes' across multiple datasets to initialize variable-sized Vision Transformer models, significantly reducing training costs and storage while maintaining performance comparable to pretrain-finetune methods.

0 favorites 0 likes
#neural-architecture-search

RF-DETR: Neural Architecture Search for Real-Time Detection Transformers

Papers with Code Trending · 2025-11-12 Cached

RF-DETR introduces a lightweight detection transformer that uses weight-sharing neural architecture search to achieve state-of-the-art real-time object detection, outperforming prior methods on COCO and Roboflow100-VL while running up to 20x faster.

0 favorites 0 likes
← Back to home

Submit Feedback