---
title: Can overfitted deep neural networks in adversarial training generalize? - An approximation viewpoint
url: https://www.ml-quant.com/papers/arxiv/2401.13624/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-09-26
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2401.13624
source_url: http://arxiv.org/abs/2401.13624
featured: 2024-01-30
citations: 2
topic: ML & AI Methods
---


# Can overfitted deep neural networks in adversarial training generalize? - An approximation viewpoint

The study offers a theoretical insight into the robust overfitting issue in adversarial training on over-parameterized deep neural networks (DNNs), showing that overfitting can be avoided and a robust generalization gap is unavoidable, with the model capacity requirement depending on the target function's smoothness.

- Source: http://arxiv.org/abs/2401.13624
- Identifier: arXiv:2401.13624
- Released: 2024-01-24
- First featured: Quant Letter No. 35 (2024-01-30): https://www.ml-quant.com/issues/2024-01-30/
- Citations (Semantic Scholar): 2
- Published in: not yet
- Topic: ML & AI Methods

## Related

- [Edge Directionality Improves Learning on Heterophilic Graphs](https://www.ml-quant.com/papers/arxiv/2305.10498/): The study presents Directed Graph Neural Network (Dir-GNN), a new deep learning framework for directed graphs that surpasses traditional models in heterophilic benchmarks.
- [Modular Duality in Deep Learning](https://www.ml-quant.com/papers/arxiv/2410.21265/): The article presents a new theory of modular dualization for general neural networks, providing a theoretical basis for fast and scalable training algorithms, potentially leading to a new generation of optimizers for neural architectures.
- [Historical calibration of SVJD models with deep learning](https://www.ml-quant.com/papers/ssrn/4650097/): The paper suggests using deep neural networks to calibrate parameters of Stochastic Volatility Jump Diffusion models, proving to be more accurate, robust, and faster than other methods.
- [Evaluating Adversarial Robustness: A Comparison Of FGSM, Carlini-Wagner Attacks, And The Role of Distillation as Defense Mechanism](https://www.ml-quant.com/papers/arxiv/2404.04245/): The study investigates adversarial attacks on Deep Neural Networks for image classification, highlighting the Fast Gradient Sign Method and the Carlini-Wagner approach, and suggests defensive distillation as a defense, effective against FGSM but vulnerable to CW attacks.
- [GraphCNNpred: A stock market indices prediction using a Graph based deep learning system](https://www.ml-quant.com/papers/arxiv/2407.03760/): Stock Prediction: The paper introduces a graph neural network-based convolutional neural network model for predicting stock market prices, using custom feature engineering on diverse data sources.
- [Topological Generalization Bounds for Discrete-Time Stochastic Optimization Algorithms](https://www.ml-quant.com/papers/arxiv/2407.08723/): The research proposes a new set of topology-based complexity notions that correlate with the generalization gap in deep neural networks, offering a computationally efficient way to predict generalization without test data, and surpassing existing topological bounds across various datasets and models.
