{"url":"/method/non","slug":"non","name":"NON","full_name":"Network On Network","full_name_withheld":false,"description_markdown":"Network On Network (NON) is practical tabular data classification model based on deep neural network to provide accurate predictions. Various deep methods have been proposed and promising progress has been made. However, most of them use operations like neural network and factorization machines to fuse the embeddings of different features directly, and linearly combine the outputs of those operations to get the final prediction. As a result, the intra-field information and the non-linear interactions between those operations (e.g. neural network and factorization machines) are ignored. Intra-field information is the information that features inside each field belong to the same field. NON is proposed to take full advantage of intra-field information and non-linear interactions. It consists of three components: field-wise network at the bottom to capture the intra-field information, across field network in the middle to choose suitable operations data-drivenly, and operation fusion network on the top to fuse outputs of the chosen operations deeply","description_state":"present","introduced_year":null,"introduced_by":{"title":null,"paper":null,"first_author":null,"n_authors":0,"url_abs":null,"archive_paper_url":null},"source":{"url":"https://arxiv.org/abs/2005.10114v2","title":"Network On Network for Tabular Data Classification in Real-world Applications","url_on_a_paper_host":true},"code_snippet_url":null,"code_snippet_url_on_a_code_host":false,"categories":[{"area":"General","area_id":"general","collection":"Deep Tabular Learning","url":"/methods/category/deep-tabular-learning","pwc_aliases":[]}],"n_papers_tagged":389,"archive_num_papers":null,"papers_newest_first":[{"paper":null,"title":"Deep Semantic Segmentation for Multi-Source Localization Using Angle of Arrival Measurements","date":"2025-06-11","arxiv_id":"2506.10107","n_code_links":0,"syntology":null},{"paper":null,"title":"Wavelet Scattering Transform and Fourier Representation for Offline Detection of Malicious Clients in Federated Learning","date":"2025-06-11","arxiv_id":"2506.09674","n_code_links":0,"syntology":null},{"paper":null,"title":"Model-based Neural Data Augmentation for sub-wavelength Radio Localization","date":"2025-06-05","arxiv_id":"2506.06387","n_code_links":0,"syntology":null},{"paper":"/paper/cdr-agent-intelligent-selection-and-execution","title":"CDR-Agent: Intelligent Selection and Execution of Clinical Decision Rules Using Large Language Model Agents","date":"2025-05-29","arxiv_id":"2505.23055","n_code_links":1,"syntology":null},{"paper":null,"title":"Seeing the Threat: Vulnerabilities in Vision-Language Models to Adversarial Attack","date":"2025-05-28","arxiv_id":"2505.21967","n_code_links":0,"syntology":null},{"paper":null,"title":"LLM assisted web application functional requirements generation: A case study of four popular LLMs over a Mess Management System","date":"2025-05-23","arxiv_id":"2505.18019","n_code_links":0,"syntology":null},{"paper":null,"title":"Convergence of Adam in Deep ReLU Networks via Directional Complexity and Kakeya Bounds","date":"2025-05-21","arxiv_id":"2505.15013","n_code_links":0,"syntology":null},{"paper":null,"title":"The Triad of Modern Democracies: Money, Identity, and Information in Shaping Power and Legitimacy","date":"2025-05-14","arxiv_id":"2505.09124","n_code_links":0,"syntology":null},{"paper":"/paper/towards-scalable-surrogate-models-based-on","title":"Towards scalable surrogate models based on Neural Fields for large scale aerodynamic simulations","date":"2025-05-14","arxiv_id":"2505.14704","n_code_links":1,"syntology":null},{"paper":null,"title":"The Steganographic Potentials of Language Models","date":"2025-05-06","arxiv_id":"2505.03439","n_code_links":0,"syntology":null},{"paper":null,"title":"End-to-end fully-binarized network design: from Generic Learned Thermometer to Block Pruning","date":"2025-05-05","arxiv_id":"2505.13462","n_code_links":0,"syntology":null},{"paper":null,"title":"Multi-Step Consistency Models: Fast Generation with Theoretical Guarantees","date":"2025-05-02","arxiv_id":"2505.01049","n_code_links":0,"syntology":null},{"paper":null,"title":"Spatiotemporal Emotional Synchrony in Dyadic Interactions: The Role of Speech Conditions in Facial and Vocal Affective Alignment","date":"2025-04-29","arxiv_id":"2505.13455","n_code_links":0,"syntology":null},{"paper":null,"title":"Generalizing the Levins metapopulation model to time varying colonization and extinction rates","date":"2025-04-29","arxiv_id":"2504.20396","n_code_links":0,"syntology":null},{"paper":null,"title":"Score-Based Deterministic Density Sampling","date":"2025-04-25","arxiv_id":"2504.18130","n_code_links":0,"syntology":null},{"paper":null,"title":"IlluSign: Illustrating Sign Language Videos by Leveraging the Attention Mechanism","date":"2025-04-15","arxiv_id":"2504.10822","n_code_links":0,"syntology":null},{"paper":null,"title":"Age-of-information minimization under energy harvesting and non-stationary environment","date":"2025-04-07","arxiv_id":"2504.04916","n_code_links":0,"syntology":null},{"paper":null,"title":"Integrated LLM-Based Intrusion Detection with Secure Slicing xApp for Securing O-RAN-Enabled Wireless Network Deployments","date":"2025-04-01","arxiv_id":"2504.00341","n_code_links":0,"syntology":null},{"paper":"/paper/design-and-evaluation-of-neural-network-based","title":"Novel Deep Neural OFDM Receiver Architectures for LLR Estimation","date":"2025-03-26","arxiv_id":"2503.20500","n_code_links":1,"syntology":null},{"paper":null,"title":"Finite-Time Bounds for Two-Time-Scale Stochastic Approximation with Arbitrary Norm Contractions and Markovian Noise","date":"2025-03-24","arxiv_id":"2503.18391","n_code_links":0,"syntology":null},{"paper":null,"title":"Mechanistic Interpretability of Fine-Tuned Vision Transformers on Distorted Images: Decoding Attention Head Behavior for Transparent and Trustworthy AI","date":"2025-03-24","arxiv_id":"2503.18762","n_code_links":0,"syntology":null},{"paper":null,"title":"Robustness of deep learning classification to adversarial input on GPUs: asynchronous parallel accumulation is a source of vulnerability","date":"2025-03-21","arxiv_id":"2503.17173","n_code_links":0,"syntology":null},{"paper":null,"title":"Structural and Practical Identifiability of Phenomenological Growth Models for Epidemic Forecasting","date":"2025-03-21","arxiv_id":"2503.17135","n_code_links":0,"syntology":null},{"paper":null,"title":"Estimating stationary mass, frequency by frequency","date":"2025-03-17","arxiv_id":"2503.12808","n_code_links":0,"syntology":null},{"paper":null,"title":"Non Line-of-Sight Optical Wireless Communication using Neuromorphic Cameras","date":"2025-03-14","arxiv_id":"2503.11226","n_code_links":0,"syntology":null},{"paper":"/paper/automatic-quality-control-in-multi-centric","title":"Automatic quality control in multi-centric fetal brain MRI super-resolution reconstruction","date":"2025-03-13","arxiv_id":"2503.10156","n_code_links":3,"syntology":null},{"paper":null,"title":"From Linear to Spline-Based Classification:Developing and Enhancing SMPA for Noisy Non-Linear Datasets","date":"2025-03-13","arxiv_id":"2503.10545","n_code_links":0,"syntology":null},{"paper":null,"title":"Beyond Black-Box Benchmarking: Observability, Analytics, and Optimization of Agentic Systems","date":"2025-03-09","arxiv_id":"2503.06745","n_code_links":0,"syntology":null},{"paper":null,"title":"PLS-based approach for fair representation learning","date":"2025-02-22","arxiv_id":"2502.16263","n_code_links":0,"syntology":null},{"paper":null,"title":"A Self-Supervised Reinforcement Learning Approach for Fine-Tuning Large Language Models Using Cross-Attention Signals","date":"2025-02-14","arxiv_id":"2502.10482","n_code_links":0,"syntology":null}],"papers_shown":30,"tasks":[{"task":"/task/federated-learning","name":"Federated Learning","papers":17},{"task":"/task/language-modelling","name":"Language Modelling","papers":11},{"task":"/task/time-series-1","name":"Time Series","papers":10},{"task":"/task/regression-1","name":"regression","papers":10},{"task":"/task/reinforcement-learning-2","name":"reinforcement-learning","papers":10},{"task":"/task/clustering","name":"Clustering","papers":8},{"task":"/task/decision-making","name":"Decision Making","papers":8},{"task":"/task/denoising","name":"Denoising","papers":8},{"task":"/task/reinforcement-learning-1","name":"Reinforcement Learning (RL)","papers":8},{"task":"/task/retrieval","name":"Retrieval","papers":8},{"task":"/task/classification-1","name":"Classification","papers":7},{"task":"/task/fairness","name":"Fairness","papers":7},{"task":"/task/language-modeling","name":"Language Modeling","papers":7},{"task":"/task/management","name":"Management","papers":7},{"task":"/task/object-detection","name":"Object Detection","papers":7},{"task":"/task/reinforcement-learning","name":"Reinforcement Learning","papers":7},{"task":"/task/object-detection-1","name":"object-detection","papers":7},{"task":"/task/binary-classification","name":"Binary Classification","papers":6},{"task":"/task/deep-learning","name":"Deep Learning","papers":6},{"task":"/task/semantic-segmentation","name":"Semantic Segmentation","papers":6}],"tasks_shown":20,"n_tasks":280,"usage_by_year":[{"year":"2020","papers":2},{"year":"2021","papers":22},{"year":"2022","papers":104},{"year":"2023","papers":111},{"year":"2024","papers":109},{"year":"2025","papers":41}],"row_source":"embedded","archive":{"source":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","archive_url":"https://paperswithcode.com/method/non"},"syntology_read_at":"2026-09-24T18:15:14+00:00"}