LGMLJan 9, 2018

Convexification of Neural Graph

arXiv:1801.02901v22.22 citations
Originality Incremental advance
AI Analysis

This addresses optimization challenges in neural networks for researchers, though it appears incremental as it builds on existing graph-based methods.

The paper tackles the non-convexity of neural architectures by decomposing them into nodes and proving that tree-structured graphs are nearly convex under certain conditions, with experiments validating the theory in two applications.

Traditionally, most complex intelligence architectures are extremely non-convex, which could not be well performed by convex optimization. However, this paper decomposes complex structures into three types of nodes: operators, algorithms and functions. Iteratively, propagating from node to node along edge, we prove that "regarding the tree-structured neural graph, it is nearly convex in each variable, when the other variables are fixed." In fact, the non-convex properties stem from circles and functions, which could be transformed to be convex with our proposed \textit{\textbf{scale mechanism}}. Experimentally, we justify our theoretical analysis by two practical applications.

Foundations

The foundational work for this paper's niche, ranked by how specifically the neighbourhood builds on it — not by global fame.

Your Notes