Mathematical Circuits in Neural Nets

By: Sean Osier

Summary

In this project, I:

Derive by hand the optimal configurations (architecture and weights) of "vanilla" neural networks (multilayer perceptrons; ReLU activations) that implement basic mathematical functions (e.g. absolute value, minimum of two numbers, etc.)
Identify "features" and "circuits" of these networks that are reused repeatedly across networks modeling different mathematical functions
Verify these theoretical results empirically (in code)

What follows is a brief introduction to this work. For full details, please see:

My presentation of this work on YouTube (also linked at the bottom of this page)
Or if you prefer to go at your own pace, the slides I walk through in that video

Motivation

Olah et al. make three claims about the fundamental interpretability of neural networks:

They demonstrate these claims in the context of image models:

Features / Circuits:

Universality:

This work demonstrate the same concepts apply in the space of neural networks modeling basic mathematical functions.

Results

Specifically, I show that the optimal network for calculating the minimum of two arbitrary numbers is fully constructed from smaller "features" and "circuits" used across even simpler mathematical functions. Along the way, I explore:

"Positiveness" and "Negativeness" Detectors
Identity Circuits (i.e. f(x) = x)
Negative Identity Circuits (i.e. f(x) = -x)
Subtraction Circuits (i.e. f(x1, x2) = x1 - x2)
"Greaterness" Detectors
And More

Minimum Network:

I also demonstrate that each of these theoretical results hold in practice. The code for these experiments can be found in the Jupyter Notebook in this repo.

Full Details

For full details, please see the PDF presenation in this repo.

Watch the Video! (Click image below)

You can watch a full walkthrough of the presentation by clicking the image below:

Name		Name	Last commit message	Last commit date
Latest commit History 17 Commits
.gitignore		.gitignore
Mathematical_Circuits.ipynb		Mathematical_Circuits.ipynb
Mathematical_Circuits_in_Neural_Networks.pdf		Mathematical_Circuits_in_Neural_Networks.pdf
README.md		README.md
Video_Thumbnail.png		Video_Thumbnail.png

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

Mathematical Circuits in Neural Nets

Summary

Motivation

Results

Full Details

Watch the Video! (Click image below)

About

Releases

Packages

Languages

sosier/Mathematical_Circuits_in_Neural_Nets

Folders and files

Latest commit

History

Repository files navigation

Mathematical Circuits in Neural Nets

Summary

Motivation

Results

Full Details

Watch the Video! (Click image below)

About

Resources

Stars

Watchers

Forks

Releases

Packages 0

Languages

Packages