Independent AI research lab
New methods for how networks are built and trained.
Mathalyse develops optimisers and architectures, and the mathematics that explains why they work. Each method starts from a mechanism, and each claim is sized to its evidence.
Gradient descentloss
Momentumloss
Adamloss
Three optimisers released from the same point on one non-convex loss surface. The deepest well sits to the right of centre. The update rule decides which well each run reaches. Click the surface to release all three from a new point.