Gehaltener Vortrag - Details

Original Vortragstitel:	Modern Hopfield Networks
Sprache des Vortragstitels:	Englisch
Original Tagungtitel:	International Conference on Machine Vision and Machine Learning (MVML21)
Sprache des Tagungstitel:	Englisch
Original Kurzfassung:	We propose a new paradigm for deep learning by equipping each layer of a deep learning architecture with modern Hopfield networks. The new paradigm is a new powerful concept comprising functionalities like pooling, memory, and attention for each layer. Associative memories date back to the 1960/70s and became popular through Hopfield Networks in 1982. Recently, we saw a renaissance of Hopfield Networks, the modern Hopfield Networks, with a tremendously increased storage capacity and an extremely fast convergence. We generalize modern Hopfield Networks with exponential storage capacity to continuous patterns. Their update rule ensures global convergence to local energy minima and they converge in one update step with exponentially low error. Surprisingly, the transformer attention mechanism is equal to the update rule of our new modern Hopfield Network with continuous states. The new modern Hopfield network can be integrated into deep learning architectures as layers to allow the storage of and access to raw input data, intermediate results, or learned prototypes. These Hopfield layers enable new ways of deep learning, beyond fully-connected, convolutional, or recurrent networks, and provide pooling, memory, association, and attention mechanisms. We demonstrate the broad applicability of the Hopfield layers across various domains. Hopfield layers improved state-of-the-art on three out of four considered multiple instance learning problems as well as on immune repertoire classification with several hundreds of thousands of instances. On the UCI benchmark collections of small classification tasks, where deep learning methods typically struggle, Hopfield layers yielded a new state-of-the-art when compared to different machine learning methods. Finally, Hopfield layers achieved state-of-the-art on two drug design datasets.
Sprache der Kurzfassung:	Englisch
Vortragstyp:	Hauptvortrag / Eingeladener Vortrag auf einer Tagung
Vortragsdatum:	30.07.2021
Vortragsort:	Österreich
Details zum Vortragsort:	Online
Vortragende:	Sepp Hochreiter
Forschungseinheiten:	LIT Artificial Intelligence Lab Institut für Machine Learning

Wissenschaftsgebiete:	Artificial Intelligence (ÖSTAT:102001) Bildverarbeitung (ÖSTAT:102003) Bioinformatik (ÖSTAT:102004) Bioinformatik (ÖSTAT:106005) Biomathematik (ÖSTAT:101004) Approximationstheorie (ÖSTAT:101031) Biostatistik (ÖSTAT:106007) Computational Intelligence (ÖSTAT:102032) Computerunterstützte Diagnose und Therapie (ÖSTAT:305901) Data Mining (ÖSTAT:102033) Dynamische Systeme (ÖSTAT:101027) Embedded Systems (ÖSTAT:202017) Human-Computer Interaction (ÖSTAT:102013) Informatik (ÖSTAT:102) Künstliche Neuronale Netze (ÖSTAT:102018) Machine Learning (ÖSTAT:102019) Mathematische Modellierung (ÖSTAT:101028) Mathematische Statistik (ÖSTAT:101029) Medizinische Informatik (ÖSTAT:305905) Medizinische Statistik (ÖSTAT:305907) Numerische Mathematik (ÖSTAT:101014) Operations Research (ÖSTAT:101015) Optimierung (ÖSTAT:101016) Robotik (ÖSTAT:202035) Sensorik (ÖSTAT:202036) Signalverarbeitung (ÖSTAT:202037) Spieltheorie (ÖSTAT:101017) Statistik (ÖSTAT:101018) Statistische Physik (ÖSTAT:103029) Stochastik (ÖSTAT:101019) Wahrscheinlichkeitstheorie (ÖSTAT:101024) Zeitreihenanalyse (ÖSTAT:101026)

fodok.jku.at

Benutzerbetreuung: Sandra Winzer, letzte Änderung:

Johannes Kepler Universität (JKU) Linz, Altenbergerstr. 69, A-4040 Linz, Austria
Telefon + 43 732 / 2468 - 9121, Fax + 43 732 / 2468 - 29121, Internet www.jku.at, Impressum