healthcarereimagined

Envisioning healthcare for the 21st century

  • About
  • Economics

At NeurIPS, what’s old is new again – Amazon

Posted by timmreardon on 12/20/2023
Posted in: Uncategorized.

MACHINE LEARNING

At NeurIPS, what’s old is new again

Amazon Scholar and NeurIPS advisory board member Richard Zemel on what robustness and responsible AI have in common, what AI can still learn from neuroscience, and the emerging topics that interest him most.

By Larry Hardesty

December 13, 2023

Share

The current excitement around large language models is just the latest aftershock of the deep-learning revolution that started in 2012 (or maybe 2010), but Columbia professor and Amazon Scholar Richard Zemel was there before the beginning. As a PhD student at the University of Toronto in the late ’80s and early ’90s, Zemel wrote his dissertation on representation learning in unsupervised machine learning systems for Geoffrey Hinton, one of the three “godfathers of deep learning”.

Zemel is also on the advisory board of the main conference in the field of deep learning, the Conference on Neural Information Processing (NeurIPS), which takes place this week. His breadth of experience gives him a rare perspective on the field of deep learning — both how far it’s come and where it’s going.

“It’s come a very long way in some sense, in terms of the scope of problems that are relevant and the whole real-world applicability of it,” Zemel says. “But a lot of the same problems still exist. There are just many more facets than there used to be.”

For example, Zemel says, take the concept of robustness, the ability of a machine learning model to maintain performance when the data it sees at inference time differs from the data it was trained on, because of noise, drift in the data distribution, or the like.

“One of the original neural-net applications was ALVINN, the automated land vehicle in a neural network, in the late ’80s,” Zemel says. “It was a neural net that had 29 hidden units, and it was an answer to DARPA’s self-driving challenge. It was a big success for neural nets at the time.

“Robustness came up there because they were worried about the car going off the road, and they didn’t have any training examples of that. They worked out how to augment the data with those kinds of training examples. So thirty years ago, robustness was seen as an important question, and some ideas came up.”

Today, data augmentation remains one of the main ways to ensure robustness. But as Zemel says, the problem of robustness has new facets.

Neural attentive circuits.16x9.png

Related content

NeurIPS: Why causal-representation learning may be the future of AI

“For instance, we can consider algorithmic fairness as a form of robustness,” he says. “It’s robustness with respect to particular groups. A lot of the methods that are used for that are methods that have also been developed for robustness, and vice versa. For example, they’re formulated as trying to develop a prediction that has some invariance properties. And it could be that you’re not just developing a prediction: in the deep-learning world, you’re trying to develop a representation that has these properties. The final layer of representation should be invariant. Think of multiclass object recognition: anything that has a label of class K should have a very similar kind of distribution over representations, no matter what environment it comes from.”

With generative-AI models, Zemel says, evaluating robustness becomes even more difficult. In practice, the most common machine learning model has, until recently, been the classifier, which outputs the probabilities that a given input belongs to each of several classes. One way to gauge a classifier’s robustness is to determine whether its predicted probabilities — its confidence in its classifications — accurately reflects its performance on data. If the model is overconfident, it probably won’t generalize well to new settings.

But with generative AI models, there’s no such confidence metric to appeal to.

“If now the system is busy writing sentences, what does the uncertainty mean?” Zemel asks. “How do you talk about uncertainty? The whole question about building robust, properly confident, responsible systemsbecomes that much harder in the in the era where generative models are actually working well.”

The neural analogy

NeurIPS was first held in 1986, and in the early years, the conference was as much about neuroscientists using computational tools to model the brain as about computer scientists using brain-like models to do computation.

“The neural part of it has been drowned out by the engineering side of things,” Zemel says, “but there’s always been a lively interest in it. And there’s been some loose — and not so loose — inspiration that has gone that way.”

Today’s generative-AI models, for instance, are usually transformer models, whose signature component is the attention mechanism that decides which aspects of the input to focus on when generating outputs.

Amazon Scholars Michael I. Jordan and Michael Kearns and Amazon distinguished scientist Bernhard Scholkopf NeurIPS Amazon Science.jpg

Related content

NeurIPS luminaries on the future of AI

“Some of that work actually has its roots in cognitive science and to some extent in neuroscience,” Zemel says. “Neuroscience and cognitive science have studied attention for a long time now, particularly spatial attention: what do you focus on when viewing a scene? We have also been considering spatial attention in our models. About a decade ago, we were working on image captioning, and the idea was that when the system was generating the text for the caption, you could see what part of the image it was attending to. When it was entering the next word, it was focusing on some part of the image.

“It’s a little different than the attention in the transformers, where they took it a step further, as one layer can attend to activities in another layer of a network. It’s a similar idea, but it was a natural deep-learning version — learning applied to that same idea.”

Recently, Zemel says, computer scientists seem to be showing a renewed interest in what neuroscience and cognitive science have to teach them.

“I think it’s coming back as people try to scale up the systems and make them work with less data, or as the models become bigger and bigger, and it’s very inefficient and sometimes impossible to back-propagate through the whole system,” he says. “Brains have interesting structure at different scales. There are different kinds of neurons that have different functions, and we don’t really have that in our neural nets. And there’s no clear place where there’s short-term memory and long-term memory that are thought to be important parts of the brain. Maybe there are ways of getting that kind of architectural scaffolding structure that could be useful in improving neural nets and improving machine learning.”

New frontiers

As Zemel considers the future of deep learning, two areas of research strike him as particularly intriguing.

Mike Jordan.jpg

Related content

ICASSP: Michael I. Jordan’s “alternative view on AI”

“One of them is this area called mechanistic interpretability,” he says. “Can you both understand and affect what’s going on inside these systems? One way of demonstrating that you understand what’s going on is to make some change and predict what that change is. I’m not talking about understanding what a particular unit or a particular neuron does. It’s more like, we’d like to be able to make this change to the generative model; how do we achieve that without adding new data or post hoc processing? Can you actually go in and change how the network behaves?

“The other one is this idea that we talked about: can we add inductive biases, add structure to the system, add some sort of knowledge — it could be a logic, it could be a probability —to enable these systems to become much more efficient, to learn with less data, with less energy? There are just so many problems that are now open and unsolved that I think it’s a great time to be doing research in the area.”

Article link: https://www.amazon.science/blog/at-neurips-whats-old-is-new-again?

Share this:

  • Click to share on X (Opens in new window) X
  • Click to share on Facebook (Opens in new window) Facebook
  • Click to share on LinkedIn (Opens in new window) LinkedIn
Like Loading...

Related

Posts navigation

← The top 10 MIT Sloan articles of 2023
Generative AI research from MIT Sloan →
  • Search site

  • Follow healthcarereimagined on WordPress.com
  • Recent Posts

    • Hype Correction – MIT Technology Review 12/15/2025
    • Semantic Collapse – NeurIPS 2025 12/12/2025
    • The arrhythmia of our current age – MIT Technology Review 12/11/2025
    • AI: The Metabolic Mirage 12/09/2025
    • When it all comes crashing down: The aftermath of the AI boom – Bulletin of the Atomic Scientists 12/05/2025
    • Why Digital Transformation—And AI—Demands Systems Thinking – Forbes 12/02/2025
    • How artificial intelligence impacts the US labor market – MIT Sloan 12/01/2025
    • Will quantum computing be chemistry’s next AI? 12/01/2025
    • Ontology is having its moment. 11/28/2025
    • Disconnected Systems Lead to Disconnected Care 11/26/2025
  • Categories

    • Accountable Care Organizations
    • ACOs
    • AHRQ
    • American Board of Internal Medicine
    • Big Data
    • Blue Button
    • Board Certification
    • Cancer Treatment
    • Data Science
    • Digital Services Playbook
    • DoD
    • EHR Interoperability
    • EHR Usability
    • Emergency Medicine
    • FDA
    • FDASIA
    • GAO Reports
    • Genetic Data
    • Genetic Research
    • Genomic Data
    • Global Standards
    • Health Care Costs
    • Health Care Economics
    • Health IT adoption
    • Health Outcomes
    • Healthcare Delivery
    • Healthcare Informatics
    • Healthcare Outcomes
    • Healthcare Security
    • Helathcare Delivery
    • HHS
    • HIPAA
    • ICD-10
    • Innovation
    • Integrated Electronic Health Records
    • IT Acquisition
    • JASONS
    • Lab Report Access
    • Military Health System Reform
    • Mobile Health
    • Mobile Healthcare
    • National Health IT System
    • NSF
    • ONC Reports to Congress
    • Oncology
    • Open Data
    • Patient Centered Medical Home
    • Patient Portals
    • PCMH
    • Precision Medicine
    • Primary Care
    • Public Health
    • Quadruple Aim
    • Quality Measures
    • Rehab Medicine
    • TechFAR Handbook
    • Triple Aim
    • U.S. Air Force Medicine
    • U.S. Army
    • U.S. Army Medicine
    • U.S. Navy Medicine
    • U.S. Surgeon General
    • Uncategorized
    • Value-based Care
    • Veterans Affairs
    • Warrior Transistion Units
    • XPRIZE
  • Archives

    • December 2025 (8)
    • November 2025 (9)
    • October 2025 (10)
    • September 2025 (4)
    • August 2025 (7)
    • July 2025 (2)
    • June 2025 (9)
    • May 2025 (4)
    • April 2025 (11)
    • March 2025 (11)
    • February 2025 (10)
    • January 2025 (12)
    • December 2024 (12)
    • November 2024 (7)
    • October 2024 (5)
    • September 2024 (9)
    • August 2024 (10)
    • July 2024 (13)
    • June 2024 (18)
    • May 2024 (10)
    • April 2024 (19)
    • March 2024 (35)
    • February 2024 (23)
    • January 2024 (16)
    • December 2023 (22)
    • November 2023 (38)
    • October 2023 (24)
    • September 2023 (24)
    • August 2023 (34)
    • July 2023 (33)
    • June 2023 (30)
    • May 2023 (35)
    • April 2023 (30)
    • March 2023 (30)
    • February 2023 (15)
    • January 2023 (17)
    • December 2022 (10)
    • November 2022 (7)
    • October 2022 (22)
    • September 2022 (16)
    • August 2022 (33)
    • July 2022 (28)
    • June 2022 (42)
    • May 2022 (53)
    • April 2022 (35)
    • March 2022 (37)
    • February 2022 (21)
    • January 2022 (28)
    • December 2021 (23)
    • November 2021 (12)
    • October 2021 (10)
    • September 2021 (4)
    • August 2021 (4)
    • July 2021 (4)
    • May 2021 (3)
    • April 2021 (1)
    • March 2021 (2)
    • February 2021 (1)
    • January 2021 (4)
    • December 2020 (7)
    • November 2020 (2)
    • October 2020 (4)
    • September 2020 (7)
    • August 2020 (11)
    • July 2020 (3)
    • June 2020 (5)
    • April 2020 (3)
    • March 2020 (1)
    • February 2020 (1)
    • January 2020 (2)
    • December 2019 (2)
    • November 2019 (1)
    • September 2019 (4)
    • August 2019 (3)
    • July 2019 (5)
    • June 2019 (10)
    • May 2019 (8)
    • April 2019 (6)
    • March 2019 (7)
    • February 2019 (17)
    • January 2019 (14)
    • December 2018 (10)
    • November 2018 (20)
    • October 2018 (14)
    • September 2018 (27)
    • August 2018 (19)
    • July 2018 (16)
    • June 2018 (18)
    • May 2018 (28)
    • April 2018 (3)
    • March 2018 (11)
    • February 2018 (5)
    • January 2018 (10)
    • December 2017 (20)
    • November 2017 (30)
    • October 2017 (33)
    • September 2017 (11)
    • August 2017 (13)
    • July 2017 (9)
    • June 2017 (8)
    • May 2017 (9)
    • April 2017 (4)
    • March 2017 (12)
    • December 2016 (3)
    • September 2016 (4)
    • August 2016 (1)
    • July 2016 (7)
    • June 2016 (7)
    • April 2016 (4)
    • March 2016 (7)
    • February 2016 (1)
    • January 2016 (3)
    • November 2015 (3)
    • October 2015 (2)
    • September 2015 (9)
    • August 2015 (6)
    • June 2015 (5)
    • May 2015 (6)
    • April 2015 (3)
    • March 2015 (16)
    • February 2015 (10)
    • January 2015 (16)
    • December 2014 (9)
    • November 2014 (7)
    • October 2014 (21)
    • September 2014 (8)
    • August 2014 (9)
    • July 2014 (7)
    • June 2014 (5)
    • May 2014 (8)
    • April 2014 (19)
    • March 2014 (8)
    • February 2014 (9)
    • January 2014 (31)
    • December 2013 (23)
    • November 2013 (48)
    • October 2013 (25)
  • Tags

    Business Defense Department Department of Veterans Affairs EHealth EHR Electronic health record Food and Drug Administration Health Health informatics Health Information Exchange Health information technology Health system HIE Hospital IBM Mayo Clinic Medicare Medicine Military Health System Patient Patient portal Patient Protection and Affordable Care Act United States United States Department of Defense United States Department of Veterans Affairs
  • Upcoming Events

Blog at WordPress.com.
  • Reblog
  • Subscribe Subscribed
    • healthcarereimagined
    • Join 154 other subscribers
    • Already have a WordPress.com account? Log in now.
    • healthcarereimagined
    • Subscribe Subscribed
    • Sign up
    • Log in
    • Copy shortlink
    • Report this content
    • View post in Reader
    • Manage subscriptions
    • Collapse this bar
 

Loading Comments...
 

    %d