Single-model uncertainty quantification in neural network potentials does not consistently outperform model ensembles


JSON Export

{
  "revision": 14, 
  "id": "1696", 
  "created": "2023-03-15T16:55:05.205326+00:00", 
  "metadata": {
    "doi": "10.24435/materialscloud:55-sd", 
    "status": "published", 
    "title": "Single-model uncertainty quantification in neural network potentials does not consistently outperform model ensembles", 
    "mcid": "2023.73", 
    "license_addendum": null, 
    "_files": [
      {
        "description": "Training dataset for silica", 
        "key": "silica_train.xyz", 
        "size": 117051475, 
        "checksum": "md5:d1e53f5238e084691cec2f23d485f154"
      }, 
      {
        "description": "Testing dataset for silica", 
        "key": "silica_test.xyz", 
        "size": 29593592, 
        "checksum": "md5:3996912931795a2a3245e2cbb4e4cade"
      }, 
      {
        "description": "Training dataset for ammonia", 
        "key": "ammonia_train.xyz", 
        "size": 38493, 
        "checksum": "md5:4e24ab2ab0d26246a5fff157182d76e1"
      }, 
      {
        "description": "Testing dataset for ammonia", 
        "key": "ammonia_test.xyz", 
        "size": 98014, 
        "checksum": "md5:7bb78eefb69c724dc7276f916c1eb70c"
      }, 
      {
        "description": "Description of the files and units", 
        "key": "README.md", 
        "size": 2187, 
        "checksum": "md5:599dd5689f3b96ac856db3120bf33dae"
      }
    ], 
    "owner": 585, 
    "_oai": {
      "id": "oai:materialscloud.org:1696"
    }, 
    "keywords": [
      "Uncertainty quantification", 
      "neural network interatomic potentials", 
      "single deterministic neural networks", 
      "adversarial sampling", 
      "silica glass", 
      "ammonia"
    ], 
    "conceptrecid": "1695", 
    "is_last": false, 
    "references": [
      {
        "type": "Preprint", 
        "doi": "10.48550/arXiv.2305.01754", 
        "url": "https://doi.org/10.48550/arXiv.2305.01754", 
        "citation": "A. R. Tan, S. Urata, S. Goldman, J. C. B. Dietschreit, R. Gomez-Bombarelli, arXiv:2305.01754 (2023)"
      }
    ], 
    "publication_date": "May 04, 2023, 16:18:10", 
    "license": "Creative Commons Attribution 4.0 International", 
    "id": "1696", 
    "description": "Neural networks (NNs) often assign high confidence to their predictions, even for points far out-of-distribution, making uncertainty quantification (UQ) a challenge. When they are employed to model interatomic potentials in materials systems, this problem leads to unphysical structures that disrupt simulations, or to biased statistics and dynamics that do not reflect the true physics. Differentiable UQ techniques can find new informative data and drive active learning loops for robust potentials. However, a variety of UQ techniques, including newly developed ones, exist for atomistic simulations and there are no clear guidelines for which are most effective or suitable for a given case. In this work, we examine multiple UQ schemes for improving the robustness of NN interatomic potentials (NNIPs) through active learning. In particular, we compare incumbent ensemble-based methods against strategies that use single, deterministic NNs: mean-variance estimation, deep evidential regression, and Gaussian mixture models. We explore three datasets ranging from in-domain interpolative learning to more extrapolative out-of-domain generalization challenges: rMD17, ammonia inversion, and bulk silica glass. Performance is measured across multiple metrics relating model error to uncertainty. Our experiments show that none of the methods consistently outperformed each other across the various metrics. Ensembling remained better at generalization and for NNIP robustness;  MVE only proved effective for in-domain interpolation, while GMM was better out-of-domain; and evidential regression, despite its promise, was not the preferable alternative in any of the cases. More broadly, cost-effective, single deterministic models cannot yet consistently match or outperform ensembling for uncertainty quantification in NNIPs.", 
    "version": 1, 
    "contributors": [
      {
        "email": "atan14@mit.edu", 
        "affiliations": [
          "Department of Materials Science and Engineering, Massachusetts Institute of Technology (MIT), Cambridge, Massachusetts, United States of America"
        ], 
        "familyname": "Tan", 
        "givennames": "Aik Rui"
      }, 
      {
        "email": "shingo.urata@agc.com", 
        "affiliations": [
          "Innovative Technology Laboratories, AGC Inc., Yokohama, Japan"
        ], 
        "familyname": "Urata", 
        "givennames": "Shingo"
      }, 
      {
        "email": "samlg@mit.edu", 
        "affiliations": [
          "Computational and Systems Biology, Massachusetts Institute of Technology (MIT), Cambridge, Massachusetts, United States of America"
        ], 
        "familyname": "Goldman", 
        "givennames": "Samuel"
      }, 
      {
        "email": "jdiet@mit.edu", 
        "affiliations": [
          "Department of Materials Science and Engineering, Massachusetts Institute of Technology (MIT), Cambridge, Massachusetts, United States of America"
        ], 
        "familyname": "Dietschreit", 
        "givennames": "Johannes C. B."
      }, 
      {
        "email": "rafagb@mit.edu", 
        "affiliations": [
          "Department of Materials Science and Engineering, Massachusetts Institute of Technology (MIT), Cambridge, Massachusetts, United States of America"
        ], 
        "familyname": "G\u00f3mez-Bombarelli", 
        "givennames": "Rafael"
      }
    ], 
    "edited_by": 576
  }, 
  "updated": "2023-11-21T08:54:41.689323+00:00"
}