Skip to main content
v2026.11,772 entries · CC-BY 4.0

Direct comparison

NOMAD vs Materials Data Facility

Compare NOMAD and Materials Data Facility: governance, hosted data types, scale, and which materials-science repository fits your project.

Written and maintained by CASRAI Editorial Board

Last updated

Ask CASRAI · included with Regulatory Radar

Ask about NOMAD vs Materials Data Facility

Ask CASRAI answers research-administration questions and cites the passages behind every claim — and says so when the corpus does not cover something, instead of guessing. It comes with a Regulatory Radar subscription at $29 a month, alongside the daily digest of regulatory changes and the dashboard of what changed.

150 questions a day, on this site, over the API, or inside your own tools through the CASRAI MCP server.

Everything CASRAI publishes — this page, the dictionary, the guides and the news — stays free to read, with no account and no card.

How do NOMAD, Materials Data Facility compare side by side?

The table below compares NOMAD, Materials Data Facility across 11 procurement-relevant dimensions, from what it defines through relationship to mgi.

Side-by-side comparison

DimensionNOMADMaterials Data Facility
What it definesA repository and processing platform for computational materials-science data (simulation input/output)A general publication and discovery service for structured materials datasets, computational or experimental
GovernanceFAIR-DI e.V. (German non-profit association); development led by FAIRmat, an NFDI/DFG-funded consortiumGlobus Labs (University of Chicago Dept. of Computer Science + Argonne National Laboratory), via the Computation Institute
HostingMax Planck Computing and Data Facility (MPCDF), Garching, GermanyMulti-petabyte storage at Argonne National Laboratory and NCSA (UIUC), transfer via Globus
Origin2013-2014, Humboldt-Universität zu Berlin and the Fritz Haber Institute of the Max Planck SocietyLaunched in support of the US Materials Genome Initiative, with CHiMaD (NIST-funded center) as a founding partner
Primary data typeRaw/derived computational data: DFT, GW, molecular dynamics, and other simulation outputAny structured materials dataset — computational, experimental, or mixed
Code-aware parsingYes — automatically parses and normalizes files from 60+ simulation codes into a common schemaNo dedicated code parsing — relies on depositor-supplied metadata
Scale (indicative)On the order of 19 million uploaded entries covering several million distinct materials (grows continuously)Thousands of published datasets; scale reported per-dataset rather than per-entry
Search strengthDeep property-level search across normalized simulation output; OPTIMADE API for cross-database structure queriesDataset-level discovery driven by depositor metadata
Large-file / HPC transferStandard web upload/APIGlobus-based, built for high-throughput, resumable transfer of large, high-file-count datasets
Typical best fitDepositing/searching raw simulation output; training ML models on normalized dataPublishing a citable, DOI'd, mixed computational+experimental dataset; moving large datasets to/from HPC
Relationship to MGIIndependent European effort, frequently cited alongside MGI databases as a parallel FAIR-for-materials initiativeBuilt directly with Materials Genome Initiative funding and CHiMaD partnership

Common questions

Common questions about NOMAD vs Materials Data Facility

Can I use both NOMAD and MDF for the same project?

+

Yes, and it's common practice. Many researchers deposit raw simulation output in NOMAD for its automatic parsing and cross-code search, then publish a curated, DOI'd dataset bundle (which may reference the NOMAD entries) through MDF for citation in a paper's data-availability statement.

Which one should I list in a Data Management Plan for computational materials-science work?

+

If your output is primarily raw simulation files from a supported code, list NOMAD as the primary repository. If you need a single DOI for a mixed or experimental dataset, or you're already using Globus for transfer, list MDF. Both are indexed in re3data and are recognized domain repositories for materials science.

Are NOMAD and MDF free to use?

+

Yes, both are free for researchers to deposit and publish data. NOMAD is funded through FAIR-DI e.V./FAIRmat and German research-infrastructure funding; MDF runs on Globus's non-profit infrastructure, which offers free core services for academic use.

Are NOMAD and MDF interoperable?

+

They are not merged into a single system, but both participate in the broader FAIR-materials-data ecosystem — NOMAD exposes an OPTIMADE-compliant API used by several other structure databases, and both repositories are indexed in general discovery layers such as re3data. Cross-referencing a NOMAD entry ID from an MDF-published dataset record (or vice versa) in a paper's data-availability statement is common practice rather than a formal integration.

Referenced across the research world

University of Cambridge logoColumbia University logoCrossref logoUniversity of Edinburgh logoHarvard University logoUniversity of Oxford logoPrinceton University logoStanford School of Medicine logoUniversity College London logoORCID logoUniversity of Cambridge logoColumbia University logoCrossref logoUniversity of Edinburgh logoHarvard University logoUniversity of Oxford logoPrinceton University logoStanford School of Medicine logoUniversity College London logoORCID logo
  • University of Cambridge logo
  • Columbia University logo
  • Crossref logo
  • University of Edinburgh logo
  • Harvard University logo
  • University of Oxford logo
  • Princeton University logo
  • Stanford School of Medicine logo
  • University College London logo
  • ORCID logo

View CASRAI adoption →

Regulatory Radar

Stop finding out after the fact

$29/month, cancel anytime. Daily digest updates from our analysis, a dashboard holding the same items, and a cited assistant for everything they raise.

  • Federal Register, Federal Register+, Grants.gov, Regulations.gov, NSF News, UKRI, plus CASRAI’s own published content.
  • 72,264 indexed passages, and every answer cites the ones it drew on.