IN047-08
AMP: An Automated Metadata Pipeline
Abstract:
The Automated Metadata Pipeline (AMP) project seeks to develop a fully-automated metadata pipeline that integrates machine learning and ontologies to generate syntactically and semantically consistent metadata recordsthat advance FAIR objectives and support Earth science research for a diverse group of stakeholders ranging from scientists to policy makers. AMP uses machine learning techniques to auto-generate semantically consistent, variable-level metadata records for NASA data products and, in collaboration with the ARtificial Intelligence for Ecosystem Services (ARIES) developer and user communities, we demonstrate the value of robust, semantically consistent metadata in addressing usability and scalability issues for data providers and metadata curators; improving the semantic interoperability of NASA data products; and demonstrating the benefits of semantically interoperable, FAIR data across communities of practice.