Analysis and visualization are fundamental components of the data-driven decision-making process. In the health and environment fields, various platforms exist for massive data processing to generate information products that assist decision-makers in crafting public policies. These policies are informed by observed data trends to mitigate potential epidemiological impacts on the population. However, most existing solutions focus on either storage, processing, or visualization of data separately, complicating the implementation of comprehensive analysis depending on the data domain. In this work, we present ONCA, a generic, scalable platform for massive data processing, designed to facilitate data analysis using a microservices architecture. ONCA integrates mechanisms for data processing, the creation of observatories, the publication of information products, and user queries, seamlessly automating the interconnection of these components. Designed as a distributed system for deployment on cloud environments, ONCA enables collaboration and information sharing among organizations. This paper details the proposal of the ONCA platform and presents preliminary results from generating information products using environmental data from the Mexican territory.

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Towards the Implementation of ONCA: A Generic, Scalable, and Massive Data Processing Platform for Information Discovery and Analytics

  • Melesio Crespo-Sanchez,
  • Hugo G. Reyes-Anastacio,
  • J.L. Gonzalez-Compean,
  • Jaqueline Calderon-Hernadez,
  • Ignacio Castillo-Barrios

摘要

Analysis and visualization are fundamental components of the data-driven decision-making process. In the health and environment fields, various platforms exist for massive data processing to generate information products that assist decision-makers in crafting public policies. These policies are informed by observed data trends to mitigate potential epidemiological impacts on the population. However, most existing solutions focus on either storage, processing, or visualization of data separately, complicating the implementation of comprehensive analysis depending on the data domain. In this work, we present ONCA, a generic, scalable platform for massive data processing, designed to facilitate data analysis using a microservices architecture. ONCA integrates mechanisms for data processing, the creation of observatories, the publication of information products, and user queries, seamlessly automating the interconnection of these components. Designed as a distributed system for deployment on cloud environments, ONCA enables collaboration and information sharing among organizations. This paper details the proposal of the ONCA platform and presents preliminary results from generating information products using environmental data from the Mexican territory.