Towards the Implementation of ONCA: A Generic, Scalable, and Massive Data Processing Platform for Information Discovery and Analytics
摘要
Analysis and visualization are fundamental components of the data-driven decision-making process. In the health and environment fields, various platforms exist for massive data processing to generate information products that assist decision-makers in crafting public policies. These policies are informed by observed data trends to mitigate potential epidemiological impacts on the population. However, most existing solutions focus on either storage, processing, or visualization of data separately, complicating the implementation of comprehensive analysis depending on the data domain. In this work, we present ONCA, a generic, scalable platform for massive data processing, designed to facilitate data analysis using a microservices architecture. ONCA integrates mechanisms for data processing, the creation of observatories, the publication of information products, and user queries, seamlessly automating the interconnection of these components. Designed as a distributed system for deployment on cloud environments, ONCA enables collaboration and information sharing among organizations. This paper details the proposal of the ONCA platform and presents preliminary results from generating information products using environmental data from the Mexican territory.