- Build architecture, develop big data mining tools, process, transform big data and manage big data (Big Data).
transform big data and manage big data (Big Data).
- Deploy and develop applications, data processing models in the big data ecosystem (Big Data) (Hadoop, Spark, Kafka,…) to ensure provision of a unified, stable and highly efficient data storage, processing and mining platform.
a unified, stable and highly efficient data storage, processing and mining platform.
- Research and design table systems, storage areas, data pipelines, compression standards and data tiering… to serve data provision for mining and analysis needs of requirements and deployment projects on Big Data infrastructure.
data tiering… to serve data provision for mining and analysis needs of requirements and deployment projects on Big Data infrastructure.
- Understand, analyze, evaluate and process semi-structured and unstructured data sources; build data connection solutions that are logically correct, fast and stable.
stable.
- Research, develop and plan synchronization and real-time (near realtime) data processing systems, data streaming on Oracle, Kafka Streams,… platforms to serve real-time data mining needs.
Kafka Streams,… platforms to serve real-time data mining needs.
- Research, develop and plan cloud systems (Google Cloud) to serve data storage and processing to reduce load on on-premise systems.
on on-premise systems.
- Evaluate technical solutions as well as architecture of data processing flows to ensure meeting requirements for performance, high availability and scalability.
scalability.
- Build and update documents related to design, deployment, development and operation of big data storage and processing systems