Data preparation and analysis using Big Data technologies (Cloudera, Impala, Hive, Spark, Kafka, MinIO, SuperSet, etc.).
· Application of Big Data technologies (Cloudera, Spark, Databricks, Snowflake, etc.) and cloud services (Azure, AWS, GCP).
· Creating data structures, implementing data transfers, and collaborating on data integrations.
· Preparing data analyses and working with metadata (modeling and technical design) - SQL, NoSQL databases.
· Building data processing pipelines, deploying them to production, creating and monitoring reports.
· Contributing to data integration deliverables (data to and from analytical platforms and data systems - on-prem or cloud).
· Programming parallel and scalable applications for processing data structures.
· DevOps - code development, programming and debugging, deployment and code maintenance.
· Supporting deployment of various applications and tools.
· Implementing tasks across various economic sectors (banking, telco, manufacturing, retail, utilities, R&D, etc.).
· As opportunities arise, working on infrastructure and admin tasks focused on Cloudera/BD/cloud and integration of new components.
· Contributing to tasks focused on real-time data processing and visualization. Collaborating on the creation of data catalogs, data dictionaries, or metadata descriptions.
· Collaborating on the creation of data marts and data layers.
· Assisting with the preparation of customer presentations, primarily visualizing analysis outputs and interpreting findings.
· Supporting pre-sales activities.
· Tracking and studying new trends in clouds and Big Data.
· Sharing experiences and work results within the team. Later preparing training and workshops for customers and colleagues.
Meal allowance