AI SummaryVerified by Aipplify AI
The vacancy is well-defined but lacks compensation details, affecting overall attractiveness.
AI quality score6.3 / 10
Check Match โ Just drop your CV
See your fit for Data Engineer in seconds.
Overview
Wildberries is looking for a Data Engineer to ensure the stable and secure operation of their Data Platform using Trino, Spark, S3, and Apache Iceberg.
Responsibilities
- โขEnsure stable, productive, and secure operation of the Data Platform based on Trino, Spark, S3, and Apache Iceberg, including administration and management of the access role model, documenting changes in the project.
- โขConfigure, update, monitor, and tune Trino clusters.
- โขSet up connectors (Iceberg, S3).
- โขOptimize query performance (resource groups, query analysis).
- โขOptimize Iceberg performance (partitioning, clustering, metadata management).
- โขDevelop and implement centralized role models for data and resource access on the platform.
Requirements
- โขUnderstanding of the interaction between Spark, Iceberg, and S3.
- โขExperience in operating Apache Iceberg (administration of Iceberg format tables, configuration and use of Hive Metastore).
- โขUnderstanding and application of: compaction, expiration snapshots, time travel, schema evolution.
- โขSkills in Linux, Bash, Python for automation.
- โขExperience managing access policies and permissions through Ranger in S3 and Iceberg.
- โขExperience in developing and implementing a centralized role model for data and resource access on the platform.
- โขExperience in administering Greenplum or ClickHouse (installation, configuration, optimization, integration with S3/Iceberg).
Loading similar jobs...