Wildberries

Data Engineer

6.0/10
Wildberries
Not specified
Remote
mid
about 2 hours ago
AI SummaryVerified by Aipplify AI

The vacancy is well-defined but lacks compensation details, affecting overall attractiveness.

AI quality score6.3 / 10

Check Match โ€” Just drop your CV

See your fit for Data Engineer in seconds.

Overview

Wildberries is looking for a Data Engineer to ensure the stable and secure operation of their Data Platform using Trino, Spark, S3, and Apache Iceberg.

Responsibilities

  • โ€ขEnsure stable, productive, and secure operation of the Data Platform based on Trino, Spark, S3, and Apache Iceberg, including administration and management of the access role model, documenting changes in the project.
  • โ€ขConfigure, update, monitor, and tune Trino clusters.
  • โ€ขSet up connectors (Iceberg, S3).
  • โ€ขOptimize query performance (resource groups, query analysis).
  • โ€ขOptimize Iceberg performance (partitioning, clustering, metadata management).
  • โ€ขDevelop and implement centralized role models for data and resource access on the platform.

Requirements

  • โ€ขUnderstanding of the interaction between Spark, Iceberg, and S3.
  • โ€ขExperience in operating Apache Iceberg (administration of Iceberg format tables, configuration and use of Hive Metastore).
  • โ€ขUnderstanding and application of: compaction, expiration snapshots, time travel, schema evolution.
  • โ€ขSkills in Linux, Bash, Python for automation.
  • โ€ขExperience managing access policies and permissions through Ranger in S3 and Iceberg.
  • โ€ขExperience in developing and implementing a centralized role model for data and resource access on the platform.
  • โ€ขExperience in administering Greenplum or ClickHouse (installation, configuration, optimization, integration with S3/Iceberg).
Loading similar jobs...