Sign up to save this job, get alerts, and apply with an optimized CV.
Data Engineer with GCP m f x
Job description
Do you want to develop your skills in cloud technologies and work with real data? Join our Data & Analytics team, where we build and develop solutions based on GCP. Work with experts, develop your career in Data Engineering, Big Data, or Machine Learning, and have a real impact on projects. Your Tasks: * Designing, implementing, and maintaining scalable data pipelines based on Google Cloud Platform. * Working with BigQuery as the main data warehouse: data modeling, query and cost optimization, ensuring performance and reliability of solutions. * Integrating data from various sources (files, databases, APIs, events) and processing and transforming it. * Orchestrating data workflows using Apache Airflow / Cloud Composer. * Creating and maintaining CI/CD solutions for data pipelines and infrastructure. * Managing cloud infrastructure using the Infrastructure as Code approach (Terraform). * Ensuring data quality, pipeline monitoring, and rapid incident response. * Collaborating with analytical, BI, and product teams to deliver stable and well-documented data. * Participating in the development of data architecture and jointly defining data engineering best practices. Requirements: * Minimum 4 years of experience as a Data Engineer or in a similar role working with data in a production environment. * Very good knowledge of Google Cloud Platform, especially: BigQuery (data modeling, query optimization) and Cloud Storage. * Ability to design, build, and maintain data pipelines (batch and/or streaming). * Very good knowledge of SQL and Python in the context of data processing and orchestration. * Experience in workflow orchestration (Apache Airflow / Cloud Composer). * Practice in implementing CI/CD for data solutions, e.g., GitHub Actions, GitLab CI, Cloud Build. * Knowledge of the Infrastructure as Code approach, specifically Terraform. * Previous work with large data volumes, considering the performance and reliability of solutions. * Fluent communication in English. * Requirement to be located in Poland and fluent knowledge of the Polish language. Nice to have: * Practical experience in streaming data processing (e.g., Dataflow / Apache Beam, Pub/Sub). * Proficiency in Apache Spark / PySpark when working with large data volumes. * Competencies in data transformation and modeling using tools such as dbt. * Ability to work with diverse data platforms (e.g., Databricks, Snowflake, MS Fabric). * Familiarity with tools and best practices in Data Governance, Data Lineage, and Data Quality.
Required skills
databricks
snowflake
english
ci/cd
monitoring
incident response
reliability
polish
sql
data modeling
python
github actions
big data
data engineer
gcp
google cloud platform
terraform
gitlab ci
dbt
machine learning
data governance
data warehouse
infrastructure as code
data quality
data engineering
pyspark
performance
data integration
data architecture
data pipelines
cost optimization
data processing
cloud storage
bigquery
bi
apache airflow
data transformation
product teams
apache beam
cloud composer
pub/sub
dataflow
cloud build
query optimization
apache spark
data & analytics
batch processing
data lineage
workflow orchestration
ms fabric
large data volumes
streaming processing
Sign up to apply
Create a free account to apply for this job and get access to:
- AI-powered CV optimization for this specific job
- Save jobs and create custom alerts
- See your CV match score for each job
Company information
- Company
- Sii
- Location
-
Polska, mazowieckie, Warszawa
Poland - Posted
- 2 weeks ago
Interested in this position?
Create your free account and tailor your CV to match this job.