... projektowej. Mile widziane, Doświadczenie w realizacji programów transformacyjnych związanych z architekturą danych i platformami chmurowymi., Znajomość zagadnień Data Governance, Data Management lub Data Mesh. To oferujemy, Długoterminowa współpraca, Szkolenia techniczne, certyfikaty i podnoszenie kwalifikacji, Mentoring Competence ...
... experience in areas like: - Strong proficiency in Python and AWS services - Development of APIs and microservices - Building and maintaining ETL/ELT pipelines - Database technologies, Data modeling and schema design (ideally MySQL) - Experience with continuous delivery toolchains, automated testing, build tools, and CI/CD ...
Senior Data Engineer (Databricks & Python) Miejsce pracy: Warszawa Technologies we use Expected - Python - PySpark - Spark SQL - Delta Lake About the project We are looking for an experienced Senior Data Engineer to join our Labeling Data Platform team. The platform supports the collection, processing, transformation, governance, ...
... performance tuning (PostgreSQL, MySQL, SQL Server, Oracle, etc.) - Data warehousing concepts and dimensional modeling - Data management fundamentals: data modeling, data quality, metadata management, data warehouse/lake patterns, distributed systems - Version control using Git and CI/CD practices - Data governance, data quality, ...
... platforms, Databricks Lakehouse architecture, and regulated data environments. This is how we work, at the client's site Your responsibilities, Develop and maintain data pipelines using Azure Databricks, Python, PySpark, Spark SQL, and Delta Lake., Design and implement solutions following the Bronze → Silver → Gold medallion architecture., ...
... performance tuning (PostgreSQL, MySQL, SQL Server, Oracle, etc.), Data warehousing concepts and dimensional modeling, Data management fundamentals: data modeling, data quality, metadata management, data warehouse/lake patterns, distributed systems, Version control using Git and CI/CD practices, Data governance, data quality, ...
... podejście streaming-first, a nie wyłącznie klasyczne ETL/batch processing,, •potrafi wpływać na kierunek techniczny wielu zespołów. Twój zakres obowiązków, Platforma Databricks, •Pełnienie roli Tech Leada dla rozwoju platformy Databricks Lakehouse., •Wsparcie zespołów w budowie produktów dataowych (best practices, wzorce, narzędzia)., ...
... data pipelines integrating data from diverse sources (databases, APIs, streaming platforms). Collaborate with cross‑functional teams including data engineers, data scientists, and cloud architects to deliver end‑to‑end data platform solutions. Engage at all critical stages of project development with business analysts, source ...
... Wrocławiu Stawka: do 130 zł/h netto na b2b Poszukujemy doświadczonego Project Managera do projektu związanego ze zmianą zasilania hurtowni danych oraz Cloud Data Lake danymi z nowo utworzonych domen biznesowych. Kluczowe jest doświadczenie w projektach migracji danych oraz budowy i zasilania hurtowni danych/CDL. Zakres ...
... za rozwój architektury danych oraz rozwiązań chmurowych opartych o GCP,, wykorzystywany stos technologiczny: Google Cloud Platform (GCP), BigQuery, Kubernetes, SQL, Data Lake, ETL/ELT, Big Data, Real-Time Processing, Feature Store,, lokalizacja projektu: Warszawa ( Centrum),, tryb realizacji usług: hybrydowy, z dużą elastycznością. ...
... and analysts, Build Python services and APIs (FastAPI) that expose data to downstream applications, and integrate with third-party and internal APIs, Develop data transformation logic in SQL and dbt on Snowflake, and build data models that support fast, reliable querying, Write unit, integration, and data-contract tests, ...
... processes for data pipelines and applications. Proficiency in Python and PySpark for data processing, transformation, and analysis. Demonstrated experience in using Databricks for data engineering tasks, including working with Delta Lake and Spark SQL. We offer P&G-sized projects and access to world leading IT partners and technologies ...
Job Description: Data Engineer, LDWH Migration Programme Level: Senior / Mid-Senior Main Responsibilities: Develop and maintain data pipelines using Azure Databricks. Implement ETL processes, ensuring data quality and integrity. Read and comprehend complex SQL stored procedures. Translate architecture direction into engineering ...
... połączeń z systemami poza ekosystem CRM • implementacja mechanizmów kontroli jakości oraz walidacji danych Wymagania • Minimum 4-6 lat doświadczenia na stanowisku Data Engineer, preferencyjnie w sektorze o dużej skali danych • doświadczenie w projektowaniu i modelowaniu struktur danych (hurtownie danych, DataLake-ów) na potrzeby ...
... roli Data Engineer lub podobnej. Doświadczenie komercyjne w pracy w środowisku AWS oraz technologiami Spark, Python i SQL. Praktyczna znajomość architektury Data Mesh oraz rozwiązań do przetwarzania dużych zbiorów danych. Doświadczenie w budowie skalowalnych i wysokowydajnych rozwiązań data engineeringowych w środowisku ...
... experience in Data Engineering or Software Engineering. Degree in computer science, engineering or equivalent technical field experience. Proficient in cloud-based data processing and storage technologies like Databricks, Spark, Airflow. Fluent in SQL and proficient in at least one modern programming language (e.g., Python, Scala, ...
... Warehousing and Business Intelligence. Good knowledge of common Python libraries and frameworks (Pandas, FastAPI, kubeflow pipelines, GCP libraries). Good knowledge of SQL databases, knowledge about noSQL would be an asset. Experience in working with Linux and Git version control system (GitLab). Knowledge of English at a level ...
Senior Data Engineer Miejsce pracy: Warszawa Technologies we use Expected - Google Cloud Platform - Terraform - Git - GitHub Actions - SQL - Python Optional - Docker - Kubernetes About the project We are looking for a Senior Data Engineer who will act as a technical mentor, define our architecture, and bring best practices ...
... keep reading. Your responsibilities Design & Build • Build and improve GCP data solutions using BigQuery, Cloud Composer, and Cloud Storage • Build and maintain data pipelines and transformations with Python, dbt, and Airflow • Model data across Data Warehouses, Lakehouses and Data Lakes within an established architecture ...
... consistency. What you will be doing : - Review clinical trial data captured in electronic data capture systems - Identify missing, inconsistent, or out-of-range data - Generate, track, and resolve data queries - Ensure data aligns with study protocols and data management plans. - Perform ongoing data quality checks and trend ...