Data Engineer (Scala) - Data Assets w Kraków, Polska - Jobeax
Opis oferty
Data Engineer (Scala) - Data Assets w Kraków, Polska
ZdalniePraca z dowolnego miejsca
HybrydowoPołączenie biura i pracy zdalnej
KontraktUmowa czasowa lub freelancing
14600 - 20825 zł
Poland, Kraków
Data Engineer (Scala) - Data Assets w Kraków, Polska is listed on Jobeax. Browse 110,000+ vacancies available.
Data Engineer (Scala) - Data Assets Location: Warszawa, PL, 00-841 Kraków, PL, 31-503 Company: Allegro sp. z o.o. Team: Technology Contract Type: Employee Important things for you: Flexible working hours in the hybrid model (4/1) - working hours start between 7:00 a.m. and 10:00 a.m. We also have 30 days of occasional remote work. The salary range for this position depending on the skill set is as follows (contract of employment, tax-deductible cost): PLN gross 14 600 - 20 825 Annual bonus based on your annual performance and company results. Our team is based in Warsaw and Cracow. About the team: As part of the Data & AI area, we implement projects based on the practical "data science" and "artificial intelligence" applications of an unprecedented scale in Poland. Data & AI is a group of over 150 experienced engineers organized into over a dozen teams with various specializations. Some of them build dedicated tools for creating and launching BigData processes or implementing ML models for the entire organization. Others work closer to the client and are responsible for the implementation of the search engine, creating recommendations, building a buyer profile or developing an experimental platform. There are also research teams in the area whose aim is to find solutions to non-trivial problems requiring the use of machine learning. We are looking for Data Engineers who want to build highly scalable and fault-tolerant data processing for millions of Allegro customers. The platform processes (batch and realtime) 5 billion clickstream events every day (up to 150k / sec, ) from all Allegro sites and Allegro mobile applications. This is a hybrid solution using Google Cloud Platform (GCP) services like DataProc, GKE, Beam, BigQuery, Pubsub or Dataflow (scio). We are looking for people who: Are programming in languages such as Scala or Java, Python Have strong understanding of distributed systems, data storage and processing framework like dbt, Spark or Apache Beam Have knowledge of GCP (especially Dataproc, Dataflow and Composer) or other public cloud environments like Azure or AWS Use good practices (clean code, code review, TDD, CI/CD) Navigate efficiently within Unix/Linux systems Possess a positive attitude and team-working skills Are eager for personal development and keeping their knowledge up to date Know English at B2 level What's in it for you: Well-located offices (with e.g. fully equipped kitchens, bicycle parking, terraces full of greenery) and excellent work tools (e.g., raised desks, ergonomic chairs, interactive conference rooms). A 16" or 14" MacBook Pro or corresponding Dell with Windows (if you don't like Macs) and all the necessary accessories. A wide selection of fringe benefits in a cafeteria plan - you choose what you like (e.g., medical, sports or lunch packages, insurance, purchase vouchers). English classes that we pay for related to the specific nature of your job. A training budget, inter-team tourism (see more here), hackathons, and an internal learning platform where you will find multiple trainings. An additional day off for volunteering, which you can use alone, with a team, or with a larger group of people connected by a common goal. Social events for Allegro people - Spin Kilometers, Family Day, Fat Thursday, Advent of Code, and many other occasions we enjoy. And that's just the beginning! You can read more about the benefits here. #goodtobehere means that: You will join a team you can count on - we work with top-class specialists who have knowledge- and experience-sharing in their DNA. You will love our level of autonomy in team organization, the space for continuous development, and the opportunity to try new things. You get to choose which technology solves the problem and you are responsible for what you create. You will value our Developer Experience and the full platform of tools and technologies that make creating software easier. We rely on an internal ecosystem based on self-service and widely used tools such as Kubernetes, Docker, Consul, GitHub, and GitHub Actions. Thanks to this, you can contribute to Allegro from your very first days on the job. You will be equipped with modern AI tools to automate repetitive tasks, allowing you to focus on developing new services and refining existing ones (also leveraging AI support). You will create solutions that will be used (and loved!) by your friends, family and millions of our customers. You will meet the Allegro Scale, which starts with over 1000 microservices, an open-source data bus (Hermes) with 300K+ rps, a Service Mesh with 1M+ rps, tens of petabytes of data, and production-used machine learning. You will become part of Allegro Tech - We speak at industry conferences, cooperate with tech communities, run our own blog (it's been over 10 years!), record podcasts, lead guilds, and we organize our own internal conference - the Allegro Tech Meeting. We create solutions we love (and can) to talk about! Send us your CV and… see you at Allegro!
... cloud-native patternsBuild and optimize ETL/ELT workflows to support analytical and operational use casesImplement data modeling strategies aligned with modern data lakehouse architecturesEnsure data quality, governance, and monitoring across all data assetsDeliver secure, scalable, and highly available data solutions following ...
DCG Poland is looking for experienced professionals to develop distributed BigData processing pipelines. Candidates should have 7+ years of expertise in Spark and Scala, along with substantial experience in Python and Robot Framework. You will collaborate with DevOps and QA teams in an Agile environment, maintaining a focus ...
... will develop and maintain a data platform and reporting processes operating in a complex Big Data environment. Responsibilities: Design, develop and maintain Data Engineering solutions using Hadoop and Apache Spark. Build scalable solutions for processing large volumes of data. Analyse complex business requirements and ...
... roli Data Engineer lub podobnej. Doświadczenie komercyjne w pracy w środowisku AWS oraz technologiami Spark, Python i SQL. Praktyczna znajomość architektury Data Mesh oraz rozwiązań do przetwarzania dużych zbiorów danych. Doświadczenie w budowie skalowalnych i wysokowydajnych rozwiązań data engineeringowych w środowisku ...
... Engineering. Degree in computer science, engineering or equivalent technical field experience. Proficient in cloud-based data processing and storage technologies like Databricks, Spark, Airflow. Fluent in SQL and proficient in at least one modern programming language (e.g., Python, Scala, Java). Proactive self-starter with a proven ...
... in Python. Requirements: Higher education in Computer Science, Mathematics / Statistics, Telecommunications or related. Ability to work with large amounts of data in an operational environment. Experience in data engineering for AI/ML solutions hosted on-prem or in the cloud (preferred GCP). Experience in Data Warehousing ...
Responsibilities As a Data Engineer, you will design, build, and operate scalable data platforms and pipelines supporting analytics and machine learning use cases. You will transform raw data into reliable, high-quality data products and contribute to architectural and platform-level decisions. - Design and implement batch ...
Data Engineers (Mid & Senior) — Krakow (Hybrid) Are you looking to work on modern, high-throughput data platforms without the legacy drag? We are scaling a dedicated Data Engineering team in Krakow to drive a major enterprise cloud consolidation onto a modern GCP & Medallion (Bronze/Silver/Gold) architecture. If you thrive ...
Join Unit8 SA in Warsaw, Poland, as a proficient Software Engineer. You will design, build, and maintain data pipelines and contribute to critical data & AI projects. This role involves collaborating with engineers and clients, requiring strong Python skills and a problem-solving mindset. Enjoy a flexible work environment, ...
... dobre praktyki programistyczne. Współpraca z międzynarodowym, zwinnym zespołem. Job description For our client we are looking for an experienced Senior Data Engineer to join a project focused on building and developing scalable, cloud-based data solutions. Responsibilities Design, develop, and maintain scalable data pipelines. ...
Infrastructure Data Engineer Dołącz do globalnego zespołu naszego klienta pracującego nad zaawansowaną platformą analityczną w obszarze cyberbezpieczeństwa na stanowisko Infrastructure Data Engineer. Lokalizacja: praca w modelu hybrydowym 6 dni na miesiąc z biura w Krakowie Współpraca: B2B O projekcie Zespół Cybersecurity ...
... komercyjne na stanowisku Data Engineer, - pracowałeś z chmurą publiczną (AWS i/lub GCP i/lub Azure), - masz praktyczne doświadczenie z Apache Airflow, - znasz Databricks i ekosystem Spark, - swobodnie poruszasz się w SQL oraz jednym z języków programowania (Python/Scala), - potrafisz pracować samodzielnie w środowisku rozproszonym. ...
... business requirements into scalable technical solutions Proactive mindset with strong ownership and accountability Passion for automation, innovation, and engineering best practices Experience sharing knowledge through training, mentoring, and documentation Nice to Have Experience in product-oriented data platforms Exposure ...
... implementing scalable, production-grade data pipelines, data models, and data processing frameworks in cloud environments., Leading the technical design of cloud data architectures, contributing to standards, guidelines, and long-term platform strategy., Driving best practices in data engineering: quality assurance, lineage, ...
... to set the direction, •Review the work of other engineers and say clearly when something needs reworking, •Stay up-to-date with trends and new technologies in data engineering, , DevOps & Quality, •Work with Terraform, GitHub and CI/CD pipelines across the full deployment lifecycle, •Ensure high standards of data quality, ...
... resolving issues reported through organizational communication channels (Slack, email, OpsGenie) Job requirements At least 5 years of commercial experience in data engineering. Code in Python (65% of tasks in the team) and JVM languages (Java, Kotlin, Scala; 35%). Know about Google Cloud Platform (BigQuery, Composer, Dataproc, ...
... ubezpieczenie na życie Pełna odpowiedzialność za warstwę danych i infrastrukturę - realne decyzje architektoniczne Dostęp do konferencji i szkoleń z obszaru DevOps, data engineering i systemów rozproszonych Jak będzie wyglądać proces rekrutacji Aplikacja – po otrzymaniu CV wracamy z informacją zwrotną Rozmowa wstępna (HR) – wzajemne ...
... CDC. We are looking for a mid/senior Data Platform Engineer (backend) to join our team in Warsaw. Desired skills include: experience designing and maintaining data pipelines using ETL and/or ELT approaches, with the ability to reason about trade-offs rather than defaulting to a single pattern understanding of data pipeline ...
... into data engineering solutions. • Design, build, and maintain end-to-end data pipelines for ingestion, transformation, and processing using Azure services, Databricks, IICS, and Azure Synapse. • Develop advanced transformation logic in Databricks and IICS to support complex business use cases. • Knowledge of modern Data ...