Site Reliability Engineer w Warszawa, Polska - Jobeax
Opis oferty
Site Reliability Engineer w Warszawa, Polska
KontraktUmowa czasowa lub freelancing
$2k USD
Poland, Warszawa
Site Reliability Engineer w Warszawa, Polska is listed on Jobeax. Browse 110,000+ vacancies available.
We are seeking a highly skilled and motivated Site Reliability Engineer (SRE) to join our team. In this critical role, you will collaborate closely with software developers and operations teams to ensure high reliability, scalability, and efficiency of our systems, with a strong focus on meeting and exceeding customer expectations. Your expertise will be crucial in deploying, maintaining, and automating our infrastructure and application environments to ensure seamless user experiences. Your proactive involvement will be key to enhancing system reliability, optimizing resource utilization, and ensuring continuous improvement in our operational practices. You will have the opportunity to lead the adoption of AI-enabled platform capabilities, such as generative AI (GenAI), autonomous agents, and AIOps, driving innovation and operational excellence. Your responsibilities will include defining and tracking Service Level Objectives (SLOs), managing error budgets, and reducing toil through automation. You will play a pivotal role in driving the success of technology initiatives, maximizing their impact across the organization, and ensuring that solutions consistently meet the high standards our customers expect. Responsibilities Collaborate with development, security, quality, and operation teams to implement SRE practices and ensure system reliabilityDefine and support required level of reliability, availability, and performance for services and applicationsDesign and deliver Cloud-based solutions tailored to client needsTroubleshoot, mitigate, and support fixing of the infrastructure and application issues in a timely mannerImplement a monitoring system for the infrastructure and application reliabilityGuide adoption of AI technologies on the platform (GenAI, AI agents, AIOps) to improve operationsCommunicate technical concepts clearly to both engineering teams and management stakeholders Requirements Bachelor’s degree in Computer Science, Engineering, or a related field 3+ years of hands-on experience in Site Reliability Engineering or related rolesProven experience in any cloud (AWS/GCP/Azure) Experience with implementing SRE practices such as SLO/SLI, Error budgets, Postmortems, Reducing Toil, capacity planning, and Incident ManagementPython or other scripting/programming language Strong background in monitoring tools Proficiency in CI/CD tools, infrastructure as code, and configuration managementSolid knowledge of container orchestration technologies (Kubernetes, Docker) English language proficiency at an Upper-Intermediate level (B2) or higher Nice to have Certification in Kubernetes, AWS/GCP/Azure, or similar technologiesProven experience in DevOpsExpertise in deployment and management of LLMs, including technologies like RAGKnowledge of managing and optimizing AI/ML models in production environments, including basic deployment, monitoring, and maintenanceNative AI cloud services: AWS Bedrock, Google Vertex AI, Azure AIExperience in designing, building, and operating AI agents and agentic frameworksCoding Agents: Claude Code, OpenCode, Cursor, Trae, Antigravity We offer We gather like-minded people:Top tech minds driving innovation in AI, cloud and digital platform modernizationSupportive team and agile, startup-like cultureHybrid by design mode and opportunity to work remotely within PolandChance to work abroad for up to 60 days annuallyBusiness-driven relocation opportunitiesWe provide growth opportunities:Career development programsThought leadership, mentoring, soft skills and well-being programsCertification (Anthropic, Gemini, GCP, Azure, AWS)English classesWe cover it all:Stable payParticipation in the Employee Stock Purchase Plan with a 15% discountBenefits package (health insurance, multisport, shopping vouchers)Referral bonuses up to $2,000Offices featuring entertainment and relaxation zones, table tennis and football, free snacks, coffee and moreCorporate, social and well-being eventsPlease, note:Benefits listed above are available to employees onlyWe are open for working with Contractors. Terms of B2B cooperation agreements are agreed individuallyWe will reach out to selected candidates exclusively EPAM is global leader in AI transformation engineering and integrated consulting, serving Forbes Global 2000 companies and ambitious startups. With over thirty years of expertise in custom software, product and platform engineering, we empower our clients to become AI-Native enterprises, driving measurable value from innovation and digital investments.
... Manager (m/f/d) in energy plant construction, you will be responsible for the successful execution of construction projects at domestic and international job sites in the field of energy infrastructure (HVDC, FACTS). You will lead the on-site construction team and ensure that construction work is carried out on schedule, ...
... improvements in tools, frameworks, test coverage and workflows, Propose innovations using AI, automation or data insights to increase efficiency and product reliability. Cross‑Team Collaboration - work closely with software engineers, product engineers, data/analytics teams, partners and contribute to requirement reviews, ...
... response, root cause analysis, and post-incident review processes Requirements Minimum 7 years of professional experience in software development, DevOps, and/or Site Reliability EngineeringMinimum 3 years of experience building and maintaining CI/CD pipelines and SRE automation within cloud environments at scaleExperience ...
... the core Azure Databricks data platform and infrastructure that enables analytics, reporting and AI solutions across Arla. As an Experienced Data Platform Engineer, you will play a key role in shaping the technical direction of the team, mentoring colleagues and ensuring we continue to deliver a modern, enterprise-grade ...
... the next generation of cloud server architecture? Do you enjoy hands-on bring-up and deep dives into data plane behavior? Join our Server Team! Our Software Engineering team designs, qualifies, and scales the server platforms that power Akamai Cloud. We partner closely with OEM/ODM manufacturers to deliver reliable, efficient, ...
... features, from technical design and scoping through implementation, launch, and production support., Collaborate closely with product managers, researchers, Site Reliability Engineers (SREs), and other engineers to understand requirements, translate them into technical specifications, and deliver robust and scalable solutions., ...
... payments. - Support study specific audit activities related to investigator payments when required. - Resolve site and study level payment escalations to maintain site satisfaction and operational continuity. - Ensure appropriate contracts, work orders, and funding are in place to support the execution of site payments. Candidate ...
... cross-cloud exposure), Databricks, Spark, robust CI/CD, and high-impact data solutions used across international organizations., , You'll join a team of passionate engineers who value technical excellence, collaboration, and continuous growth. Team size, 4-7 This is how we work, in house, at the client's site, you focus on a single ...
... teams to embed reliability and operational excellence across the software development lifecycle. What we are looking for 10+ years of experience in Software Engineering, DevOps, Site Reliability Engineering, or Infrastructure Engineering. Strong hands-on experience with AWS and cloud infrastructure design, operations, and ...
Site Reliability Engineer Warsaw | Up to 40% working from home possible | Reference 8133 We’re looking for a seasoned DevOps/SRE professional to drive automation, reliability, and operational excellence across our platforms. You will own CI/CD pipelines, release processes, incident response, and collaborate closely with ...
... abilities, coding expertise, urgent issue resolution, and dedication to maintaining service availability are essential for success in this role. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling robust programmatic tooling and infrastructure-as-code utilities in Python to eliminate ...
... improves the reliability of our systems. You'll work with various technologies as we release brand new applications and modernize our existing tooling. As a Senior Site Reliability Engineer, you will be: - Providing support and mentorship for other SRE engineers within the team - Deploying and maintaining the platform and tools ...
... które inspirują ludzi do poznawania świata i wspólnie zamieniamy te inspiracje w doświadczenia., , Do obszaru infrastruktury w eSky Group chcielibyśmy zaprosić Site Reliability Engineer’a. Razem z 3 SRE będziesz odpowiadać za zapewnienie niezawodności naszych rozwiązań poprzez budowanie i utrzymanie wysoko dostępnej infrastruktury ...
... overhead through automation. Over time, you’ll have opportunities to influence platform architecture, lead reliability initiatives, and establish Site Reliability Engineering best practices across the organization. Responsibilities - Serve as the platform subject matter expert, mentoring engineering teams on reliability, scalability, ...
STN Inc is seeking a Site Reliability Engineer to own reliability, observability, and incident response for the GPU One platform. You will define SLOs, build the observability stack, and lead major incidents to resolution. The role requires strong SRE/DevOps background, proficiency in Go or Python, and hands-on Kubernetes ...
Site Reliability Engineer (SRE) DevOps Support at Citi is at the cusp of a major transformation. The production support team is transitioning to a complete Site Reliability Engineering (SRE) model, leveraging established SRE principles and the latest technologies in Generative AI. We are adopting agile methodologies to ...
... across hundreds of servers and ~50 services. Drive observability in cooperation with development team. Establish and enforce IaC practices, CI/CD pipeline reliability, and change management processes. Participate in the on‑call rotation alongside backend developers. Respond to and lead incident resolution, run post‑mortems ...
Site Reliability Engineer (SRE) DevOps Support at Citi is at the cusp of a major transformation. The production support team is transitioning to a complete Site Reliability Engineering (SRE) model, leveraging established SRE principles and the latest technologies in Generative AI. We are adopting agile methodologies to ...
... Job Description As a Software Engineer on the Emergency Call Management site reliability engineering (ECM-SRE) team you will join a team of talented software engineers who work directly with product and engineering teams to constantly improve reliability across our suite of public safety products. Your responsibilities will ...