Data Scientist - Python & Cloud Data Platforms (f/m/x) w Gdańsk, Polska
Sii Sp. z o.o.
Gdańsk, Polska
Data Scientist – Python & Cloud Data Platforms (f/m/x)
Miejsce pracy: Gdańsk
Technologies we use
Expected
Cloud Computing
Python
Data Platforms
XGBoost
Pandas
NumPy
Scikit-learn
Optional
Palantir Foundry
Snowflake
Microsoft Fabric
Natural Language Processing
LLM
MLflow
Power BI
Apache Spark
Time Series Forecasting
About the project
We’re looking for data scientists who turn ambiguous business questions into models and analyses that actually get used. You’ll frame the problem, wrangle the data, build and validate models, and — just as importantly — explain what the numbers mean to people who don’t speak Python. Your day-to-day will revolve around the classic Python data stack: pandas, NumPy, and scikit-learn.
You’ll also work on modern enterprise data platforms — running notebooks and lakehouse workloads in Microsoft Fabric, building pipelines and operational applications in Palantir Foundry, and deploying analytics on Azure, AWS, or GCP. The science matters, but so does making it work inside a real organization’s data ecosystem.
Your responsibilities
Translate business problems into analytical solutions: define hypotheses, choose metrics, and select the right modeling approach for the question at hand
Explore, clean, and prepare data using Python (pandas, NumPy), working with structured and semi-structured sources of varying quality
Build, validate, and tune machine learning models for classification, regression, forecasting, segmentation, and recommendation using scikit-learn, XGBoost, and statsmodels
Design and analyze experiments: A/B tests, statistical hypothesis testing, and causal analysis that hold up to scrutiny
Work hands-on with enterprise data platforms such as Microsoft Fabric (lakehouses, notebooks, semantic models) and Palantir Foundry (pipelines, ontology, operational workflows)
Communicate findings through clear visualizations, dashboards, and narratives tailored to technical and non-technical stakeholders alike
Collaborate with data engineers on data availability and quality, and with ML/AI engineers to move models from notebook to production
Monitor deployed models for drift and degradation, and own the retraining and improvement cycle
Our requirements
At least 4 years in data science or applied analytics, with models and analyses that made it past the prototype stage
Strong Python skills across the standard data stack (pandas, NumPy, scikit-learn)
Solid statistical foundations: hypothesis testing, regression, experimental design, and knowing when a result is real versus noise
Knowledge of at least one cloud/data platform and its data science components — Azure (Microsoft Fabric, Azure Machine Learning), AWS (SageMaker), GCP (Vertex AI, BigQuery ML), Snowflake (Snowpark ML, Cortex), Databricks (MLflow, Mosaic AI), or Palantir (Foundry Code Workspaces, Foundry ML, AIP)
Ability to communicate analytical results clearly to business stakeholders and influence decisions with data
Familiarity with working autonomously while collaborating effectively with data engineers, architects, and product teams
Fluent English, both written and spoken
Fluent Polish required
Residing in Poland required
Optional
Production experience with Microsoft Fabric or Palantir Foundry certification (Foundry Data Engineer / Data Scientist tracks)
Experience with time series forecasting, NLP, recommender systems, or exposure to LLM-based workflows
Familiarity with MLOps practices: MLflow, experiment tracking, model versioning, and CI/CD for ML
Experience with Power BI or other BI tools, and distributed processing with Spark/PySpark
... Crime (Anti Money Laundering, Combating Terrorism Financing, Fraud Risk Management, FATCA and implementation of financial sanctions, etc.). We are looking for: Data Scientist Your future role: - Be Part of the Data Science (R&D) Team and Take Data Throughout Its Full Lifecycle - Research, Exploration, Preprocessing, Development, ...
... opportunities with the possibility to rotate between projects ,[Work with the ML team on feature engineering, model selection, training, validation and testing, Analyse datasets and assess how useful they are for the project, Collaborate closely with data scientists and data engineers] Requirements: SQL, Python, PySpark, pandas, Machine ...
... skills - Written and verbal proficiency in English is required Nice to have: - Experience in a commercial, or education technology environment - Familiarity with cloud platforms such as AWS, Azure, or GCP - Experience with modern data engineering tools and frameworks - Knowledge of Agile methodologies and product development ...
... opportunities with the possibility to rotate between projects ,[Work with the ML team on feature engineering, model selection, training, validation and testing, Analyse datasets and assess how useful they are for the project, Collaborate closely with data scientists and data engineers] Requirements: SQL, Python, PySpark, pandas, Machine ...
... incremental embedding, storage tiering — and report the unit economics. – Write runbooks so the client’s own team can operate this after handover. SKILLS – Strong Python and solid SQL – Unstructured-data pipelines: transcripts, mail, chat, documents — parsing, normalisation, deduplication – Embedding / retrieval infrastructure: ...
... Posiadanie certyfikatów z obszaru zarządzania danymi (Data Engineer, Data Analyst, Data Scientist, Data Engineer, Data Architect, w szczególności na platformie Google Cloud Platform) - Umiejętność wizualizacji danych w narzędziach BI (Qlik Sense/Looker Studio/Tableau) - Znajomość zagadnień Data Governance Dla naszego Klienta lidera ...
... oraz evaluation frameworks., Implementacja mechanizmów guardrails, safety, PII handling, audit logs oraz human-in-the-loop., Budowanie produkcyjnego kodu w Pythonie z wykorzystaniem dobrych praktyk dotyczących testowania, packagingu i jakości kodu., Współpraca z Data Engineers, Data Scientists, Software Engineers oraz ...