XTB

Data Engineer

XTB Warsaw, Masovian Voivodeship, Poland

Financial Services · 1,001-5,000 employees

20 h ago
Remote data-engineer Mid (2-5 yrs) Full-time Poland
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

You will be responsible for building and maintaining scalable data integration mechanisms and processing pipelines using SQL, Python, and Apache Spark. Additionally, you will contribute to architectural decisions, develop CI/CD pipelines, and ensure high data quality and reliability across the centralized platform.

What they look for

Data Engineering MS SQL Server SSIS Python Apache Spark PySpark SQL CI/CD Infrastructure as Code Data Modeling ETL/ELT Data Warehouse Design REST gRPC Airflow Dagster

Requirements

Candidates must have at least 3 years of experience in data engineering and proficiency in MS SQL Server, SSIS, and Python. Strong analytical skills, experience with data modeling, and familiarity with ETL/ELT processes and CI/CD principles are essential for this role.

Benefits

Training budget Birthday day off Parental day off Equipment tailored to needs Private medical care Group insurance E-learning platform access Wellbeing platform access Private therapy sessions Remote work

Full description

XTB is a global company from the financial industry, focusing on online trading of financial instruments. We are the largest FinTech in Poland and a leader in Central and Eastern Europe, and the range of our operations covers several countries, including Asia and South America. At XTB, we focus on the development of our employees, giving them opportunities to gain knowledge and skills in various fields, as well as offering a number of training and development programs. If you are looking for challenges and want to gain valuable experience in an international business environment, XTB is the right place for you.

We are a certified Great Place to Work company.

We are looking for a Data Engineer to join the Core Data Platform team to co-create and develop a centralized data platform used by teams across the company. In this role, you will be responsible for building scalable data integration mechanisms from various source systems, developing shared platform components and standards, and ensuring high data quality, reliability, and consistency.

\n

Responsibilities

  • Implementing, maintaining and support of MS SQL Server warehouse and SSIS processes and jobs on a daily basis.
  • Designing, building, and maintaining scalable data processing pipelines using SQL, Python, and Apache Spark / PySpark.
  • Working with raw data and designing methods for its integration into the data platform.
  • Developing CI/CD pipelines for data engineering solutions.
  • Creating and maintaining Infrastructure as Code solutions.
  • Integrating the data platform with other systems and applications.
  • Participating in architectural decision-making and defining engineering standards.
  • Conducting code reviews, documenting solutions, and sharing knowledge with team members.

Requirements

  • At least 3 years of experience as a Data Engineer or in a related software engineering role.
  • At least 1 year of experience with MS SQL Server, SSIS and SQL Server Agent technologies - including the deployment of new features.
  • Analyzing existing legacy code for improvement and transfer to the newer technology.
  • Flexibility to be enrolled into the Databricks oriented project at some point of the cooperation.
  • Proficiency in Python and the ability to write clean, testable production code.
  • Hands-on experience with Apache Spark.
  • Advanced SQL skills (query optimization on large datasets).
  • Strong understanding of data warehouse design principles and data modeling.
  • Practical knowledge of ETL/ELT processes.
  • Familiarity with data quality, monitoring, and pipeline reliability.
  • Experience working with relational databases, including an understanding of how they work, data modeling, and optimization.
  • Understanding of REST and gRPC standards for system integration.
  • Experience with workflow orchestration tools (e.g., Airflow, Dagster, or similar).
  • Knowledge of CI/CD pipeline building and maintenance principles.
  • Strong analytical problem-solving skills and attention to detail.
  • Effective communication skills and the ability to collaborate in a team.
  • Openness to learning and exploring new technologies and methodologies.

Nice to have

  • Experience with dbt.
  • Experience with Kafka, Pub/Sub, or other event streaming systems.
  • Experience building near-real-time / real-time pipelines.
  • Experience working with Infrastructure-as-Code tools.
  • Understanding of CDC (Change Data Capture) and data integration from transactional systems.
  • Experience with Databricks or other lakehouse platforms.

What we offer

  • Real influence on the development of the company and the product.
  • Work in an experienced team that is happy to share its knowledge.
  • A clear vision of development thanks to regular feedback and clear career paths.
  • Regular team-building meetings.

Benefits

  • A training budget for courses and conferences that interest you.
  • An extra day off on your birthday.
  • An extra day off for parents.
  • Equipment tailored to your needs.
  • Private medical care and group insurance.
  • Access to an e-learning platform for learning English and a benefits platform.
  • Access to a wellbeing platform and the opportunity to take advantage of workshops and private therapy sessions.
  • Remote work, from the office in Warsaw or from a coworking space in your city.

\n

Similar roles