Capco logo

Capco

Mid/Senior Data Engineer (Kraków/GCP) at Capco

Poland - CracowFull-timeData & AnalyticsPosted 10 days ago
Apply with Pipeline

About the Role

<h1 data-pm-slice="1 1 []">MID/SENIOR DATA ENGINEER – CAPCO POLAND</h1> <p><em>We offer a flexible collaboration model based on a B2B contract, with the opportunity to work on innovative AI and automation initiatives for leading financial institutions.</em></p> <p>At <strong>Capco Poland</strong>, we’re not just another consultancy – we’re the spark behind digital transformation in the financial world. As a global leader in technology and management consulting, we help our clients tackle complex challenges across banking, payments, capital markets, wealth, and asset management.</p> <p>Our secret?<br>A culture that’s fast, flexible, and fiercely entrepreneurial. We move quickly, think creatively, and always put our people first.</p> <p>We’re passionate about growth – both for our clients and ourselves – and that means attracting talented professionals who want to develop their skills, take ownership, and make a real impact.</p> <p>We’re proud to be:</p> <ul data-spread="false"> <li> <p>Trailblazers in banking, payments, capital markets, wealth, and asset management</p> </li> <li> <p>Champions of an agile, nimble, and innovative work environment</p> </li> <li> <p>Dedicated to building a team of talented professionals who share our drive and vision</p> </li> </ul> <h2>THE ROLE</h2> <p>We are looking for a <strong>Mid Data Engineer</strong> to join our growing data engineering team and contribute to building scalable, reliable data solutions for our financial services clients.</p> <p>You will work with modern data technologies and cloud platforms, developing and maintaining data pipelines, processing large datasets, and supporting the delivery of enterprise-scale data solutions.</p> <p>This is a great opportunity for a Data Engineer who already has hands-on commercial experience and wants to further develop their expertise in <strong>Python, Apache Spark, Hadoop, Linux, and Google Cloud Platform (GCP)</strong> while working on complex international projects.</p> <h2>WHAT YOU’LL DO</h2> <ul data-spread="false"> <li> <p>Design, develop, and maintain scalable data pipelines and data processing solutions.</p> </li> <li> <p>Develop data transformation and processing workflows using <strong>Python and Apache Spark</strong>.</p> </li> <li> <p>Work with large-scale datasets in distributed environments using <strong>Hadoop and related technologies</strong>.</p> </li> <li> <p>Build and support cloud-based data solutions on <strong>Google Cloud Platform (GCP)</strong>.</p> </li> <li> <p>Develop reliable ingestion processes integrating data from multiple source systems.</p> </li> <li> <p>Implement data transformations, validation rules, and data quality checks.</p> </li> <li> <p>Troubleshoot data pipeline issues and support performance optimization.</p> </li> <li> <p>Work with <strong>Linux-based environments</strong>, including scripting, deployment, and operational activities.</p> </li> <li> <p>Collaborate with Data Engineers, Architects, Analysts, and other project stakeholders to translate business requirements into technical solutions.</p> </li> <li> <p>Participate in code reviews and follow software engineering and data engineering best practices.</p> </li> <li> <p>Create and maintain technical documentation covering data flows, dependencies, configurations, and operational procedures.</p> </li> <li> <p>Support deployment, testing, stabilization, and ongoing maintenance of data solutions.</p> </li> </ul> <h2>WHAT WE’RE LOOKING FOR</h2> <ul data-spread="false"> <li> <p>2–4+ years of commercial experience in <strong>Data Engineering</strong> or a similar role.</p> </li> <li> <p>Good hands-on programming skills in <strong>Python</strong>.</p> </li> <li> <p>Practical experience with <strong>Apache Spark</strong>, including building and maintaining data processing jobs.</p> </li> <li> <p>Experience working with <strong>Hadoop</strong> or distributed data processing ecosystems.</p> </li> <li> <p>Good knowledge of <strong>Linux</strong> and command-line environments.</p> </li> <li> <p>Commercial experience with <strong>Google Cloud Platform (GCP)</strong> and relevant data services.</p> </li> <li> <p>Good understanding of <strong>ETL/ELT processes, data pipelines, and data transformation concepts</strong>.</p> </li> <li> <p>Working knowledge of <strong>SQL</strong> and relational data concepts.</p> </li> <li> <p>Understanding of data quality, monitoring, and troubleshooting practices.</p> </li> <li> <p>Familiarity with Git and modern software development practices.</p> </li> <li> <p>Ability to work effectively in an Agile environment and collaborate with distributed teams.</p> </li> <li> <p>Good communication skills and <strong>English at a minimum B2 level</strong>.</p> </li> </ul> <h2>NICE TO HAVE</h2> <ul data-spread="false"> <li> <p>Experience with GCP services such as <strong>BigQuery, Cloud Storage, Dataproc, Dataflow, or Pub/Sub</strong>.</p> </li> <li>Experience in Financial/Banking domain</li> <li> <p>Experience with orchestration tools such as <strong>Apache Airflow</strong>.</p> </li> <li> <p>Familiarity with CI/CD processes for data solutions.</p> </li> <li> <p>Knowledge of data modelling and data warehouse concepts.</p> </li> <li> <p>Experience working with financial services or banking clients.</p> </li> <li> <p>Familiarity with containerization technologies such as Docker or Kubernetes.</p> </li> </ul> <h2>ONLINE RECRUITMENT PROCESS</h2> <ol data-spread="false"> <li> <p>Screening call with the Recruiter</p> </li> <li> <p>Hiring Manager Technical Interview</p> </li> <li> <p>Client Interview</p> </li> <li> <p>Feedback / Offer</p> </li> </ol> <p><strong>#LI-HYBRID</strong></p> <p>&nbsp;</p>