Mid/Senior Data Engineer (Kraków/GCP)
Job Description
<h1 data-pm-slice="1 1 []">MID/SENIOR DATA ENGINEER – CAPCO POLAND</h1> <p><em>We offer a flexible collaboration model based on a B2B contract, with the opportunity to work on innovative AI and automation initiatives for leading financial institutions.</em></p> <p>At <strong>Capco Poland</strong>, we’re not just another consultancy – we’re the spark behind digital transformation in the financial world. As a global leader in technology and management consulting, we help our clients tackle complex challenges across banking, payments, capital markets, wealth, and asset management.</p> <p>Our secret?<br>A culture that’s fast, flexible, and fiercely entrepreneurial. We move quickly, think creatively, and always put our people first.</p> <p>We’re passionate about growth – both for our clients and ourselves – and that means attracting talented professionals who want to develop their skills, take ownership, and make a real impact.</p> <p>We’re proud to be:</p> <ul data-spread="false"> <li> <p>Trailblazers in banking, payments, capital markets, wealth, and asset management</p> </li> <li> <p>Champions of an agile, nimble, and innovative work environment</p> </li> <li> <p>Dedicated to building a team of talented professionals who share our drive and vision</p> </li> </ul> <h2>THE ROLE</h2> <p>We are looking for a <strong>Mid Data Engineer</strong> to join our growing data engineering team and contribute to building scalable, reliable data solutions for our financial services clients.</p> <p>You will work with modern data technologies and cloud platforms, developing and maintaining data pipelines, processing large datasets, and supporting the delivery of enterprise-scale data solutions.</p> <p>This is a great opportunity for a Data Engineer who already has hands-on commercial experience and wants to further develop their expertise in <strong>Python, Apache Spark, Hadoop, Linux, and Google Cloud Platform (GCP)</strong> while working on complex international projects.</p> <h2>WHAT YOU’LL DO</h2> <ul data-spread="false"> <li> <p>Design, develop, and maintain scalable data pipelines and data processing solutions.</p> </li> <li> <p>Develop data transformation and processing workflows using <strong>Python and Apache Spark</strong>.</p> </li> <li> <p>Work with large-scale datasets in distributed environments using <strong>Hadoop and related technologies</strong>.</p> </li> <li> <p>Build and support cloud-based data solutions on <strong>Google Cloud Platform (GCP)</strong>.</p> </li> <li> <p>Develop reliable ingestion processes integrating data from multiple source systems.</p> </li> <li> <p>Implement data transformations, validation rules, and data quality checks.</p> </li> <li> <p>Troubleshoot data pipeline issues and support performance optimization.</p> </li> <li> <p>Work with <strong>Linux-based environments</strong>, including scripting, deployment, and operational activities.</p> </li> <li> <p>Collaborate with Data Engineers, Architects, Analysts, and other project stakeholders to translate business requirements into technical solutions.</p> </li> <li> <p>Participate in code reviews and follow software engineering and data engineering best practices.</p> </li> <li> <p>Create and maintain technical documentation covering data flows, dependencies, configurations, and operational procedures.</p> </li> <li> <p>Support deployment, testing, stabilization, a