Descrição da vaga
<h1 data-pm-slice="1 1 []">MID/SENIOR DATA ENGINEER – CAPCO POLAND</h1>
<p><em>We offer a flexible collaboration model based on a B2B contract, with the opportunity to work on innovative AI and automation initiatives for leading financial institutions.</em></p>
<p>At <strong>Capco Poland</strong>, we’re not just another consultancy – we’re the spark behind digital transformation in the financial world. As a global leader in technology and management consulting, we help our clients tackle complex challenges across banking, payments, capital markets, wealth, and asset management.</p>
<p>Our secret?<br>A culture that’s fast, flexible, and fiercely entrepreneurial. We move quickly, think creatively, and always put our people first.</p>
<p>We’re passionate about growth – both for our clients and ourselves – and that means attracting talented professionals who want to develop their skills, take ownership, and make a real impact.</p>
<p>We’re proud to be:</p>
<ul data-spread="false">
<li>
<p>Trailblazers in banking, payments, capital markets, wealth, and asset management</p>
</li>
<li>
<p>Champions of an agile, nimble, and innovative work environment</p>
</li>
<li>
<p>Dedicated to building a team of talented professionals who share our drive and vision</p>
</li>
</ul>
<h2>THE ROLE</h2>
<p>We are looking for a <strong>Mid Data Engineer</strong> to join our growing data engineering team and contribute to building scalable, reliable data solutions for our financial services clients.</p>
<p>You will work with modern data technologies and cloud platforms, developing and maintaining data pipelines, processing large datasets, and supporting the delivery of enterprise-scale data solutions.</p>
<p>This is a great opportunity for a Data Engineer who already has hands-on commercial experience and wants to further develop their expertise in <strong>Python, Apache Spark, Hadoop, Linux, and Google Cloud Platform (GCP)</strong> while working on complex international projects.</p>
<h2>WHAT YOU’LL DO</h2>
<ul data-spread="false">
<li>
<p>Design, develop, and maintain scalable data pipelines and data processing solutions.</p>
</li>
<li>
<p>Develop data transformation and processing workflows using <strong>Python and Apache Spark</strong>.</p>
</li>
<li>
<p>Work with large-scale datasets in distributed environments using <strong>Hadoop and related technologies</strong>.</p>
</li>
<li>
<p>Build and support cloud-based data solutions on <strong>Google Cloud Platform (GCP)</strong>.</p>
</li>
<li>
<p>Develop reliable ingestion processes integrating data from multiple source systems.</p>
</li>
<li>
<p>Implement data transformations, validation rules, and data quality checks.</p>
</li>
<li>
<p>Troubleshoot data pipeline issues and support performance optimization.</p>
</li>
<li>
<p>Work with <strong>Linux-based environments</strong>, including scripting, deployment, and operational activities.</p>
</li>
<li>
<p>Collaborate with Data Engineers, Architects, Analysts, and other project stakeholders to translate business requirements into technical solutions.</p>
</li>
<li>
<p>Participate in code reviews and follow software engineering and data engineering best practices.</p>
</li>
<li>
<p>Create and maintain technical documentation covering data flows, dependencies, configurations, and operational procedures.</p>
</li>
<li>
<p>Support deployment, testing, stabilization, and ongoing maintenance of data solutions.</p>
</li>
</ul>
<h2>WHAT WE’RE LOOKING FOR</h2>
<ul data-spread="false">
<li>
<p>2–4+ years of commercial experience in <strong>Data Engineering</strong> or a similar role.</p>
</li>
<li>
<p>Good hands-on programming skills in <strong>Python</strong>.</p>
</li>
<li>
<p>Practical experience with <strong>Apache Spark</strong>, including building and maintaining data processing jobs.</p>
</li>
<li>
<p>Experience working with <strong>Hadoop</strong> or distributed data processing ecosystems.</p>
</li>
<li>
<p>Good knowledge of <strong>Linux</strong> and command-line environments.</p>
</li>
<li>
<p>Commercial experience with <strong>Google Cloud Platform (GCP)</strong> and relevant data services.</p>
</li>
<li>
<p>Good understanding of <strong>ETL/ELT processes, data pipelines, and data transformation concepts</strong>.</p>
</li>
<li>
<p>Working knowledge of <strong>SQL</strong> and relational data concepts.</p>
</li>
<li>
<p>Understanding of data quality, monitoring, and troubleshooting practices.</p>
</li>
<li>
<p>Familiarity with Git and modern software development practices.</p>
</li>
<li>
<p>Ability to work effectively in an Agile environment and collaborate with distributed teams.</p>
</li>
<li>
<p>Good communication skills and <strong>English at a minimum B2 level</strong>.</p>
</li>
</ul>
<h2>NICE TO HAVE</h2>
<ul data-spread="false">
<li>
<p>Experience with GCP services such as <strong>BigQuery, Cloud Storage, Dataproc, Dataflow, or Pub/Sub</strong>.</p>
</li>
<li>Experience in Financial/Banking domain</li>
<li>
<p>Experience with orchestration tools such as <strong>Apache Airflow</strong>.</p>
</li>
<li>
<p>Familiarity with CI/CD processes for data solutions.</p>
</li>
<li>
<p>Knowledge of data modelling and data warehouse concepts.</p>
</li>
<li>
<p>Experience working with financial services or banking clients.</p>
</li>
<li>
<p>Familiarity with containerization technologies such as Docker or Kubernetes.</p>
</li>
</ul>
<h2>ONLINE RECRUITMENT PROCESS</h2>
<ol data-spread="false">
<li>
<p>Screening call with the Recruiter</p>
</li>
<li>
<p>Hiring Manager Technical Interview</p>
</li>
<li>
<p>Client Interview</p>
</li>
<li>
<p>Feedback / Offer</p>
</li>
</ol>
<p><strong>#LI-HYBRID</strong></p>
<p> </p>