Intermediate Data Engineer
Job Description
<p><span style="font-size: 12pt;"><strong>About us</strong></span></p> <p>AB InBev is the leading global brewer and one of the world’s top 5 consumer product companies. With over 500 beer brands, we’re number one or two in many of the world’s top beer markets, including North America, Latin America, Europe, Asia, and Africa.</p> <p><span style="font-size: 12pt;"><strong>About BEES</strong></span></p> <p>At BEES, our ambition is – and always will be – to put customers at the heart of everything we do, making their lives easier and their businesses more profitable. Through our B2B e-commerce and SaaS platform, we bring the power of digital to small and medium-sized retailers, unlocking new growth opportunities for all.</p> <p><strong>What you'll do:</strong></p> <ul> <li data-start="1260" data-end="1353">Implement <em class="Highlight ht572e16cb-c053-46ed-af7c-2d84cb6a5668">ETL</em>/<em class="Highlight ht572e16cb-c053-46ed-af7c-2d84cb6a5668">ELT</em> solutions and data integration between multiple systems and data sources.</li> <li data-start="1356" data-end="1457">Design, implement, and maintain data pipelines to ingest, store, and process large volumes of data.</li> <li data-start="1460" data-end="1530">Collaborate with other teams to ensure data security and compliance.</li> <li data-start="1533" data-end="1645">Design, develop, and implement solutions to optimize the performance and scalability of data processing systems.</li> </ul> <p data-start="1647" data-end="1683"><span data-ccp-props="{"201341983":0,"335559739":0,"335559740":240}"> </span><strong>What you'll need:</strong></p> <ul> <li data-start="1687" data-end="1816">Bachelor's degree in Computer Science, Computer Engineering, Information Systems, Systems Analysis and Development, or similar;</li> <li data-start="1819" data-end="1842">Intermediate English;</li> <li data-start="1845" data-end="1965">Apply knowledge of data pipeline concepts and tools to implement data transformation, cleaning, and aggregation tasks.</li> <li data-start="1968" data-end="2070">Use data pipeline frameworks and libraries to automate data processing tasks and optimize workflows.</li> <li data-start="2073" data-end="2191">Follow best practices for developing data pipelines, including version control, testing, and thorough documentation.</li> <li data-start="2194" data-end="2301">Apply programming skills in <em class="Highlight ht572e16cb-c053-46ed-af7c-2d84cb6a5668">Python</em>, <em class="Highlight ht572e16cb-c053-46ed-af7c-2d84cb6a5668">PySpark</em>, Scala, and SQL to effectively manipulate and transform data.</li> <li data-start="2304" data-end="2426">Understand and utilize cloud computing platforms and services offered by providers such as AWS, <em class="Highlight ht572e16cb-c053-46ed-af7c-2d84cb6a5668">Azure</em>, and Google Cloud.</li> <li data-start="2429" data-end="2592">Develop data pipelines using orchestration frameworks (e.g., Apache Airflow, Luigi, Mage, <em class="Highlight ht572e16cb-c053-46ed-af7c-2d84cb6a5668">Databricks</em> Workflows), integrating with <em class="Highlight ht572e16cb-c053-46ed-af7c-2d84cb6a5668">PySpark</em> and/or Scala as needed.</li> <li data-start="2595" data-end="2674">Understand and apply software design principles to data engineering projects.</li> <li data-start="2677" data-end="2755">Apply data modeling techniques to design efficient and scalable data models.</li> <li data-start="2758" data-end="2868">Optimize <em class="Highlight ht572e16cb-c053-46ed-af7c-2d84cb6a5668">PySpark</em> and SQL query perform