Publicado el 17 ago 2026 · Confirmamos el 06 sept 2026 que sigue activo
¿Esta empresa es tuya?This is a Full-Time Role (40 hours per week, 5 days per week) with no option for part-time work. While this is a remote-first opportunity, the candidate filling this role must be a resident of Pennsylvania, New York, or Brazil at the start of employment. Additionally, they must be within commuting distance of our office in Philadelphia, New York City, or São Paulo.
Please visit our Careers page to review all opportunities and submit your application for the role(s) that best fit your location and work authorization.
The Data & AI team is responsible for building B Lab's data platform, delivering on business priorities, and scaling AI adoption across the organization. The team reports directly to the CTDO. The Data & ML Platforms team owns the data platform, ML infrastructure, and the foundational data models the rest of the Data & AI team depends on — the supply-side capability layer that Business & Data Priorities and AI Enablement build on. About the Opportunity As a Junior Data Engineer within the Data & ML Platforms pillar, you build and maintain the data pipelines and infrastructure the rest of the Data & AI team depends on. Your first priority is unblocking the onboarding of new data sources — currently one of the team's top hiring priorities, paused pending this hire. You will absorb data engineering work currently split between the Senior Machine Learning Engineer and the Pillar Lead, freeing them to focus on ML infrastructure and platform strategy respectively. You will work closely with the Senior Analytics Engineer on core data modeling and with the Data Governance Lead on data standards, definitions, and compliance. Core Responsibilities
Data Pipeline Development & New Source Onboarding (55%): Own Pipeline Development End-to-End: Design, build, and maintain robust, scalable ETL/ELT pipelines that reliably deliver clean data to the platform. Add New Data Sources: Evaluate, scope, and integrate new data sources as they're identified. Partner Across Pillars: Work continuously with Network Priorities, Regional Enablement and AI Enablement to understand incoming data needs and translate them into pipeline work. Monitor & Troubleshoot: Proactively identify and resolve pipeline failures and data quality issues before they affect downstream users. Platform Support & Cross-Team Collaboration (35%): Support Foundational Data Models: Work with the Senior Analytics Engineer to maintain the core data models the rest of the team depends on. Ensure Data Availability for Consumers: Make sure the data needed by Data Analysts and the Senior Machine Learning Engineer is reliably available. Follow Data Governance Standards: Apply the data standards, definitions, and sensitivity classifications set by the Data Governance Lead. Strategic Innovation & Business Impact (10%): Evaluate Pipeline Tooling: Explore and pilot new ETL/ELT tools or approaches that could improve onboarding speed or pipeline reliability. Quantify Impact: Track and articulate how pipeline reliability and onboarding speed affect downstream analytics and ML work. Major Objectives/Project for the role in the first 6-12 months Add new data sources to the Data Platform Improve data infrastructure allowing for less downtime and more proactive monitoring Enable transformation of data and data models
Improve dependency handling between tables Improve data labelling and documentation for AI usage About You A BA/BS in Computer Science, Information Technology, or a related field strongly preferred Minimum of 2+ years of experience in data engineering Experience working in a DevOps-oriented culture that prioritizes continuous integration and continuous deployment Proficiency with Git and collaborative version control workflows (e.g., branching, pull requests, code review) Experience with Infrastructure as Code (e.g., Terraform, CloudFormation, or CDK) for provisioning and managing cloud infrastructure
Proven experience in designing and deploying data solutions Experience designing, building, and onboarding new data sources into ETL/ELT pipelines
Ability and desire to take product/project ownership Proficiency in SQL and experience with scripting languages such as Python, Java, or Scala Experience with data pipeline and workflow management tools Strong knowledge of big data tools and frameworks such as Hadoop, Spark, or Hive is a plus Experience using AI coding assistants and other AI tools to improve development speed and productivity
Excellent communication skills Compensation Details B Lab has a compensation plan that includes: A yearly salary in the range of R$122,100 - R$146,300 (not including the 13th salary) Sick & other leave in accordance with Brazilian statutory leave allowance Company provided laptop We also offer other benefits that are based on company policy and are not included in your contract and therefore are subject to change or addition as our organization works to support our staff. Paid time off during organization-wide closures for wellness Professional Development and time off: 40 hours paid time off with access to professional development after 1 year of service Paid time off for volunteering - after one year of service One time home office set-up allowance Additional perks you may qualify for: monthly home office allowance, monthly food allowance & monthly health insurance reimbursement Remote-first workplace This will be a CLT contract. This is a
Full-Time Role (40 hours per week, 5DWW) with no option for part-time work.
This job ad is for São Paulo, Brazil. While this is a remote-first opportunity, the candidate filling this role must hold Brazilian work authorization without any time limitations or any other restrictions, and they must be a resident of Brazil at the start of employment. Additionally, they must be within commuting distance of São Paulo. If you wish to be based in one of our other
Crea una cuenta gratis para ver el empleo completo y postularte.