Senior Data Engineer at AIA Contract Documents | Torre

Senior Data Engineer

Emma highlights
This highlight was written by Emma’s AI. Ask Emma to edit it.
Full-time

Legal agreement: Employment

Compensation
USD150k - 180k/year
location_on
Remote (anywhere)
Match
skeleton-gauges
You have opted out of job matches in .
To undo this, go to the 'Skills and Interests' section of your preferences.
Review preferences
Shared by
Emma of Torre.ai
3 months ago

Requirements and responsibilities


About ACDFor more than a century, AIA Contract Documents (ACD) has supported architecture, engineering, and construction professionals by delivering a shared industry standard to align parties on a project.What began in 1888 with the development of standardized construction contracts has evolved into a comprehensive suite of contract tools and foundational workflows that not only shape how the industry works today, but uniquely position ACD to help firms navigate construction’s growing complexity.In a world driven by scale, fragmentation, and AI-generated decisions that prioritize speed over a clear understanding of risk, ACD serves as a trusted anchor, ensuring project participants can reduce disputes and negotiations, while achieving faster alignment and more predictable outcomes for the future.AIA Contract Documents is seeking a Senior Data Engineer to support the buildout of data infrastructure powering upcoming AI initiatives. This role will play a critical part in designing and structuring data systems that enable scalable, high-quality inputs for AI and machine learning models.The ideal candidate brings deep experience in modern data engineering practices, strong familiarity with Databricks, and the ability to operate independently in a fast-evolving environment with limited oversight.Key ResponsibilitiesDesign and implement scalable data models optimized for AI/ML use cases, ensuring data is structured for effective model training and inferenceArchitect and manage data pipelines using Databricks, including orchestration, job scheduling, and workflow optimizationDevelop and maintain robust ETL/ELT processes to support data ingestion, transformation, and delivery across systemsLeverage Databricks Asset Bundles to manage deployment of data assets (pipelines, jobs, notebooks, and files) across environmentsCollaborate with cross-functional teams to align data architecture with AI initiative requirements and business objectivesEnsure data quality, integrity, and governance standards are met across all pipelines and datasetsContribute to CI/CD practices using Azure DevOps (ADO), with a future transition to GitHub-based workflowsParticipate in Agile development processes, including sprint planning, stand-ups, and iterative deliveryOther duties as assignedQualifications5+ years of experience in Data Engineering or related fieldStrong expertise in Databricks, including pipeline development, orchestration, and data architectureExperience designing data models to support AI/ML applicationsProficiency in Python (PySpark) and SQLHands-on experience with Azure-based data environmentsExperience with CI/CD tools such as Azure DevOps (ADO)Ability to work independently with minimal oversight and ramp quickly in a fast-paced environmentExperience with Databricks Asset Bundles for deployment and environment management, preferredFamiliarity with transitioning CI/CD workflows to GitHub, preferredExperience working in Agile development environments, preferredExposure to AI/ML workflows and data requirements for model development, preferredWhat Success Looks LikeEfficient, scalable data pipelines supporting AI initiativesWell-structured, high-quality datasets optimized for machine learningStreamlined deployment and orchestration of Databricks assets across environmentsStrong collaboration with stakeholders to deliver data solutions aligned with business needs