Deloitte

Principal AI Data Engineer

โญ - Featured Role | Apply direct with Data Freelance Hub
This role is for a "Principal AI Data Engineer" in London, with a 4-month contract starting August 2026, offering a pay rate via Rockford Payroll. Key skills include expertise in Databricks, AI/GenAI systems, Python, and Azure cloud environments.
๐ŸŒŽ - Country
United States
๐Ÿ’ฑ - Currency
$ USD
-
๐Ÿ’ฐ - Day rate
Unknown
-
๐Ÿ—“๏ธ - Date
July 31, 2026
๐Ÿ•’ - Duration
3 to 6 months
-
๐Ÿ๏ธ - Location
On-site
-
๐Ÿ“„ - Contract
Unknown
-
๐Ÿ”’ - Security
Unknown
-
๐Ÿ“ - Location detailed
City Of London, England, United Kingdom
-
๐Ÿง  - Skills detailed
#Mathematics #MLflow #Statistics #GitHub #Data Architecture #Scala #Azure cloud #Azure #pydantic #Replication #Delta Lake #DevOps #Datasets #Langchain #Data Engineering #Agile #Docker #Deployment #Data Ingestion #PyTorch #Research Skills #ADF (Azure Data Factory) #Databricks #Spark (Apache Spark) #Python #SQL (Structured Query Language) #GIT #Programming #Cloud #"ETL (Extract #Transform #Load)" #AI (Artificial Intelligence) #Qlik #Data Science
Role description
Contract role: Principal AI Data Engineer Contract Location: London, 5 days onsite weekly Contract Start Date: August 2026 Contract Duration: 4 months Payroll provider: Rockford Payroll Info for Contingent Workers โ€“ Rockford Pay Job Description: Key Responsibilities โ€ข Develop and evaluate AI/GenAI/AgenticAI prototypes using tools like Copilot Studio, AI Foundry and Copilot Analyst Agent, Mosiac AI, Genie, AgentBricks, MLflow with a focus on quick wins and enterprise integration. โ€ข Build and tune Retrieval-Augmented Generation (RAG) systems, including embedding model selection, prompt engineering, and traceable evaluation. โ€ข Design and deploy basic AI agents using frameworks such as LangChain, AutoGen, and smolagents โ€ข Communicate complex AI concepts clearly to business stakeholders and cross-functional teams. โ€ข Collaborate on E platform enhancements and work within its current limitations. โ€ข Deploy models and applications using Azure OpenAI, Azure AI Foundry, Databricks Mosaic Gateway, and Docker. โ€ข Follow DevOps best practices including CI/CD pipelines, testing, linting, and GitHub workflows. โ€ข Write modular, reusable code using OOP design patterns in Python (Pydantic, PyTorch, etc.). โ€ข Operate in agile teams and contribute to sprint planning, reviews, and retrospectives. โ€ข Deliver hands on GenAI/AgenticAI systems used directly by commercial teams within Trading & Supply, taking solutions from prototype to production โ€ข Apply engineering skills (emphasis on Databricks) and research skills across experimentation, rapid prototyping, and iterative delivery. Someone who puts emphasis on reproducibility and open source, manages large-scale text and structured datasets on Databricks. โ€ข Build AI capability, manage stakeholders and communicate effectively to ensure alignment between business needs and AI solutions, and a quick understanding of commercial operations that happen in T&S โ€ข Design and run evaluation and testing frameworks for GenAI systems, including benchmarking, reproducibility checks, and structured model assessments โ€ข Build solutions using Databricks infrastructure, Genie, MLflow (deployment and tracing and evaluations), LangChain, and LangGraph, and integrate them into scalable AI workflows and architectures โ€ข Contribute to system planning, architectural design, and structured testing to ensure long term reliability, performance, and maintainability โ€ข Preferably also someone who can set the building blocks and lead building out the backlog Required Skills โ€ข Bachelor or Master or equivalent in Statistics, Mathematics, Econometrics or similar discipline with at least 8-12 yearsโ€™ experience on data science/AI projects. โ€ข Deep understanding of LLM families (GPT, Llama, Claude, Mistral) and their reasoning capabilities. โ€ข S trong experience with Databricks- DLT, Delta Lake concepts, UC governance. โ€ข Solid understanding of streaming technologies (e.g., Spark Structured Streaming, Autoloader) โ€ข Programming skills in Python, SQL, or Scala. โ€ข Proficiency in data modelling, ETL/ELT processes, and data architecture. โ€ข Strong analytical background with problem-solving skills. โ€ข Performance tuning concepts like watermarking, late data handling, parallelism & checkpointing. โ€ข Hands-on expertise in ADF, and Qlik Replicate for data ingestion and replication. โ€ข Experience working in Azure cloud environments. โ€ข Experience with GenAI evaluation frameworks and benchmarking methodologies. โ€ข Experience in MS Copilot, AI Foundry , Databricks (MosiacAI, MLflow, Agentbricks, Genie) โ€ข Strong Git practices and collaborative coding standards. โ€ข A passion for and expertise in practicing data science to solve real-world problems. โ€ข Excellent oral and written communication skills. โ€ข Strong interpersonal skills and enthusiasm for teamwork, as well as the ability to work independently. โ€ข Familiarity with the enterprise AI platforms and governance models is a plus. โ€ข Strong decision-making abilities, using data-driven insights to make informed choices that align with organizational goals. โ€ข Skills in managing conflicts and facilitating effective resolutions to maintain a positive and productive team dynamic. โ€ข Ability to engage with and manage expectations of various stakeholders, including executives, project managers, and other teams. โ€ข Proficiency in identifying potential risks in data projects and implementing strategies to mitigate them. โ€ข Strong commitment and ownership of project delivery.