

VySystems
Data Architect
β - Featured Role | Apply direct with Data Freelance Hub
This role is for a Data Architect focused on Databricks, with a contract length of "X months" and a pay rate of "$X/hour." Required skills include SQL, Python, Delta Lake, and cloud platform experience (AWS/Azure/GCP). Databricks certifications are preferred.
π - Country
United States
π± - Currency
$ USD
-
π° - Day rate
Unknown
-
ποΈ - Date
August 13, 2026
π - Duration
Unknown
-
ποΈ - Location
Unknown
-
π - Contract
Unknown
-
π - Security
Unknown
-
π - Location detailed
New Jersey, United States
-
π§ - Skills detailed
#AWS S3 (Amazon Simple Storage Service) #Security #Terraform #Batch #Azure ADLS (Azure Data Lake Storage) #Strategy #Scala #Compliance #S3 (Amazon Simple Storage Service) #Azure DevOps #Databricks #MLflow #Jenkins #Data Engineering #AutoScaling #PySpark #Kafka (Apache Kafka) #GCP (Google Cloud Platform) #Airflow #Spark (Apache Spark) #Apache Spark #ML (Machine Learning) #AWS (Amazon Web Services) #Data Pipeline #SQL (Structured Query Language) #Delta Lake #ADLS (Azure Data Lake Storage) #Data Governance #ADF (Azure Data Factory) #Data Architecture #Infrastructure as Code (IaC) #Automation #DevOps #Python #GitHub #Azure #Cloud
Role description
We are seeking a highly skilled Databricks Architect to lead the design and implementation of next-generation data platforms built on the Lakehouse paradigm.
This role goes beyond pipeline developmentβyou will own the Databricks platform architecture end-to-end, driving scalability, governance, performance, and cost optimization across enterprise data ecosystems.
Databricks Platform Architecture
β’ Architect and implement enterprise-scale Databricks environments (dev/test/prod)
β’ Define workspace strategy, cluster policies, and job orchestration frameworks
β’ Design secure, scalable Lakehouse architecture using Delta Lake
Data Engineering & Processing
β’ Build and optimize high-performance data pipelines using Apache Spark (PySpark / Scala)
β’ Implement Medallion Architecture (Bronze, Silver, Gold)
β’ Develop batch and real-time streaming pipelines using Structured Streaming
Governance, Security & Compliance
β’ Implement fine-grained access control using Unity Catalog
β’ Define enterprise-wide data governance, lineage, and auditing frameworks
β’ Ensure compliance with security and regulatory standards
Performance & Cost Optimization
β’ Optimize workloads using partitioning, caching, and Photon engine
β’ Design cost-efficient cluster strategies (autoscaling, spot instances, DBU optimization)
β’ Monitor and improve query and pipeline performance at scale
DevOps & Automation
β’ Implement CI/CD pipelines for Databricks using Azure DevOps, Jenkins, or GitHub Actions
β’ Enable Infrastructure as Code using Terraform or equivalent
β’ Standardize reusable frameworks and engineering best practices
Cloud Integration
β’ Architect Databricks solutions on AWS / Azure / GCP
β’ Integrate with cloud-native services (AWS S3/Glue, Azure ADLS/ADF, etc.)
Must-Have Skills
β’ Advanced proficiency in SQL and Python
β’ Hands-on experience with Delta Lake, Medallion Architecture, and Unity Catalog
β’ Strong experience with at least one cloud platform (AWS / Azure / GCP)
β’ Experience with CI/CD tools (Azure DevOps, Jenkins, GitHub Actions)
β’ Knowledge of workflow orchestration tools (Airflow or equivalent)
Good to Have
β’ Databricks certifications (Associate / Professional)
β’ Experience with Databricks SQL & dashboards
β’ Exposure to ML pipelines (MLflow)
β’ Experience with Kafka or event-driven architectures
β’ Domain experience in Finance / Retail / Healthcare
What Makes You a Great Fit
β’ Strong architectural mindset with ability to design scalable data platforms
β’ Deep understanding of performance tuning and cost optimization
β’ Ability to balance technical depth with business impact
β’ Excellent communication and stakeholder management skills
We are seeking a highly skilled Databricks Architect to lead the design and implementation of next-generation data platforms built on the Lakehouse paradigm.
This role goes beyond pipeline developmentβyou will own the Databricks platform architecture end-to-end, driving scalability, governance, performance, and cost optimization across enterprise data ecosystems.
Databricks Platform Architecture
β’ Architect and implement enterprise-scale Databricks environments (dev/test/prod)
β’ Define workspace strategy, cluster policies, and job orchestration frameworks
β’ Design secure, scalable Lakehouse architecture using Delta Lake
Data Engineering & Processing
β’ Build and optimize high-performance data pipelines using Apache Spark (PySpark / Scala)
β’ Implement Medallion Architecture (Bronze, Silver, Gold)
β’ Develop batch and real-time streaming pipelines using Structured Streaming
Governance, Security & Compliance
β’ Implement fine-grained access control using Unity Catalog
β’ Define enterprise-wide data governance, lineage, and auditing frameworks
β’ Ensure compliance with security and regulatory standards
Performance & Cost Optimization
β’ Optimize workloads using partitioning, caching, and Photon engine
β’ Design cost-efficient cluster strategies (autoscaling, spot instances, DBU optimization)
β’ Monitor and improve query and pipeline performance at scale
DevOps & Automation
β’ Implement CI/CD pipelines for Databricks using Azure DevOps, Jenkins, or GitHub Actions
β’ Enable Infrastructure as Code using Terraform or equivalent
β’ Standardize reusable frameworks and engineering best practices
Cloud Integration
β’ Architect Databricks solutions on AWS / Azure / GCP
β’ Integrate with cloud-native services (AWS S3/Glue, Azure ADLS/ADF, etc.)
Must-Have Skills
β’ Advanced proficiency in SQL and Python
β’ Hands-on experience with Delta Lake, Medallion Architecture, and Unity Catalog
β’ Strong experience with at least one cloud platform (AWS / Azure / GCP)
β’ Experience with CI/CD tools (Azure DevOps, Jenkins, GitHub Actions)
β’ Knowledge of workflow orchestration tools (Airflow or equivalent)
Good to Have
β’ Databricks certifications (Associate / Professional)
β’ Experience with Databricks SQL & dashboards
β’ Exposure to ML pipelines (MLflow)
β’ Experience with Kafka or event-driven architectures
β’ Domain experience in Finance / Retail / Healthcare
What Makes You a Great Fit
β’ Strong architectural mindset with ability to design scalable data platforms
β’ Deep understanding of performance tuning and cost optimization
β’ Ability to balance technical depth with business impact
β’ Excellent communication and stakeholder management skills






