VySystems

Data Architect

⭐ - Featured Role | Apply direct with Data Freelance Hub
This role is for a Data Architect focused on Databricks, with a contract length of "X months" and a pay rate of "$X/hour." Required skills include SQL, Python, Delta Lake, and cloud platform experience (AWS/Azure/GCP). Databricks certifications are preferred.
🌎 - Country
United States
πŸ’± - Currency
$ USD
-
πŸ’° - Day rate
Unknown
-
πŸ—“οΈ - Date
August 13, 2026
πŸ•’ - Duration
Unknown
-
🏝️ - Location
Unknown
-
πŸ“„ - Contract
Unknown
-
πŸ”’ - Security
Unknown
-
πŸ“ - Location detailed
New Jersey, United States
-
🧠 - Skills detailed
#AWS S3 (Amazon Simple Storage Service) #Security #Terraform #Batch #Azure ADLS (Azure Data Lake Storage) #Strategy #Scala #Compliance #S3 (Amazon Simple Storage Service) #Azure DevOps #Databricks #MLflow #Jenkins #Data Engineering #AutoScaling #PySpark #Kafka (Apache Kafka) #GCP (Google Cloud Platform) #Airflow #Spark (Apache Spark) #Apache Spark #ML (Machine Learning) #AWS (Amazon Web Services) #Data Pipeline #SQL (Structured Query Language) #Delta Lake #ADLS (Azure Data Lake Storage) #Data Governance #ADF (Azure Data Factory) #Data Architecture #Infrastructure as Code (IaC) #Automation #DevOps #Python #GitHub #Azure #Cloud
Role description
We are seeking a highly skilled Databricks Architect to lead the design and implementation of next-generation data platforms built on the Lakehouse paradigm. This role goes beyond pipeline developmentβ€”you will own the Databricks platform architecture end-to-end, driving scalability, governance, performance, and cost optimization across enterprise data ecosystems. Databricks Platform Architecture β€’ Architect and implement enterprise-scale Databricks environments (dev/test/prod) β€’ Define workspace strategy, cluster policies, and job orchestration frameworks β€’ Design secure, scalable Lakehouse architecture using Delta Lake Data Engineering & Processing β€’ Build and optimize high-performance data pipelines using Apache Spark (PySpark / Scala) β€’ Implement Medallion Architecture (Bronze, Silver, Gold) β€’ Develop batch and real-time streaming pipelines using Structured Streaming Governance, Security & Compliance β€’ Implement fine-grained access control using Unity Catalog β€’ Define enterprise-wide data governance, lineage, and auditing frameworks β€’ Ensure compliance with security and regulatory standards Performance & Cost Optimization β€’ Optimize workloads using partitioning, caching, and Photon engine β€’ Design cost-efficient cluster strategies (autoscaling, spot instances, DBU optimization) β€’ Monitor and improve query and pipeline performance at scale DevOps & Automation β€’ Implement CI/CD pipelines for Databricks using Azure DevOps, Jenkins, or GitHub Actions β€’ Enable Infrastructure as Code using Terraform or equivalent β€’ Standardize reusable frameworks and engineering best practices Cloud Integration β€’ Architect Databricks solutions on AWS / Azure / GCP β€’ Integrate with cloud-native services (AWS S3/Glue, Azure ADLS/ADF, etc.) Must-Have Skills β€’ Advanced proficiency in SQL and Python β€’ Hands-on experience with Delta Lake, Medallion Architecture, and Unity Catalog β€’ Strong experience with at least one cloud platform (AWS / Azure / GCP) β€’ Experience with CI/CD tools (Azure DevOps, Jenkins, GitHub Actions) β€’ Knowledge of workflow orchestration tools (Airflow or equivalent) Good to Have β€’ Databricks certifications (Associate / Professional) β€’ Experience with Databricks SQL & dashboards β€’ Exposure to ML pipelines (MLflow) β€’ Experience with Kafka or event-driven architectures β€’ Domain experience in Finance / Retail / Healthcare What Makes You a Great Fit β€’ Strong architectural mindset with ability to design scalable data platforms β€’ Deep understanding of performance tuning and cost optimization β€’ Ability to balance technical depth with business impact β€’ Excellent communication and stakeholder management skills