Envision Technology Solutions

Java Spark Engineer

⭐ - Featured Role | Apply direct with Data Freelance Hub
This role is for a Java Spark Engineer, a long-term contract position based in Berkeley Heights, NJ. Requires 7+ years Java experience, 5+ years with Apache Spark, strong SQL skills, and expertise in distributed systems. Onsite work, 5 days a week.
🌎 - Country
United States
πŸ’± - Currency
$ USD
-
πŸ’° - Day rate
Unknown
-
πŸ—“οΈ - Date
August 4, 2026
πŸ•’ - Duration
Unknown
-
🏝️ - Location
On-site
-
πŸ“„ - Contract
W2 Contractor
-
πŸ”’ - Security
Unknown
-
πŸ“ - Location detailed
Berkeley Heights, NJ
-
🧠 - Skills detailed
#Data Governance #Java #YARN (Yet Another Resource Negotiator) #Data Architecture #"ETL (Extract #Transform #Load)" #Delta Lake #Kubernetes #Computer Science #Apache Spark #Leadership #Cloud #Scala #Alation #Kafka (Apache Kafka) #Data Pipeline #Batch #Spark (Apache Spark) #Storage #SQL (Structured Query Language) #Strategy
Role description
Dear Applicant, Please let me know if you are interested. Position: Java Spark Engineer Location: Berkeley heights, NJ (5 days onsite per week) Hire Type: long term Contract Job Description: Responsibilities: β€’ Architect and build scalable, fault-tolerant data pipelines using Apache Spark (Java) β€’ Lead design of batch and streaming ETL/ELT systems handling large data volumes β€’ Deep-dive performance tuning: partitioning strategy, memory management, shuffle/skew optimization, job cost reduction β€’ Set coding standards and lead code/design reviews across the team β€’ Drive technical decisions on data architecture, storage formats, and pipeline orchestration β€’ Mentor mid-level and junior engineers; act as a technical escalation point β€’ Partner with product, analytics, and platform teams to translate requirements into scalable systems β€’ Own production reliability β€” on-call ownership, incident response, root-cause analysis for pipeline failures β€’ Evaluate and introduce new tools/frameworks where they improve the system β€’ Contribute to capacity planning and cost optimization for cluster infrastructure Required Qualifications β€’ Bachelor’s or Master’s degree in Computer Science, Engineering, or related field β€’ 7+ years of professional Java development experience β€’ 5+ years hands-on experience with Apache Spark in production environments β€’ Expert-level understanding of distributed systems: fault tolerance, data locality, shuffle mechanics, resource management β€’ Proven track record designing systems processing terabyte+ scale data β€’ Strong SQL skills and deep familiarity with columnar storage formats (Parquet, ORC, Avro, Delta Lake/Iceberg) β€’ Experience with cluster managers (YARN, Kubernetes) and cloud-managed Spark β€’ Proficiency with Kafka β€’ Strong grasp of CI/CD, containerization, and infrastructure-as-code practices Preferred Qualifications β€’ Experience with Flink or other stream-processing frameworks β€’ Familiarity with data governance, lineage, and quality frameworks β€’ Experience with workflow orchestration at scale β€’ Background in system design for multi-tenant or multi-region data platforms β€’ Prior experience leading a team or acting as a technical lead Soft Skills / Leadership β€’ Excellent communication β€” able to explain technical tradeoffs to non-technical stakeholders β€’ Strong mentorship and coaching ability β€’ Comfortable driving ambiguous, cross-team technical initiatives Behavioral Skills: 1. Good Communication skills 1. 5 days Work from Office at Berkley Heights, NJ 1. Team Player 1. Ability to work in a changing environment 1. Strong problem solving and analytical skills 1. Ability to work independently or within a team 1. Manage day-to-day challenges and communicate developmental risks with the technical team