MeeBoss

Data Engineer

โญ - Featured Role | Apply direct with Data Freelance Hub
This role is for a Data Engineer (GCP) in Sunnyvale, CA, for 6-12+ months at a pay rate of "X". Candidates must have expertise in big data technologies, SQL, and data modeling, with eCommerce experience preferred.
๐ŸŒŽ - Country
United States
๐Ÿ’ฑ - Currency
$ USD
-
๐Ÿ’ฐ - Day rate
560
-
๐Ÿ—“๏ธ - Date
July 23, 2026
๐Ÿ•’ - Duration
More than 6 months
-
๐Ÿ๏ธ - Location
Hybrid
-
๐Ÿ“„ - Contract
Unknown
-
๐Ÿ”’ - Security
Unknown
-
๐Ÿ“ - Location detailed
Sunnyvale, CA
-
๐Ÿง  - Skills detailed
#Azure #Hadoop #GraphQL #Python #Big Data #Data Integration #Data Modeling #Data Pipeline #Apache Kafka #Cloud #Redis #Spark SQL #Looker #Spark (Apache Spark) #Data Engineering #Tableau #GCP (Google Cloud Platform) #HDFS (Hadoop Distributed File System) #AWS (Amazon Web Services) #SQL (Structured Query Language) #Scala #Kafka (Apache Kafka) #Airflow #BigQuery #Datasets #Batch #Apache Spark #REST (Representational State Transfer) #Luigi #API (Application Programming Interface) #Apache Hive #REST API
Role description
Hi, I am looking for Data Engineer with one of our client. Please see the job details below and let me know if you would be interested in this role. If interested, please send me a copy of your resume, your contact details, your availability and a good time to connect with you. Position: Data Engineer(GCP) Location: Sunnyvale, CA /Hybrid (local to ca) Duration: 6-12+ Months Job Description: ยท Proficiency in managing and manipulating huge datasets in the order of terabytes (TB) is essential. ยท Expertise in big data technologies like Hadoop, Apache Spark (Scala preferred), Apache Hive, or similar frameworks on the cloud (GCP preferred, AWS, Azure etc.) to build batch data pipelines with strong focus on optimization, SLA adherence and fault tolerance. ยท Expertise in building idempotent workflows using orchestrators like Automic, Airflow, Luigi etc. ยท Expertise in writing SQL to analyze, optimize, profile data preferably in BigQuery or SPARK SQL ยท Strong data modeling skills are necessary for designing a schema that can accommodate the evolution of data sources and facilitate seamless data joins across various datasets ยท Ability to work directly with stakeholders to understand data requirements and translate that to pipeline development / data solution work. ยท Strong analytical and problem-solving skills are crucial for identifying and resolving issues that may arise during the data integration and schema evolution process. ยท Ability to move at rapid pace with quality and start delivering with minimal ramp up time will be crucial to succeed in this initiative. ยท Effective communication and collaboration skills are necessary for working in a team environment and coordinating efforts between different stakeholders involved in the project. Nice to have: ยท Experience building complex near real time (NRT) streaming data pipelines using Apache Kafka, Spark streaming, Kafka Connect with a strong focus on stability, scalability, and SLA adherence. ยท Good understanding of REST APIs โ€“ working knowledge on Apache Druid, Redis, Elastic search, GraphQL or similar technologies. Understanding of API contracts, building telemetry, stress testing etc. ยท Exposure in developing reports/dashboards using Looker/Tableau ยท Experience in eCommerce domain. Tech stack: Google cloud, HDFS, SPARK, Scala, Python (optional), Automic/Airflow, BigQuery, Kafka, API, Druid