Data Engineer

Contract
Hybrid
$60-70/hr
Sunnyvale, CA
Supports visa sponsorship

Job description

Hi,
 I am looking for Data Engineer with one of our client. Please see the job details below and let me know if you would be interested in this role. If interested, please send me a copy of your resume, your contact details, your availability and a good time to connect with you.

Position: Data Engineer(GCP)
Location: Sunnyvale, CA /Hybrid (local to ca)
Duration: 6-12+ Months

Job Description:

·               Proficiency in managing and manipulating huge datasets in the order of terabytes (TB) is essential.

·               Expertise in big data technologies like Hadoop, Apache Spark (Scala preferred), Apache Hive, or similar frameworks on the cloud (GCP preferred, AWS, Azure etc.) to build batch data pipelines with strong focus on optimization, SLA adherence and fault tolerance. 

·               Expertise in building idempotent workflows using orchestrators like Automic, Airflow, Luigi etc. 

·               Expertise in writing SQL to analyze, optimize, profile data preferably in BigQuery or SPARK SQL

·               Strong data modeling skills are necessary for designing a schema that can accommodate the evolution of data sources and facilitate seamless data joins across various datasets

·               Ability to work directly with stakeholders to understand data requirements and translate that to pipeline development / data solution work.

·               Strong analytical and problem-solving skills are crucial for identifying and resolving issues that may arise during the data integration and schema evolution process.

·               Ability to move at rapid pace with quality and start delivering with minimal ramp up time will be crucial to succeed in this initiative.

·               Effective communication and collaboration skills are necessary for working in a team environment and coordinating efforts between different stakeholders involved in the project.

Nice to have:

·               Experience building complex near real time (NRT) streaming data pipelines using Apache Kafka, Spark streaming, Kafka Connect with a strong focus on stability, scalability, and SLA adherence. 

·               Good understanding of REST APIs – working knowledge on Apache Druid, Redis, Elastic search, GraphQL or similar technologies. Understanding of API contracts, building telemetry, stress testing etc.

·               Exposure in developing reports/dashboards using Looker/Tableau

·               Experience in eCommerce domain. 

 

Tech stack: Google cloud, HDFS, SPARK, Scala, Python (optional), Automic/Airflow, BigQuery, Kafka, API, Druid

More information

Minimum education level

Bachelor's

Job skills

GCP

Languages

English

Company overview

company-logo
Caspex

Information Technology and Services·201-1,000 employees

Caspex: Leading International Consulting & Technology Services ggasd At Caspex, we specialize in delivering transformative technology solutions to businesses worldwide. With over 15 years of experience, our multinational team of experts provides innovative, client-focused solutions that help businesses thrive in today’s fast-paced digital landscape. Our commitment to purpose-driven innovation has positioned us as a trusted partner for organizations seeking to explore new opportunities and drive long-term growth. We offer expertise in areas such as AI-driven solutions, big data analytics, open-source development, and cloud and mobile applications, enabling businesses to streamline operations, improve decision-making, and stay ahead of the competition.