Open role

Big Data / PySpark Engineer

Thane / Hybrid · Full-time

Design and optimize distributed data pipelines and PySpark jobs for analytics, ML features, and HPC-adjacent workloads.

You will build scalable Spark pipelines, tune cluster performance, and partner with AI and cloud teams on lakehouse and high-throughput data platforms.

Experience with PySpark, distributed computing, and cloud data services is essential; HPC exposure is a strong plus.

Apply

Any email is fine. Verify with OTP, share your LinkedIn profile (required), and optionally attach a resume.

Any email is accepted (including Gmail). Verify with OTP. LinkedIn profile is required.