Open role
Big Data / PySpark Engineer
Thane / Hybrid · Full-time
Design and optimize distributed data pipelines and PySpark jobs for analytics, ML features, and HPC-adjacent workloads.
You will build scalable Spark pipelines, tune cluster performance, and partner with AI and cloud teams on lakehouse and high-throughput data platforms.
Experience with PySpark, distributed computing, and cloud data services is essential; HPC exposure is a strong plus.
Apply
Any email is fine. Verify with OTP, share your LinkedIn profile (required), and optionally attach a resume.