Description
Technology->Analytics - Packages->Python - Big Data,Technology->Big Data - Data Processing->PySpark Design, develop, and maintain scalable batch/stream data pipelines using Python and PySpark in distributed environments. Implement efficient transformations, aggregations, and joins on large datasets while ensuring performance and cost optimization. Write optimized SQL for data extraction, validation, and reconciliation across multiple sources.
Build reusable, testable modules and follow engineer…
Employer contacts (email/phone/telegram) are hidden from the public preview —
send your CV, and we will connect you directly.