Databricks, Pyspark
InfosysJob Description
Databricks, Pyspark
Technology->Analytics - Solutions->SQL Server - Analytics Technology->Big Data - Data Processing->PySpark Technology->Data Engineering->Databricks Preferred Qualifications: • Experience designing lakehouse-style architectures and organizing curated layers for analytics readiness. • Strong experience with Spark optimization techniques and handling large-scale datasets efficiently. • Ability to build reusable PySpark utilities/frameworks for ingestion, transformation, and validation patterns. • Experience with orchestration and scheduling approaches for dependable pipeline execution and recovery. • Proven track record of technical leadership: mentoring, conducting reviews, and driving engineering best practices. Key Responsibilities: • Lead the development of end-to-end data pipelines on Databricks using PySpark for batch and incremental processing. • Design scalable data models and curated datasets to support analytics and downstream consumption. • Write and optimize advanced SQL for transformations, validations, and performance-critical queries. • Implement robust data quality checks, reconciliation logic, and monitoring to ensure trusted datasets. • Tune Spark jobs for performance and cost efficiency (partitioning, caching, file formats, cluster sizing). • Establish coding standards, reusable frameworks, and review practices to improve maintainability. • Collaborate with stakeholders to translate requirements into technical designs and delivery plans. • Troubleshoot production issues, perform root-cause analysis, and drive preventive improvements. • Mentor team members and provide technical guidance across design, implementation, and optimization. Minimum Qualifications: • BTECH, MTECH, MCA, or MSC in Computer Science, Information Technology, or a related field. • 6–8 years of overall experience in data engineering or large-scale data processing roles. • Strong hands-on experience with PySpark for distributed data processing and transformation logic. • Strong hands-on experience with Databricks for building, running, and managing data workloads. • Proficiency in Advanced SQL including complex joins, window functions, and query optimization. • Experience building reliable pipelines with strong focus on data quality, performance, and stabilityJob role
Job requirements
About company
Similar jobs you can apply for
Accounts / FinanceRisk Officer
Teamspace Financial Services Private Limited
Quality Control Manager
nice Neotech Medical Systems Pvt LtdQuality Control Engineer
Buma Trident
Quality Control Inspector
Plastronics India
Sales Engineer
Creative Packaging Systems
Engineering Trainee
Gtech Drives & ControlsYou can expect a minimum salary of 0 INR. The salary offered will depend on your skills, experience and performance in the interview.
The candidate should have completed the required education and people who have 6 to 8 years are eligible to apply for this job. You can apply for more jobs in Chennai to get hired quickly.
The candidate should have sound communication skills and sound communication skills for this job.
Both Male and Female candidates can apply for this job.
No, it's not a work from home job and can't be done online. You can explore and apply for other work from home jobs in Chennai at apna.
No work-related deposit needs to be made during your employment with the company.
Go to the apna app and apply for this job. Click on the apply button and call HR directly to schedule your interview.
The last date to apply for this job is . For more details, download apna app and find Full Time jobs in Chennai . Through apna, you can find jobs in 64 cities across India. Join NOW!