Search Jobs

Search by job, company or skills

Python, Pyspark Developer

Python, Pyspark Developer

Infosys Limited
Early Applicant
  • Posted a month ago
  • Be among the first 10 applicants

Job Description

Job Description:

  • Build and scale data driven solutions that power smarter decisions
  • In this role you ll design and deliver high performance data processing pipelines using Python and PySpark working closely with data engineers analysts and product teams to turn raw data into reliable actionable insights
  • You ll contribute to a collaborative environment where clean code thoughtful design and continuous improvement are valued
  • If you enjoy solving complex data challenges optimizing distributed workloads and delivering production ready systems that make a real impact this is a great opportunity to grow your expertise while helping teams move faster with trustworthy data

Key Responsibilities:

  • Design develop and maintain scalable batch stream data pipelines using Python and PySpark in distributed environments
  • Implement efficient transformations aggregations and joins on large datasets while ensuring performance and cost optimization
  • Write optimized SQL for data extraction validation and reconciliation across multiple sources
  • Build reusable testable modules and follow engineering best practices code reviews unit testing documentation
  • Troubleshoot production issues perform root cause analysis and implement long term fixes and monitoring improvements
  • Collaborate with stakeholders to translate requirements into technical designs delivery plans and measurable outcomes
  • Ensure data quality through validation checks anomaly detection patterns and consistent schema management
  • Contribute to continuous improvement of development standards performance benchmarks and pipeline reliability

Technical Requirements:

  • Technology Analytics Packages Python Big Data Technology Big Data Data Processing PySpark

Additional Responsibilities:

  • Bachelor s degree in Computer Science Engineering or a related field or equivalent practical experience
  • 5 9 years of hands on experience in software development and or data engineering roles
  • Strong proficiency in Python with experience building production grade applications or data workflows
  • Strong proficiency in PySpark including DataFrame APIs optimization techniques and distributed processing concepts
  • Working knowledge of SQL for complex queries data analysis and validation
  • Experience delivering reliable solutions with attention to performance scalability and maintainability

Preferred Skills:

Technology->Big Data - Data Processing->PySpark

More Info

Job Type:
Industry:
Employment Type:

Key Skills

About Company

Similar Jobs

7-9 yrs
Hyderabad, India
Skills:
PysparkCloud StorageGitGcpPostgresDataprocSqlPythonBig QueryAirflowBigQuery SQL
2-5 yrs
Hyderabad, India
Skills:
ErpAutomated TestingData ModelingSqlELTData QualityGitPandasDockerAzurePythonEtlAirflowPolarsReconciliationproduction troubleshootingCI CDDuckDBfinance data
3-5 yrs
Hyderabad, India
Skills:
PysparkSqlTensorflowNosqlNumpyPandasPytorchGcpDockerDatabricksFastAPIAzureKubernetesPythonAWSMLflowTransformersDelta Lake
10-12 yrs
Hyderabad, India
Skills:
Data ArchitecturePythonMicrosoft AI StackAWS BedrockMicrosoft 365 CopilotGitHub CopilotCopilot CoworkClaude APICopilot StudioAnthropic Claude EcosystemIntegration patternsResponsible AI principles
6-8 yrs
Hyderabad, India
Skills:
AccessibilityDatadogNosqlReactGitTypescriptJavascriptJUnitTddDockerDesign PatternsPythonAWSSqlJenkinsGcpResponsive DesignSplunkFastAPICucumberAzurePerformance Testing Toolsfrontend performance optimizationsoftware engineering principlessecure coding practicesStreamlitcross-browser compatibilityGraphQL APIsCI CD pipelinesUI UX principles