Search by job, company or skills

Azure Data Architect

10-15 Years
Quick Apply
  • Posted 7 hours ago
  • Over 50 applicants have applied

Job Description

We are looking for a seasoned Data Architect with deep expertise in Spark to lead the design and implementation of modern data processing solutions. The ideal candidate will have extensive experience in distributed data processing, large-scale data pipelines, and cloud-native data platforms.

This is a strategic role focused on building scalable, fault-tolerant, and high-performance data systems.

Key Responsibilities:

  • Architect, design, and implement large-scale data pipelines using Spark (batch and streaming).
  • Optimize Spark jobs for performance, cost-efficiency, and scalability.
  • Define and implement enterprise data architecture standards and best practices.
  • Guide the transition from traditional ETL platforms to Spark-based solutions.
  • Lead the integration of Spark-based pipelines into cloud platforms (Azure Fabric/Spark pools).
  • Establish and enforce data architecture standards, including governance, lineage, and quality.
  • Mentor data engineering teams on best practices with Spark (e.g., partitioning, caching, join strategies).
  • Implement and manage CI/CD pipelines for Spark workloads using tools like GIT or DevOps.
  • Ensure robust monitoring, alerting, and logging for Spark applications.

Required Skills & Qualifications:

  • 10+ years of experience in data engineering, with 7+ years of hands-on experience with Apache Spark (PySpark/Scala).
  • Proficiency in Spark optimization techniques, Monitoring, Caching, advanced SQL, and distributed data design.
  • Experience with Spark on Databricks and Azure Fabric.
  • Solid understanding of Delta LakeSpark Structured Streaming, and data pipelines.
  • Strong experience in cloud platforms ( Azure).
  • Proven ability to handle large-scale datasets (terabytes to petabytes).
  • Familiarity with data lakehouse architectures, schema evolution, and data governance.
  • Candidate to be experienced in Power BI, with at least 3+ years of experience.

Preferred Qualifications:

  • Experience implementing real-time analytics using Spark Streaming or Structured Streaming.
  • Certifications in Databricks, Fabric or Spark would be a plus.

More Info

About Company

Aptean is one of the world’s leading providers of purpose-built, industry-specific software that helps manufacturers and distributors effectively run and grow their businesses. With both cloud and on-premise deployment options, Aptean’s products, services and unmatched expertise help businesses of all sizes to be Ready for What’s Next, Now®. Aptean is headquartered in Alpharetta, Georgia and has offices in North America, Europe and Asia-Pacific. To learn more about Aptean and the markets we serve, visit www.aptean.com.

Job ID: 121607557

Similar Jobs

Bengaluru, India

Skills:

Azure DatabricksData IntegrationAzure SynapseBatchEdwPowershell ScriptingOdsAzure DevOpsData ModellingPower BiData WarehousingAzure SqlPython ProgrammingReal-TimeAzure Data FactorySparkEtlMPP database architectureBusiness IntelligenceMicrosoft Entra IDData storage designLakehouse ImplementationSQL OptimizationADLSDelta Lake

Bengaluru, India

Skills:

SqlDatabase DesignTriggersAutomationReporting ToolsDashboard DevelopmentMicrosoft FabricStored Procedures

Bengaluru, India

Skills:

Azure Data FactorySqlAzure Synapse AnalyticsData WarehousingHadoopPysparkKafkaAzure DatabricksData SecurityEtlNosqlELTWorkflowsData IntegrationData GovernanceData TransformationAzure Data LakeMicrosoft AzureSparkAzure Blob StorageSQL WarehouseData Quality FrameworksStructured StreamingAzure SQL DatabaseCompliance StandardsDelta Live TablesUnity CatalogData Pipeline Architecture

Beware of Scammers

We don’t charge money for job offers