Databricks architect

location_onPune, Maharashtra, Indiaschedule15 घंटे पहले
sync_altकाम का तरीका:हाइब्रिड

पद का विवरण

Key Responsibilities ● · Design, develop, and maintain data pipelines using Databricks, Apache Spark, and cloud-native services. ● · Build and optimize ETL/ELT workflows for large-scale structured and unstructured data. ● · Develop data models and implement data quality, validation, and governance frameworks. ● · Integrate data from multiple sources into a unified Lakehouse architecture. ● · Optimize Spark jobs and Databricks workloads for performance, scalability, and cost efficiency. ● · Implement security controls, access management, and data governance using Unity Catalog. ● · Collaborate with business, analytics, and AI/ML teams to deliver trusted data products. ● · Monitor, troubleshoot, and resolve data pipeline issues. ● · Support CI/CD, DevOps, and infrastructure automation practices. ● · Maintain technical documentation and best practices. Required Technical Skills /Core Technologies ● · Databricks Lakehouse Platform ● · Apache Spark / PySpark ● · Delta Lake ● · SQL ● · Python ● · ETL / ELT Development ● · Data Modeling ● · Data Warehousing ● · Data Quality & Validation ● · Streaming & Real-Time Processing ● Governance & Security ● · Unity Catalog ● · Data Lineage ● · Row-Level Security ● · Access Control & Compliance ● · Data Governance Frameworks ● Cloud & DevOps ● · Azure / AWS / GCP ● · Terraform ● · GitHub Actions / Azure DevOps ● · CI/CD Pipelines ● Analytics & AI ● · Semantic Layers ● · Data Products ● · BI Platforms ● · Machine Learning Support ● · Generative AI & RAG Architectures

Key Responsibilities ● · Design, develop, and maintain data pipelines using Databricks, Apache Spark, and cloud-native services. ● · Build and optimize ETL/ELT workflows for large-scale structured and unstructured data. ● · Develop data models and implement data quality, validation, and governance frameworks. ● · Integrate data from multiple sources into a unified Lakehouse architecture. ● · Optimize Spark jobs and Databricks workloads for performance, scalability, and cost efficiency. ● · Implement security controls, access management, and data governance using Unity Catalog. ● · Collaborate with business, analytics, and AI/ML teams to deliver trusted data products. ● · Monitor, troubleshoot, and resolve data pipeline issues. ● · Support CI/CD, DevOps, and infrastructure automation practices. ● · Maintain technical documentation and best practices. Required Technical Skills /Core Technologies ● · Databricks Lakehouse Platform ● · Apache Spark / PySpark ● · Delta Lake ● · SQL ● · Python ● · ETL / ELT Development ● · Data Modeling ● · Data Warehousing ● · Data Quality & Validation ● · Streaming & Real-Time Processing ● Governance & Security ● · Unity Catalog ● · Data Lineage ● · Row-Level Security ● · Access Control & Compliance ● · Data Governance Frameworks ● Cloud & DevOps ● · Azure / AWS / GCP ● · Terraform ● · GitHub Actions / Azure DevOps ● · CI/CD Pipelines ● Analytics & AI ● · Semantic Layers ● · Data Products ● · BI Platforms ● · Machine Learning Support ● · Generative AI & RAG Architectures

Key Responsibilities ● · Design, develop, and maintain data pipelines using Databricks, Apache Spark, and cloud-native services. ● · Build and optimize ETL/ELT workflows for large-scale structured and unstructured data. ● · Develop data models and implement data quality, validation, and governance frameworks. ● · Integrate data from multiple sources into a unified Lakehouse architecture. ● · Optimize Spark jobs and Databricks workloads for performance, scalability, and cost efficiency. ● · Implement security controls, access management, and data governance using Unity Catalog. ● · Collaborate with business, analytics, and AI/ML teams to deliver trusted data products. ● · Monitor, troubleshoot, and resolve data pipeline issues. ● · Support CI/CD, DevOps, and infrastructure automation practices. ● · Maintain technical documentation and best practices. Required Technical Skills /Core Technologies ● · Databricks Lakehouse Platform ● · Apache Spark / PySpark ● · Delta Lake ● · SQL ● · Python ● · ETL / ELT Development ● · Data Modeling ● · Data Warehousing ● · Data Quality & Validation ● · Streaming & Real-Time Processing ● Governance & Security ● · Unity Catalog ● · Data Lineage ● · Row-Level Security ● · Access Control & Compliance ● · Data Governance Frameworks ● Cloud & DevOps ● · Azure / AWS / GCP ● · Terraform ● · GitHub Actions / Azure DevOps ● · CI/CD Pipelines ● Analytics & AI ● · Semantic Layers ● · Data Products ● · BI Platforms ● · Machine Learning Support ● · Generative AI & RAG Architectures

उल्लिखित कौशल


AI, Banking, Data Analytics, Enterprise Software, Healthcare, Insurance, IT Services, Logistics
Pune, Maharashtra, India

EXL Service Holdings, Inc. is an American data-analytics and digital operations company headquartered in New York City. It combines analytics, AI and business-process management to serve clients in insurance, healthcare, banking, and other industries worldwide.

इस पद के लिए आवेदन करें

Use the application link supplied with this listing to apply to EXL. Check the destination before entering personal information.