Brillio

Databricks Data Specialist - R01569707

Brillio  •  Bengaluru, IN (Onsite)  •  9 days ago
Apply
AI can make mistakes so check important info. Chat history is never stored.

Job Description

Data Specialist

Primary Skills

    Databricks Engineer

    We are seeking a highly skilled Databricks Engineer to design, develop, and optimize scalable data engineering and analytics solutions on the Databricks Lakehouse Platform The ideal candidate will possess strong expertise in Databricks, PySpark, and SQL, with hands-on experience building batch and real-time data pipelines, implementing Lakehouse architectures, and ensuring data governance and performance optimization.

    Key Responsibilities

  • Design, develop, and maintain end-to-end data pipelines using Databricks and PySpark
  • Build and implement Lakehouse architectures utilizing Bronze, Silver, and Gold data layers.
  • Develop and manage Delta Lake solutions with ACID transactions, schema enforcement, and data reliability features.
  • Create, monitor, and optimize Delta Live Tables (DLT) pipelines.
  • Implement scalable and efficient data ingestion processes using Auto Loader
  • Develop and manage real-time data processing solutions using Structured Streaming
  • Orchestrate, schedule, and monitor data workflows using Databricks Workflows
  • Design and implement Lakehouse data models to support reporting, analytics, and business intelligence requirements.
  • Establish and enforce data governance, security, and access controls using Unity Catalog
  • Optimize Spark jobs, SQL queries, and overall platform performance to ensure efficiency and scalability.
  • Collaborate with cross-functional teams, including data analysts, architects, and business stakeholders, to deliver high-quality data solutions.
  • Required Skills (Must Have)

  • Databricks Platform
  • Delta Lake
  • Delta Live Tables (DLT)
  • Unity Catalog
  • Databricks Workflows
  • PySpark and Apache Spark
  • Structured Streaming
  • Auto Loader
  • SQL
  • Lakehouse Data Modeling
  • Strong understanding of data engineering best practices and scalable data architectures
  • Preferred Skills (Good to Have)

    Azure Ecosystem

  • Azure Data Factory (ADF)
  • Azure Synapse Analytics
  • Microsoft Purview
  • Microsoft Fabric
  • AWS Ecosystem

  • AWS Glue
  • AWS Lambda
  • AWS Step Functions
  • Data Engineering & Integration

  • Apache Airflow
  • DBT
  • Fivetran
  • Informatica
  • Streaming & Analytics

  • Apache Kafka
  • Power BI
  • Data Governance

  • Collibra
  • Alation
  • GCP

  • BigQuery
  • Qualifications

  • Bachelor's or Master's degree in Computer Science, Data Engineering, Information Technology, or a related discipline.
  • Proven experience in designing and implementing cloud-based data engineering solutions and scalable data pipelines.
  • Strong analytical, troubleshooting, and problem-solving capabilities.
  • Experience working in agile and collaborative environments.
  • Excellent communication and stakeholder management skills.
  • Preferred Candidate Profile

  • Hands-on experience with modern Lakehouse architectures and enterprise-scale data platforms.
  • Strong understanding of data governance, security, and compliance frameworks.
  • Experience delivering both batch and real-time data processing solutions.
  • Ability to work independently while collaborating effectively across global teams.
  • Key Technologies

    Databricks | PySpark | Apache Spark | Delta Lake | Delta Live Tables (DLT) | Unity Catalog | Structured Streaming | Auto Loader | SQL | Lakehouse Architecture | Azure | AWS | Airflow | Kafka | Power BI

Specialization

  • Databricks Engineering: Lead Data Engineer

Job requirements

    Databricks Engineer

    We are seeking a highly skilled Databricks Engineer to design, develop, and optimize scalable data engineering and analytics solutions on the Databricks Lakehouse Platform The ideal candidate will possess strong expertise in Databricks, PySpark, and SQL, with hands-on experience building batch and real-time data pipelines, implementing Lakehouse architectures, and ensuring data governance and performance optimization.

    Key Responsibilities

  • Design, develop, and maintain end-to-end data pipelines using Databricks and PySpark
  • Build and implement Lakehouse architectures utilizing Bronze, Silver, and Gold data layers.
  • Develop and manage Delta Lake solutions with ACID transactions, schema enforcement, and data reliability features.
  • Create, monitor, and optimize Delta Live Tables (DLT) pipelines.
  • Implement scalable and efficient data ingestion processes using Auto Loader
  • Develop and manage real-time data processing solutions using Structured Streaming
  • Orchestrate, schedule, and monitor data workflows using Databricks Workflows
  • Design and implement Lakehouse data models to support reporting, analytics, and business intelligence requirements.
  • Establish and enforce data governance, security, and access controls using Unity Catalog
  • Optimize Spark jobs, SQL queries, and overall platform performance to ensure efficiency and scalability.
  • Collaborate with cross-functional teams, including data analysts, architects, and business stakeholders, to deliver high-quality data solutions.
  • Required Skills (Must Have)

  • Databricks Platform
  • Delta Lake
  • Delta Live Tables (DLT)
  • Unity Catalog
  • Databricks Workflows
  • PySpark and Apache Spark
  • Structured Streaming
  • Auto Loader
  • SQL
  • Lakehouse Data Modeling
  • Strong understanding of data engineering best practices and scalable data architectures
  • Preferred Skills (Good to Have)

    Azure Ecosystem

  • Azure Data Factory (ADF)
  • Azure Synapse Analytics
  • Microsoft Purview
  • Microsoft Fabric
  • AWS Ecosystem

  • AWS Glue
  • AWS Lambda
  • AWS Step Functions
  • Data Engineering & Integration

  • Apache Airflow
  • DBT
  • Fivetran
  • Informatica
  • Streaming & Analytics

  • Apache Kafka
  • Power BI
  • Data Governance

  • Collibra
  • Alation
  • GCP

  • BigQuery
  • Qualifications

  • Bachelor's or Master's degree in Computer Science, Data Engineering, Information Technology, or a related discipline.
  • Proven experience in designing and implementing cloud-based data engineering solutions and scalable data pipelines.
  • Strong analytical, troubleshooting, and problem-solving capabilities.
  • Experience working in agile and collaborative environments.
  • Excellent communication and stakeholder management skills.
  • Preferred Candidate Profile

  • Hands-on experience with modern Lakehouse architectures and enterprise-scale data platforms.
  • Strong understanding of data governance, security, and compliance frameworks.
  • Experience delivering both batch and real-time data processing solutions.
  • Ability to work independently while collaborating effectively across global teams.
  • Key Technologies

    Databricks | PySpark | Apache Spark | Delta Lake | Delta Live Tables (DLT) | Unity Catalog | Structured Streaming | Auto Loader | SQL | Lakehouse Architecture | Azure | AWS | Airflow | Kafka | Power BI

Brillio

About Brillio

Brillio is one of the fastest growing digital technology service providers and a partner of choice for many Fortune 1000 companies seeking to turn disruption into a competitive advantage through innovative digital adoption. Founded in 2014 as a digitally native full-service digital transformation services and consulting firm, we apply our expertise in customer experience transformation, data analytics, Artificial Intelligence (AI), platform and product engineering, cloud infrastructure, and security to help clients quickly innovate for growth, create digital products, build service platforms, and drive smarter, data-driven performance.

Headquartered in Dallas, Texas, we are powered by a diverse global team of world-class professionals across the U.S., the UK, Romania, Canada, Mexico and India, and are certified a Great Place to Work®. We help clients harness the transformative potential of the four superpowers of technology: cloud computing, Internet of Things (IoT), AI, and mobility. We bring deep expertise across the full spectrum of digital capabilities:

• Accelerating customer experience transformation to drive growth, customer advocacy, and superior customer experience

• Powering intelligent enterprises by harnessing the potential of data, analytics, and AI

• Crafting products of relevance with a product mindset and high-performance engineering

• Enabling enterprise agility with resilient cloud infrastructure and security

To learn more, please visit https://www.brillio.com/ and follow us here or @brillioglobal.

Industry
Unknown
Company Size
5,001-10,000 employees
Headquarters
Dallas, Texas
Year Founded
Unknown
Social Media