Back to Jobs
Enable Data
AI & Machine Learning Just now

AI Engineer - Production AI Applications with Azure Databricks

Enable Data
IndiaIndia
Full-time
Not Disclosed
Mid-Level

Job Description

Key Skills Required

Master these to land this role

ML and Generative AIPySparkAzure AI ServicesAzure DatabricksLangChain

Want to know if you're a match for this job?

Calculate My Match Score

This role focuses on building production-ready AI applications and deploying them on Azure Databricks and Azure cloud infrastructure. You will work end-to-end: from data ingestion and model integration to scalable deployment, monitoring, and ongoing optimization.

The expectation is to convert AI ideas into reliable, governed, and cost-efficient applications that run in production. You will design data and AI pipelines, integrate models (including ML and Generative AI), and deploy them using Databricks workflows and Azure-native services.

Success in this role requires strong hands-on experience with Azure Databricks, Python, SQL, and Azure services, along with a clear understanding of how AI systems fail in production—and how to prevent it. You will collaborate closely with data scientists, platform engineers, and business stakeholders to ensure AI applications are usable, scalable, and maintainable beyond the first release.

Key Responsibilities

  • Design and build end-to-end data and AI pipelines using Azure Databricks.
  • Develop robust ETL/ELT workflows using Python (PySpark) and SQL.
  • Implement CI/CD pipelines for Databricks deployments (jobs, notebooks, workflows).
  • Integrate Databricks with Azure services (Data Lake, Blob Storage, Key Vault, Azure OpenAI, Azure Functions, etc.).
  • Optimize jobs for performance, cost, and reliability.
  • Build reusable, modular code.
  • Collaborate with data scientists and platform teams to move models from experimentation to production.
  • Implement logging, monitoring, and error handling for production pipelines.
  • Develop and deploy ML and Generative AI models (LLMs, embeddings, RAG pipelines) for NLP, computer vision, and predictive analytics.
  • Fine-tune LLMs using LoRA/QLoRA and integrate with Azure OpenAI or Hugging Face models.
  • Implement vector search and retrieval pipelines using FAISS or Azure Cognitive Search.
  • Ensure responsible AI practices, including bias detection and model governance.

Good to Have (Strong Advantage)

  • Experience with ML and Generative AI workloads on Databricks.
  • RAG, embeddings, or inference pipelines.
  • Terraform / ARM / Bicep for infrastructure.
  • Databricks Asset Bundles.
  • Airflow or ADF orchestration.
  • Production monitoring and cost optimization experience.
  • Knowledge of LangChain or similar frameworks for AI application development.
  • Experience with Azure AI services (Azure Machine Learning, Azure Cognitive Services).

Databricks

  1. Design and build pipelines for ingesting data into Databricks from various sources like SAP, website scraping. Core technical skill will be PySpark and PySQL.
  2. Write and do scenario-based testing, edge cases, and discuss with stakeholders to finalize the necessary code changes and acceptance criteria.
  3. Understanding of the Databricks Unity Catalog permissions model and how to use OBO tokens and configure Service Principals for Databricks Agents, Genie Spaces, Machine Learning, and Foundation Models.
  4. Knowledge of Databricks Apps. Ability to build and Deploy Databricks Apps using Visual Studio Code.
  5. Sending custom notifications from Databricks using APIs to custom Team Channel Webhooks.

AI

  1. Ability to use FastAPI / React front end and build minimal Single Page Applications (SPAs) using Chat UI interfaces and Claude Models/APIs.
  2. Ability to use defined System Prompts, User Prompts, etc. (i.e., skills in prompt engineering).
  3. Basic knowledge of Token Economics and ability to use rule enforcement in prompts in addition to standard NLP language.
  4. Ability to reengineer code and map it to the requirement specifications from users.
  5. Ability to use Graphs to detect interdependencies between rules and circular references.
  6. Ability to use Claude Vision API in addition to standard text processing.
  7. Understand and develop Retrieval Augmented Generation specifically for Databricks like AI Search indexes, Hierarchical Chunking, etc.
  8. Know how to implement basic guardrails related to AI safety.
  9. Understand various caching mechanisms.

Requirements

  • Azure Databricks (jobs, workflows, clusters, Unity Catalog preferred).
  • Python (PySpark-heavy, not just pandas).
  • SQL (complex joins, window functions, analytical queries).
  • Azure Cloud (ADLS Gen2, ADF, Key Vault, IAM concepts).
  • Pipeline orchestration & CI/CD, environment promotion.
  • Azure DevOps.
  • Strong understanding of ML lifecycle and MLOps best practices.
  • Experience with model deployment using MLflow or similar frameworks.

How would you rate this job post?

See what other professionals think about this role.

banner

Enable Data is an enterprise-grade technology consultancy engineered to orchestrate massive-scale data engineering, cloud infrastructure, and advanced software solutions. Operating as a highly integrated digital ecosystem partner, the company eliminates the operational friction of complex legacy systems by seamlessly deploying modern big data frameworks, including robust Databricks and Snowflake architectures. Moving beyond rigid staff augmentation models, Enable Data provides end-to-end managed projects and specialized consulting to rapidly build predictive analytics platforms, distributed rules engines, and continuous integration (CI) pipelines. Under the hood, their highly scalable cloud methodologies natively handle complex workloads, allowing enterprises in healthcare, finance, and consumer services to execute deep machine learning algorithms and advanced data science initiatives in real-time. What sets Enable Data apart is its uncompromising dedication to frictionless digital transformation; by bridging the gap between raw data ingestion and actionable predictive intelligence, the firm empowers organizational leaders to radically accelerate their cloud adoption and maximize the value of their core digital assets.

Safety First

  • Never pay for a job application.
  • Do not share sensitive bank info.
  • Verify the client before starting work.
Learn More
AI Engineer - Production AI Applications with Azure Databricks at Enable Data | HireSkys