What is Azkaban?

Azkaban is a distributed workflow management system created by LinkedIn to tackle the challenges related to Hadoop job dependencies. We encountered situations where jobs needed to run in a specific order, which spanned various applications from ETL processes to data analytics. Following the launch of version 3.0, we established two operational configurations: the standalone "solo-server" mode and the distributed multi-executor mode. The upcoming sections will clarify the differences between these two modes. In the solo server mode, the system utilizes the embedded H2 database, and both the web server and executor server run within the same process, making it suitable for small-scale applications or experimentation. In contrast, the multiple executor mode is designed for more serious production scenarios and necessitates a more sophisticated configuration with a MySQL database set up in a master-slave structure. To improve user experience, it is advisable for the web server and executor servers to operate on different hosts, which helps ensure that upgrades and maintenance do not interfere with service continuity. This architectural choice not only boosts the scalability of Azkaban but also enhances its resilience and efficiency when managing intricate workflows. Ultimately, these operational modes provide flexibility to users while meeting a variety of workflow demands.

Screenshots and Video

Azkaban Screenshot 1

Company Facts

Company Name:
Azkaban
Company Website:
azkaban.github.io

Product Details

Deployment
SaaS
Training Options
Documentation Hub
Support
Web-Based Support

Product Details

Target Company Sizes
Individual
1-10
11-50
51-200
201-500
501-1000
1001-5000
5001-10000
10001+
Target Organization Types
Mid Size Business
Small Business
Enterprise
Freelance
Nonprofit
Government
Startup
Supported Languages
English

Azkaban Categories and Features