About This Architecture
HPE Ezmeral Data Fabric 6.2 pipeline architecture integrates Spark, Hive, Kafka, and Apache Drill for unified data ingestion, processing, and analytics. Users and applications connect via multiple access layers—Spark Submit, Beeline, JDBC/ODBC, Kafka Clients, and REST APIs—feeding data into MapR-FS volumes, MapR-DB, and Kafka brokers. YARN orchestrates compute across Spark SQL, Spark Streaming, Hive, MapReduce, and Drill, while Kerberos, LDAP, and volume ACLs enforce security. This architecture demonstrates how to consolidate batch and streaming workloads with built-in failover, tiering, and monitoring, reducing operational complexity for enterprises managing petabyte-scale data lakes. Fork and customize this diagram on Diagrams.so to document your own Ezmeral deployment topology, access patterns, or security posture.