Self‑Serving Data Marts Orchestrated by Auto ML-Governed Pipelines

Authors

  • Dr. Ashish Dibouliya

Keywords:

Self-serving data marts, AutoML, data governance, enterprise data warehousing, metadata management, AI-driven analytics

Abstract

Self-serving data marts orchestrated through AutoML-governed pipelines mark a significant advancement in enterprise analytics democratization. This architectural approach creates domain-specific information repositories with automated data preparation, feature engineering, and model development functions accessible to business users without deep technical knowledge. The governance framework applies automated quality controls, lineage tracking, and access management, ensuring data integrity throughout analytical processes. Integration with existing data warehouse systems maintains centralized governance while enabling distributed analytical capabilities, addressing specific business needs. Implementation factors include metadata standardization, processing resource allocation, and organizational change management supporting effective usage. Technical elements comprise automated data profiling, dynamic transformation creation, and continuous quality monitoring throughout pipeline operation. The orchestration layer manages complex workflows while implementing appropriate error handling and recovery mechanisms. Enterprises implementing these frameworks report significant enhancements in analytical responsiveness, resource utilization effectiveness, and business coordination compared to conventional centralized models. This balanced approach resolves conflicting requirements between governance standardization and analytical adaptability, establishing durable foundations for growing self-service functions while preserving appropriate supervision across increasingly intricate data landscapes.

References

Michael Segner (2023) Data Pipeline Architecture Explained: 6 Diagrams and Best Practices. https://www.montecarlodata.com/blog-data-pipeline-architecture-explained/

Awez Syed, Amit Kara (2021) 5 Steps to Implementing Intelligent Data Pipelines With Delta Live Tables. https://www.databricks.com/blog/2021/09/08/5-steps-to-implementing-intelligent-data-pipelines-with-delta-live-tables.html

Sadig Akhund (2023) Computing Infrastructure and Data Pipeline for Enterprise-scale Data Preparation: A Scalability Optimization Study. https://www.researchgate.net/publication/370301416

What is a Data Mart?. https://www.qlik.com/us/data-warehouse/data-mart

Automated Data Integration: Concepts & Strategies. https://nexla.com/data-engineering-best-practices/automated-data-integration/

Patrycja Zajac Dataflow vs. Datamart – when to use them to enhance your Power BI solutions?. https://10senses.com/

Downloads

Published

2025-09-18

How to Cite

Self‑Serving Data Marts Orchestrated by Auto ML-Governed Pipelines. (2025). London Journal of Research In Computer Science and Technology, 25(3), 9-28. https://journalspress.uk/index.php/LJRCST/article/view/1578