Leanpub Header

Skip to main content

Data Management Engineering

Architecture, Quality, Implementation

Data Management Engineering
Passage 1 — Chapter 9, "Data Security": a definition that sets the technical tone immediately

A hacker is someone skilled at finding undocumented techniques and loopholes in the tangled architecture of complex information systems; by intent, a "white hat" looks for such loopholes in order to strengthen the system, while a "black hat" uses them to steal or to cause harm.

Passage 2 — Chapter 14, "Metadata Management": an image that explains the whole topic in one paragraph

A vivid illustration: an enormous document archive with no index at all — the shelves are full, but there is no way to find out what sits on them short of examining every single box by hand. That is exactly what an organization looks like when it has piled up mountains of data without also taking care of metadata — the information physically exists, but it cannot be used systematically, since the only way to find what is needed is to already know where it sits. Knowledge about data is always scattered: in a large company, one person carries the structure of a single database in their head, another the rules of a single integration, a third the history of a single metric, and nobody holds the complete picture.

Passage 3 — Chapter 15, "Data Quality Management, Part 1": why "quality" is an empty word without a yardstick

A postal address missing an apartment number works perfectly well for a mass catalog mailing and works terribly for a courier who has to knock on the right door — and in both cases it is the very same row in the very same database, only the yardstick applied to it differs.

Passage 4 — Chapter 21, "Organizational Change Management": the book's closing summary

Data governance, the coordinating hub of eleven knowledge areas this book opened with, stays an empty frame until specific people come to value the new way of working with data through their own experience — which is exactly why a book that began by mapping the circle of disciplines around that hub fittingly closes not with another technique or tool, but with a conversation about the person without whose deliberate participation no structure ever becomes a practice.

Minimum price

$19.99

$24.99

You pay

Author earns

$

Also available for 1 book credit with a Reader Membership

PDF
EPUB
About

About

About the Book

Short Description

Technical practitioners usually know a single tool well — a particular DBMS, an ETL platform, a BI system — but rarely see how those tools fit together into one coherent data management system for the organization. The result is a set of decisions that are locally correct but systemically inconsistent: a data model with no shared glossary of terms behind it, an integration pipeline with no lineage tracking, a quality metric nobody ties back to business consequences. This book supplies exactly the shared vocabulary and the map of knowledge areas that usually goes missing between individual tools.

The book's approach rests on the systematized practices and recommendations of DAMA International, an international professional association, translated into the language of concrete engineering decisions — architecture-selection criteria, model-design checklists, step sequences for rolling out quality controls — rather than a retelling of management meetings or vendor marketing promises. A single running technical example, an online store whose data architecture, models, pipelines, and organizational structure run through the whole book, lets the reader see how decisions made in different chapters add up to one coherent system.

What the Book Covers

The book is organized into 7 parts and 21 chapters.

Part I. Foundations of Data Management (Chapters 1–4) — DAMA concepts and framework, data handling ethics, and data governance: strategy, implementation, tools, standards, and metrics.

Part II. Architecture and Modeling (Chapters 5–7) — data architecture, the entity-relationship approach to modeling, and alternative notations (multidimensional, object-oriented, fact-based, temporal, NoSQL) along with criteria for choosing among them.

Part III. Storage, Security, and Integration (Chapters 8–10) — database storage and operations, data security, and integration and interoperability (ETL/ELT, enterprise service bus, APIs).

Part IV. Content, Reference, and Master Data (Chapters 11–12) — document and content management, electronic discovery, reference and master data, entity resolution.

Part V. Analytics, Metadata, and Quality (Chapters 13–16) — data warehousing and business intelligence, metadata management, and measuring and implementing data quality along with its ties to other knowledge areas.

Part VI. Big Data and Data Science (Chapters 17–18) — strategy, sources, tools, techniques, and governance for working with large volumes of loosely structured data.

Part VII. Maturity, Organization, and Change (Chapters 19–21) — data management maturity assessment, organizational structure and roles, and managing organizational change during the rollout of technical initiatives.

Who This Book Is For

Data engineers, data architects, analysts, database administrators, data stewards, and developers. A basic familiarity with databases and the software development lifecycle is assumed — the book deliberately stays focused on technical implementation rather than the management level, which is the subject of a separate, companion book for executives.

What Sets This Book Apart

  • Engineering language that never dilutes the substance. Every chapter gives the technical reader concrete selection criteria — ETL versus ELT, the relational model versus NoSQL, the Inmon architectural path versus Kimball's — rather than a retelling of vendor documentation or general principles detached from an actual decision.
  • Consistent terminology from cover to cover. All twenty-one chapters are built on one shared set of concepts, consolidated into a glossary at the end of the book: a term defined once is used consistently afterward, without drift between chapters.
  • One running example that grows with the book. The same online store — with the same "Customer," "Product," and "Order" entities — follows the reader from data storage in Chapter 8 through organizational change in Chapter 21, so the technical decisions of different chapters are visible within a single coherent system rather than as disconnected examples pulled from unrelated domains.

Author

About the Author

Andrii Bogdanovych

Andrii BOGDANOVYCH combines senior public-service management experience with engineering expertise in artificial intelligence — a combination rare in the Ukrainian market, and one that directly shapes this book's approach: discussing AI governance in language equally accessible to public-sector officials, corporate boards, and technical teams.

Deputy Head for Digital Development, Digital Transformation, and Digitalization (CDTO) of the State Energy Supervision Inspectorate of Ukraine (since 2022); he previously held the equivalent position at the State Ecological Inspectorate of Ukraine, and the position of Deputy Head of the Kherson Regional State Administration — in both roles leading digitalization, cybersecurity, and critical-infrastructure protection efforts, respectively at the level of a central executive authority and at the regional level.

He is the author of four training programs: "AI Management and Governance in the Organization: NIST AI RMF 1.0 and ISO/IEC 42001," data governance for executives, data governance for technical practitioners, and building organizational cybersecurity under NIST CSF 2.0 and ISO/IEC 27001.

A practicing Python/AI developer, he designs RAG systems and autonomous agentic solutions built on LLM APIs, publishes and maintains open-source libraries on PyPI, and administers his own server infrastructure. He holds a Master's degree in Public Administration (Taras Shevchenko National University of Kyiv), a Master's degree in Law (Academy of Advocacy of Ukraine), and a Bachelor's degree in Computer Science (Vadym Hetman Kyiv National Economic University).

The Leanpub 60 Day 100% Happiness Guarantee

Within 60 days of purchase you can get a 100% refund on any Leanpub purchase, in two clicks.

See full terms...

Earn $8 on a $10 Purchase, and $16 on a $20 Purchase

We pay 80% royalties on purchases of $7.99 or more, and 80% royalties minus a 50 cent flat fee on purchases between $0.99 and $7.98. You earn $8 on a $10 sale, and $16 on a $20 sale. So, if we sell 5000 non-refunded copies of your book for $20, you'll earn $80,000.

(Yes, some authors have already earned much more than that on Leanpub.)

In fact, authors have earned over $15 million writing, publishing and selling on Leanpub.

Learn more about writing on Leanpub

Free Updates. DRM Free.

If you buy a Leanpub book, you get free updates for as long as the author updates the book! Many authors use Leanpub to publish their books in-progress, while they are writing them. All readers get free updates, regardless of when they bought the book or how much they paid (including free).

Most Leanpub books are available in PDF (for computers) and EPUB (for phones, tablets and Kindle). The formats that a book includes are shown at the top right corner of this page.

Finally, Leanpub books don't have any DRM copy-protection nonsense, so you can easily read them on any supported device.

Learn more about Leanpub's ebook formats and where to read them

Write and Publish on Leanpub

You can use Leanpub to easily write, publish and sell in-progress and completed ebooks and online courses!

Leanpub is a powerful platform for serious authors, combining a simple, elegant writing and publishing workflow with a store focused on selling in-progress ebooks.

Leanpub is a magical typewriter for authors: just write in plain text, and to publish your ebook, just click a button. (Or, if you are producing your ebook your own way, you can even upload your own PDF and/or EPUB files and then publish with one click!) It really is that easy.

Learn more about writing on Leanpub