Data management in machine learning systems

書誌事項

Data management in machine learning systems

Matthias Boehm, Arun Kumar, Jun Yang

(Synthesis lectures on data management, #57)

Morgan & Craypool, c2019

  • : pbk

大学図書館所蔵 件 / 1

この図書・雑誌をさがす

注記

Includes bibliographical references (p.127-156)

DOI:10.2200/S00895ED1V01Y201901DTM057

内容説明・目次

内容説明

Large-scale data analytics using machine learning (ML) underpins many modern data-driven applications. ML systems provide means of specifying and executing these ML workloads in an efficient and scalable manner. Data management is at the heart of many ML systems due to data-driven application characteristics, data-centric workload characteristics, and system architectures inspired by classical data management techniques. In this book, we follow this data-centric view of ML systems and aim to provide a comprehensive overview of data management in ML systems for the end-to-end data science or ML lifecycle. We review multiple interconnected lines of work: (1) ML support in database (DB) systems, (2) DB-inspired ML systems, and (3) ML lifecycle systems. Covered topics include: in-database analytics via query generation and user-defined functions, factorized and statistical-relational learning; optimizing compilers for ML workloads; execution strategies and hardware accelerators; data access methods such as compression, partitioning and indexing; resource elasticity and cloud markets; as well as systems for data preparation for ML, model selection, model management, model debugging, and model serving. Given the rapidly evolving field, we strive for a balance between an up-to-date survey of ML systems, an overview of the underlying concepts and techniques, as well as pointers to open research questions. Hence, this book might serve as a starting point for both systems researchers and developers.

目次

Preface Acknowledgments Introduction ML Through Database Queries and UDFs Multi-Table ML and Deep Systems Integration Rewrites and Optimization Execution Strategies Data Access Methods Resource Heterogeneity and Elasticity Systems for ML Lifecycle Tasks Conclusions Bibliography Authors' Biographies

「Nielsen BookData」 より

関連文献: 1件中  1-1を表示

詳細情報

  • NII書誌ID(NCID)
    BB28119725
  • ISBN
    • 9781681734965
  • 出版国コード
    us
  • タイトル言語コード
    eng
  • 本文言語コード
    eng
  • 出版地
    [San Rafael, Calif.]
  • ページ数/冊数
    xv, 157 p.
  • 大きさ
    24 cm
  • 親書誌ID
ページトップへ