Skip to content
Home/ Getting Started with Kudu: Perform Fast Analytics on Fast Data
Getting Started with Kudu: Perform Fast Analytics on Fast Data

Getting Started with Kudu: Perform Fast Analytics on Fast Data

No customer reviews yet ISBN 9781491980255

Fast data ingestion, serving, and analytics in the Hadoop ecosystem have forced developers and architects to choose solutions using the least common denominator--either fast analytics at the cost of slow data ingestion or fast data ingestion at the cost of slow analytics. There is an answer to this problem. With the Apache Kudu column-oriented data store, you can easily perform fast analytics on fast data. This practical guide shows you how.

Begun as an internal project at Cloudera, Kudu is an open source solution compatible with many data processing frameworks in the Hadoop environment. In this book, current and former solutions professionals from Cloudera provide use cases, examples, best practices, and sample code to help you get up to speed with Kudu.

  • Explore Kudu's high-level design, including how it spreads data across servers
  • Fully administer a Kudu cluster, enable security, and add or remove nodes
  • Learn Kudu's client-side APIs, including how to integrate Apache Impala, Spark, and other frameworks for data manipulation
  • Examine Kudu's schema design, including basic concepts and primitives necessary to make your project successful
  • Explore case studies for using Kudu for real-time IoT analytics, predictive modeling, and in combination with another storage engine

About the author

Product details

Pub dateAug 21, 2018
ISBN-101491980257
ISBN-139781491980255
LanguageEnglish
Last updated 2026-04-24 21:16
$49.38 $49.99 1% off
You save $0.61 · list price $49.99
In stock — ships in 24 hours with free tracking
Delivery by Wednesday, October 14, 2026
Qty
Sign in to Add to Saved list
Free delivery on orders over $35.
15-day returns. Any reason.
Secure checkout. We never store card details.