Перейти к содержимому

Apache Spark Data Source V2 (Wenchen Fan & Gengliang Wang)

Databricks

0:00 / 0:00

Apache Spark Data Source V2 (Wenchen Fan & Gengliang Wang)

2 429 просмотров · 7 лет назад
Databricks
166 тыс. подписчиков
2 429 просмотров · 7 лет назад
Wenchen Fan, a software engineer at Databricks, and Gengliang Wang, a software engineer in Databrick, discuss how as a general computing engine, Spark can process data from various data management/storage systems, including HDFS, Hive, Cassandra and Kafka. For flexibility and high throughput, Spark defines the Data Source API, which is an abstraction of the storage layer. Learn more here: https://databricks.com/session/apache... Article you might like: https://databricks.com/session/theory... About: Databricks provides a unified data analytics platform, powered by Apache Spark™, that accelerates innovation by unifying data science, engineering and business. Read more here: https://databricks.com/product/unifie... Connect with us: Website: https://databricks.com Facebook:   / databricksinc   Twitter:   / databricks   LinkedIn:   / databricks   Instagram:   / databricksinc   Databricks is proud to announce that Gartner has named us a Leader in both the 2021 Magic Quadrant for Cloud Database Management Systems and the 2021 Magic Quadrant for Data Science and Machine Learning Platforms. Download the reports here. https://databricks.com/databricks-nam...