Apache Spark Data Source V2 (Wenchen Fan & Gengliang Wang)
Databricks
0:00 / 0:00
Apache Spark Data Source V2 (Wenchen Fan & Gengliang Wang)
2 429 просмотров · 7 лет назад
Databricks
166 тыс. подписчиков
2 429 просмотров · 7 лет назад
Wenchen Fan, a software engineer at Databricks, and Gengliang Wang, a software engineer in Databrick, discuss how as a general computing engine, Spark can process data from various data management/storage systems, including HDFS, Hive, Cassandra and Kafka. For flexibility and high throughput, Spark defines the Data Source API, which is an abstraction of the storage layer.
Learn more here: https://databricks.com/session/apache...
Article you might like: https://databricks.com/session/theory...
About: Databricks provides a unified data analytics platform, powered by Apache Spark™, that accelerates innovation by unifying data science, engineering and business.
Read more here: https://databricks.com/product/unifie...
Connect with us:
Website: https://databricks.com
Facebook: / databricksinc
Twitter: / databricks
LinkedIn: / databricks
Instagram: / databricksinc Databricks is proud to announce that Gartner has named us a Leader in both the 2021 Magic Quadrant for Cloud Database Management Systems and the 2021 Magic Quadrant for Data Science and Machine Learning Platforms. Download the reports here. https://databricks.com/databricks-nam...