Skip to content
#

spark-sql

Here are 245 public repositories matching this topic...

ApacheSpark

This repository will help you to learn about databricks concept with the help of examples. It will include all the important topics which we need in our real life experience as a data engineer. We will be using pyspark & sparksql for the development. At the end of the course we also cover few case studies.

  • Updated Sep 26, 2025
  • Python

Open-source graph query service layer for ClickHouse and Databricks, written in Rust, turning an existing database into a graph view in minutes. It also provides an embedded mode with chdb, now supporting writes.

  • Updated Oct 4, 2026
  • Python

A structured streaming was applied to the robot data from ROS-Gazebo simulation environment using Apache Spark. Data is collected in Kafka, analyzed by Apache Spark and stored in Cassandra.

  • Updated Feb 6, 2022
  • Python

This is a data processing pipeline that implements an End-to-End Real-Time Geospatial Analytics and Visualization multi-component full-stack solution, using Apache Spark Structured Streaming, Apache Kafka, MongoDB Change Streams, Node.js, React, Uber's Deck.gl and React-Vis, and using the Massachusetts Bay Transportation Authority's (MBTA) APIs …

  • Updated Dec 11, 2022
  • Python

Add this topic to your repo

To associate your repository with the spark-sql topic, visit your repo's landing page and select "manage topics."

Learn more