Thanks to visit codestin.com
Credit goes to github.com

Skip to content
#

pyspark-api

Here are 16 public repositories matching this topic...

This repo contains implementations of PySpark for real-world use cases for batch data processing, streaming data processing sourced from Kafka, sockets, etc., spark optimizations, business specific bigdata processing scenario solutions, and machine learning use cases.

  • Updated Jul 24, 2024
  • Jupyter Notebook

Traitement distribué d’images sur AWS (EMR, EC2, S3) avec PySpark et MobileNetV2 : extraction de features, PCA Spark et pipeline Big Data scalable.

  • Updated Nov 16, 2025
  • Jupyter Notebook

Improve this page

Add a description, image, and links to the pyspark-api topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the pyspark-api topic, visit your repo's landing page and select "manage topics."

Learn more