sparkmllib
Here are 50 public repositories matching this topic...
Spark Java_Examples for all modules including GraphX
-
Updated
Dec 8, 2017 - Java
Recommendation engine in Java. Based on an ALS algorithm (Apache Spark). Train a new model after N seconds.
-
Updated
Apr 21, 2023 - Java
Distributed Search and Recommendation with SpringBoot/ElasticSearch/Spark
-
Updated
Jun 1, 2021 - JavaScript
A collection of “cookbook-style” scripts for simplifying data engineering and machine learning in Apache Spark.
-
Updated
Oct 27, 2021
Data Driven Sentiment Insight into Twitter(X) Trends | Kafka | Spark | Spark MLlib | Docker
-
Updated
Jun 10, 2024 - Java
EverAnalyzer is my thesis in the Department of Digital Systems of the University of Piraeus. EverAnalyzer is a platform for collecting, preprocessing, processing and analyzing Big Data from the Twitter platform.
-
Updated
Sep 2, 2022 - HTML
Utilized SparkML and Scikit-Learn train several machine learning models for distinguishing fraudulent and legitimate transactions. The machine learning models are then utilized to make predictions on Kafka-generated real-time data streams. Built an interface for displaying these predictions in real-time using the Streamlit framework.
-
Updated
Dec 14, 2021 - Jupyter Notebook
Big Data Project - SSML - Spark Streaming for Machine Learning
-
Updated
Dec 31, 2021 - Python
Big Data Analytics Project using Apache Spark for Predicting Severity of Car Accidents in the USA
-
Updated
May 15, 2020 - Jupyter Notebook
-
Updated
Aug 27, 2021 - Jupyter Notebook
Work in-progress NBA Game Predictor using Spark
-
Updated
Oct 10, 2023 - Jupyter Notebook
We generate potential customer leads for businesses on yelp using big data and machine learning
-
Updated
Sep 15, 2017 - Java
This is a repository i have created to put up some of the knowledge i have gained around Big Data Technologies especially Spark, GraphX etc.
-
Updated
Apr 20, 2019
-
Updated
Dec 22, 2025 - Jupyter Notebook
• Developed a Recommender System for restaurants by performing analysis on data preprocessed from Yelp Dataset. • Used Altering Least Squares method with Matrix Factorization and Neighborhood Model to train and build the Recommender System. • Tested the Recommender System with multiple rounds of Cross Validation technique and 16% prediction erro…
-
Updated
May 22, 2019
Introduction to Apache Spark.
-
Updated
Apr 12, 2022 - Jupyter Notebook
Add this topic to your repo
To associate your repository with the sparkmllib topic, visit your repo's landing page and select "manage topics."