Visual, interactive queries against big databases
-
Updated
Oct 6, 2026 - Java
Visual, interactive queries against big databases
A Spring Boot web app that buys and sells cryptocurrencies from API data sources. Its quick trading and other features allow users to leverage computer power to outperform the market.
Programs conducted at for "Hadoop Administration" subject, of PGCHPCSA CDAC-ACTS, Pune by Tushar B. Kute in April 2026.
Scalable big data pipeline on GCP (Dataproc/HDFS) analyzing 3GB+ of IMDb & YouTube data using MapReduce & Trino.
《DNA元基催化与肽计算》 在进化计算中, 软件函数文件进行 DNA 语义元基索引编码的 PDE 新陈代谢优化方式, 是一种有效的进化方式.
A suite of benchmark applications for distributed data stream processing systems
Ingest Twitter tweets and process them through Kafka to Spark then HDFS
💾 Welcome to the Big Data Analytics Repository! 📚✨ Immerse yourself in a carefully curated reservoir of knowledge on Big Data Analytics. 🌐💡 Explore the intricacies of deriving insights from vast datasets and navigating powerful analytics tools. 🚀🔍
A Hadoop MapReduce project analyzing the Consumer Complaints dataset with five queries to extract insights like complaints by product, state, company, tags, and timely responses.
Workflow management system for the automated and distributed analysis of large-scale experimental data.
Programs conducted at Army Institute of Technology, Pune in training on Big Data Analytics during September 2024.
Big Data Analysis of NYC Fire Incident data to analyze casual relationship between fires, govt. inspections, socio-ecnomic factors and enviroment. Used Hadoop MapReduce for data pre-processing, Trino for complex queries and Tableau for visualizations and interactive dashboards
The current repository contains all the code developed during the Big Data processing and Analytics laboratories. Data are processed and analyzed using Hadoop and Spark
Implementing parallel processing techniques for efficient handling of Big Data through practical activities.
Easy Machine Learning is a general-purpose dataflow-based system for easing the process of applying machine learning algorithms to real world tasks.
A lossy counting algorithm implemented to determine the top trending hashtags using the Twitter API to get a continuous stream of tweets.
This repo explains the implementation of Map-Reduce Algorithm on the AirBnb data to understand the consumer satisfaction region and country wise. This is the effective use of parallel distributed computing to resolve the big data problems
Eskimo is a state of the art Big Data Infrastructure and Management Web Console to build, manage and operate Big Data 2.0 Analytics clusters on Kubernetes. This is the git repository of Eskimo Community Edition.
Reservoir Sampling for Group-By Queries in Flink Platform. Answering effectively Single Aggregate.
Big Data Pipeline | Querying Data from Hive Table Phase
To associate your repository with the big-data-analytics topic, visit your repo's landing page and select "manage topics."