Berlin, Germany

Abhishek Choudhary

Elite
@abhishek-ch

Data Engineer and Data Science Working on several Big Data Technologies like Apache Spark, Hadoop, Kafka, Imapala, Apache Beam, Hive and others

around-dataengineering. A Data Engineering & Machine Learning Knowledge Hub

1.1k

data-machinelearning-the-boring-way. Build & Learn Data Engineering,Machine Learning over Kubernetes. No Shortcut approach.

57

VectorVerse. Explore Multiple Vector Databases and chat with documents on Multiple LLM models, private LLM models

48

streamlit-healthcare-ML-App. Streamlit example showing Scikit Learn & Pyspark ML over Healthcare data ! Its simple !!

32

snowflakeGPT. A Snowflake GPT Demo using SqlAlchemy

23

SmileDetection. This is an Android Based Smile Detection Project , using OpenCV and JavaCV. It works well with android and to make it work you need to install the open cv libraries in your Android Phone.

16

Awesome_Algorithm. Collection of Interesting Algorithms

16

Kubectl-GPT. Kubernetes cli (kubectl) powered by GPT

15

mlx-video-qa. Explore the capabilities of the MLX library and leverage the genAI stack on MacOS to interact with any video.

9

dataengineering-agent. Data Engineering Agent Using Open AI Function Call

9

mimic-ai. MIMICAI - Exploring MIMIC Data Using LLM and Open-WebUI

8

cspaper-ai. Tool is to serve as an AI for Computer Science Papers, capable of referencing and extracting additional information

6

SQL_GPT. A quick dirty code to generate sql code using chatGPT

6

data-engineer-roadmap. Roadmap to becoming a data engineer in 2020

5

fbmoviesuggestion. In Facebook, user can like movies, so based on that , here its extracting those movies from each of friend's user and then it fetches the details of each of the movie using API , then this recommends the best movie to each of the user within the movies in friends' circle.

3

MachineLearning-using-R. CodesMachine Learning Source files

3

evolveML. Intention is to use different algorithms of Machine Learning in R-Programming and Python to work with various dimension and range of data. The implementation will be based on BigData framework and main point of attraction will be Spark and Hive includinh hadoop

3

hue. Let’s Big Data. Hue is an open source Web interface for analyzing data with Apache Hadoop.

2

Algorithms. Data Structures and Algorithms in Python

2

healthtracker. This is an iOS app built on top of Swift. App has an UI for monitor Health over iPhone as well as in Apple Watch

2

awesome-k8s-resources. A curated list of awesome Kubernetes tools and resources.

2

kaggle_facebook_recruiting_human_or_bot. Code for kaggle competition at https://www.kaggle.com/c/facebook-recruiting-iv-human-or-bot

2

transfusion-webui. Transfusion UI based on Open-Web-UI | Experiment

1

incubator-ignite. Mirror of Apache Ignite (Incubating)

1

spark-cs190.1x. Working of CS190.1x, Scalable Machine Learning

1

spark-docker. Spark Docker Environment for testing Purpose

1

spark. Mirror of Apache Spark

1

twitter-sentiment-analysis-tutorial-201107. Code to reproduce the simple sentiment analysis from my presentation

1

monger. Monger is an idiomatic Clojure MongoDB driver for a more civilized age: with sane defaults, batteries included, well documented, very fast

1

Flink-Examples. Apache Flink work and Examples

1

GeoSpark. A Cluster Computing System for Processing Large-Scale Spatial Data

1

spark_hungarian. Hungarian Method using Apache Spark

1

kafka-example. Simple example for reading and writing into Kafka

1

myvagrant. edX: Introduction to Big Data with Apache Spark

1

data-science-from-scratch. code for Data Science From Scratch book

1
35
Apply