100% Free Free Rank Your Site Faster — Get a Quality Backlink Today
Get Started
Philippe ROSSIGNOL
GitHub

Philippe ROSSIGNOL

@enahwe

View GitHub Profile

About

I'm a Big Data consultant mainly motivated around Spark and Hadoop technologies

Location Niort
Followers 5
Following 0
Public Repositories 4
Expertise

Skills & Technologies

Shell Java Python
GitHub Work

Projects & Repositories

Recent public projects and repositories from this profile.

Csv2Hive

Shell

Csv2Hive is an useful CSV schema finder for the Big Data. It discovers automatically schemas in big CSV files, generates the 'CREATE TABLE' statements and creates Hive tables. You don't need to writes any schemas at all. Csv2Hive is a really fast solution for integrating the whole CSV files into your DataLake.

⭐ 27 Forks: 10

KafkaGust

Java

KafkaGust is a flexible and useful tool to quickly test high data volumes with Apache Kafka. Apache Kafka that is used in the world of Big Data (e.g : with Storm, Cassandra, Hadoop, etc.), is a fast and scalable publish-subscribe messaging that can handle durably hundreds of megabytes of reads and writes per second from thousands of clients.

⭐ 7 Forks: 3

Spark-DIL-FTP

A Spark Application for FTP Data Ingestion

⭐ 0 Forks: 0
Repository View Project

Spark-DIL

Python

A Spark Library for Data Integration

⭐ 0 Forks: 0