100% Free Free Rank Your Site Faster — Get a Quality Backlink Today
Get Started
ZhenYi.Ma
GitHub

ZhenYi.Ma

@yesdata

View GitHub Profile

About

Docker+Harbor+Kubernetes CDH+Hadoop/Spark Maintenance

Followers 2
Following 2
Public Repositories 11
GitHub Work

Projects & Repositories

Recent public projects and repositories from this profile.

Hadoop

The Apache™ Hadoop® project develops open-source software for reliable, scalable, distributed computing. The Apache Hadoop software library is a framework that allows for the distributed processing of large data sets across clusters of computers using simple programming models. It is designed to scale up from single servers to thousands of machines, each offering local computation and storage. Rather than rely on hardware to deliver high-availability, the library itself is designed to detect and handle failures at the application layer, so delivering a highly-available service on top of a cluster of computers, each of which may be prone to failures.

⭐ 0 Forks: 0
Repository View Project

Spark

Latest News Spark 2.3.3 released (Feb 15, 2019) Spark 2.2.3 released (Jan 11, 2019) Spark+AI Summit (April 23-25th, 2019, San Francisco) agenda posted (Dec 19, 2018) Spark 2.4.0 released (Nov 02, 2018)

⭐ 1 Forks: 0
Repository View Project

Python

No project description available.

⭐ 0 Forks: 0
Repository View Project

yesdata.github.io

No project description available.

⭐ 0 Forks: 0
Repository View Project

Flume

Flume is a distributed, reliable, and available service for efficiently collecting, aggregating, and moving large amounts of log data. It has a simple and flexible architecture based on streaming data flows. It is robust and fault tolerant with tunable reliability mechanisms and many failover and recovery mechanisms. It uses a simple extensible data model that allows for online analytic application.

⭐ 0 Forks: 0
Repository View Project

HBase

HBase™ is the Hadoop database, a distributed, scalable, big data store. Use Apache HBase™ when you need random, realtime read/write access to your Big Data. This project's goal is the hosting of very large tables -- billions of rows X millions of columns -- atop clusters of commodity hardware. Apache HBase is an open-source, distributed, versioned, non-relational database modeled after Google's Bigtable: A Distributed Storage System for Structured Data by Chang et al. Just as Bigtable leverages the distributed data storage provided by the Google File System, Apache HBase provides Bigtable-like capabilities on top of Hadoop and HDFS.

⭐ 0 Forks: 0
Repository View Project