Hadoop Matters Blog

Big Industries' blog

Infrastructure as Code: Managing Servers in the Cloud

Managing Servers in the cloud.png

Read More →

Big Data Architectures: beyond hadoop

Big_Data_Architectures_beyond_hadoop.png

 

Read More →

Why Do I need a Data Lake?

What is a data lake?

A Data Lake is an enterprise-wide system for storing and analyzing disparate sources of data in their native formats. A Data Lake might combine sensor data, social media data, click-streams, location data, log files, and much more with traditional data from existing RDBMSes. The goal is to break the information silos in an enterprise by bringing all the data into a single place for analysis without the restrictions of schema, security, or authorization. Data Lakes are designed to store vast amounts of data, even petabytes, in local or cloud-based clusters consisting of commodity hardware.

Read More →

Emre Sevinç reviews: Cassandra: The Definitive Guide

 

Read More →

Emre Sevinç's Reviews > Designing Data-Intensive Applications

Designing_Data_Intensive_Applications.png

Read More →

Data Governance in hadoop environments

Cloudera User Group.jpg

 

                                  

 

Belgium Cloudera User Group Meetup

Big Industries, as main sponsor of the Belgium Cloudera User Group, organised on Wednesday February 8th, 2017 a Meetup in our offices at Cronos in Kontich with Data Governance in Hadoop Environments as central topic.

Read More →