Hadoop Matters Blog

Big Industries' blog

Cloudera User Group: Cloudera in the Cloud

Meetup logo.png

                                  

Read More →

Does the world need cloudera?

Cloudera_Office.png

 

Read More →

Running Cloudera on Microsoft Azure

In recent years, the cloud has offered a scalable, versatile, and efficient environment for big data workloads. In this blogpost we explore how companies are deploying Cloudera Enterprise for running production Hadoop workloads - ETL, BI, advanced analytics, Spark - on Microsoft Azure, with enterprise-class data security and governance.

Read More →

Cloudera User Group: Big Data analytics

Cloudera User Group.jpg

  

 

Belgium Cloudera User Group Meetup

Big Industries, as main sponsor of the Belgium Cloudera User Group, organised on Wednesday May 31st, 2017 a Meetup in our offices at Cronos in Kontich with Big Data Analytics as central topic.

Read More →

Apache spark market survey

top use cases for spark.jpg

 

Read More →

Why Do I need a Data Lake?

What is a data lake?

A Data Lake is an enterprise-wide system for storing and analyzing disparate sources of data in their native formats. A Data Lake might combine sensor data, social media data, click-streams, location data, log files, and much more with traditional data from existing RDBMSes. The goal is to break the information silos in an enterprise by bringing all the data into a single place for analysis without the restrictions of schema, security, or authorization. Data Lakes are designed to store vast amounts of data, even petabytes, in local or cloud-based clusters consisting of commodity hardware.

Read More →