Skip to main content

The Forrester Wave (Or: We're all the leaders)

Listen:
Forrester Research, an independent market research firm, released in February 2014 the quarterly Forrester Wave Big Data Hadoop Solutions, Q1 2014 Report [1]. The report shows this graphic, and it looks like that all major, minor and non-hadoop Vendors think they lead. It looks really funny when you follow the mainstream press news.

IBM [5] think they lead, Hortonworks [4] claim the leadership too, MapR [3] leads too, Teradata is the true leader (they say) [6]. Cloudera [2] ignores the report. The metapher is - all of the named companies are in the leader area, but nobody leads.

Forrester Wave Big Data Hadoop Solutions, Q1 2014 Report
Anyway, let us do a quick overview about the "Big Three" - Cloudera, MapR, Hortonworks.

The 3 major Hadoop firms (Horton, MapR, Cloudera) are nearly in the same position. All distributions have the sweet piece, which lets the customer decide which one fits most. And that is the most important point - the customer wins. Not the marketing noise.

Cloudera [2] depends on Apache Hadoop, has Cloudera Manager, a strong, sophisticated and great tool to manage an entire hadoop cluster, including add, relocate and remove services from a node to another. In addition to the Open Source version of Hadoop they offer Closed Source Applications on top, like Cloudera Manager Enterprise, Cloudera Navigator (Data Lineage), BDR, Snapshotting, Data Replication. But these additional services aren't OpenSource.

MapR [3] is the most convenient guy here - the press release on their website is clear, no big noise. The message: Choose what is the best for your business. Makes the company a bit friendly. MapR has 3 different solutions - M3, the free-to-use edition, M5 - the Enterprise Edition with NFS Support, Snapshotting, independent code support and M7, the Enterprise Database Edition, optimized for Low Latency and High Throughput. MapR Editions aren't Open Source, and the management console is not as feature-rich as Cloudera Manager. Additionally, the company created their own HDFS-like file system (MapR-FS), mostly written in C(++).

Hortonworks [4] is the youngest player in the market. Originally Horton comes from Yahoo and is a spin-off from the core developers on Apache Hadoop MapReduce, Apache Hadoop HDFS and Apache Hadoop Yarn. HDP, the Hortonworks Edition of Apache Hadoop, is the only 100% Open Source distribution in the market. The managing tool, Apache Ambari (incubating) is also not so feature-rich as Cloudera Manager, but it's Open Source and works well. Furthermore, Horton sells only Apache Projects in their distribution, for Data Governance Falcon, and for Security Purposes Knox.

All of  these three players have a strong support department as well as service delivery (Solution Architect), Pre- and Post Sales and a significant amount of customers.

In my eyes, I see only one true leader. Apache Hadoop. All of those "BigData" companies rely on a great idea, originally developed at Google and rebuilt by the Apache Open Source Community. This is what true leadership means - evolve and divide.

[1] http://www.forrester.com/pimages/rws/reprints/document/112461/oid/1-PBE69P
[2] http://www.cloudera.com
[3] http://www.mapr.com/forrester-wave-hadoop-distribution-comparison-and-benchmark-report
[4] http://info.hortonworks.com/ForresterWave_Hadoop.html
[5] http://www.ibmbigdatahub.com/whitepaper/forrester-wave-big-data-hadoop-solutions-q1-2014
[6] http://www.teradata.de/News-Releases/2014/Teradata-is-a-Leader-in-Big-Data-Hadoop-Solutions-in-2014/?LangType=1031 

Comments

Popular posts from this blog

Why Is Customer Obsession Disappearing?

 It's wild that even with all the cool tech we've got these days, like AI solving complex equations and doing business across time zones in a flash, so many companies are still struggling with the basics: taking care of their customers.The drama around Coinbase's customer support is a prime example of even tech giants messing up. And it's not just Coinbase — it's a big-picture issue for the whole industry. At some point, the idea of "customer obsession" got replaced with "customer automation," and now we're seeing the problems that came with it. "Cases" What Not to Do Coinbase, as main example, has long been synonymous with making cryptocurrency accessible. Whether you’re a first-time buyer or a seasoned trader, their platform was once the gold standard for user experience. But lately, their customer support practices have been making headlines for all the wrong reasons: Coinbase - Stuck in the Loop:  Users have reported being caugh...

MySQL Scaling in 2024

When your MySQL database reaches its performance limits, vertical scaling through hardware upgrades provides a temporary solution. Long-term growth, though, requires a more comprehensive approach. This involves optimizing the database strategically and integrating complementary technologies. Caching The implementation of a caching layer, such as Memcached or Redis , can result in a notable reduction in the load and an increase ni performance at MySQL. In-memory stores cache data that is accessed frequently, enabling near-instantaneous responses and freeing the database for other tasks. For applications with heavy read traffic on relatively static data (e.g. product catalogues, user profiles), caching represents a low-effort, high-impact solution. Consider a online shop product catalogue with thousands of items. With each visit to the website, the application queries the database in order to retrieve product details. By using caching, the retrieved details can be stored in Memcached (a...

What the Heck is Superposition and Entanglement?

If you’ve ever heard the words superposition or entanglement thrown around in conversations about quantum physics, you may have nodded politely while your brain quietly filed them away in the "too confusing to deal with" folder.  These aren't just theoretical quirks; they're the foundation of mind-bending tech like Google's latest quantum chip, the Willow with its 105 qubits. Superposition challenges our understanding of reality, suggesting that particles don't have definite states until observed. This principle is crucial in quantum technologies, enabling phenomena like quantum computing and quantum cryptography. What's in for us? Short, nothing at the moment. 105 qubits sounds awesome, but it would neither crack encryption nor enhance AI in the next few years. There are some use cases for Willow, like drug (protein) discovery or solving certain mathematical problems when they aren't too complicated. Right now, Google managed to turn physical qubits ...