Welcome!

Agile Computing Authors: Pat Romanski, Yeshim Deniz, Liz McMillan, Elizabeth White, Zakia Bouachraoui

Related Topics: @CloudExpo

@CloudExpo: Article

Big Data – The State of Affairs

Big Data is here to stay, but do we have the tools to efficiently process it?

Many products are available as open source or proprietary products that can handle Big Data. Which one is best fit for this task?

Today's classic RDBMSs and tools are able to quickly load the data, process it and present results in an easy to understand format.  You can use SQL or programmatic interface to process the data randomly or in batch; RDBMS's keep data safe, protected against hardware and software failures.

Standards tools and products are not able to cope with Big Data requirement, which is not dissimilar to  what is involved in processing today's regular data sets, just on a much bigger scale. Mainstream companies like telcos, financials, web companies as well as government are reaching the limit of what  can be efficiently processed by classic RDBMS techhnologies.

When it comes to picking a proper platform and tools to handle your Big Data there are a couple of possible choices:

  • Oracle Exadata - it doesn't fit economical mandate; Exadata's weak link and bottleneck is its reliance on classic Oracle RDBMS
  • NoSQL databases -  too immature, they offer no SQL or similar random access query language ( you are presently forced to write  programs to access your data ); often achieve scale-out by not implementing all elements of ACID, CAP
  • Hadoop/MapReduce and related open source ecosystem ( Pig, Hive, HBase ) -  useful for cheap data storage on commodity hardware and batch processing; they offer no efficient, non-programmatic random access
  • proprietary MPP databases running on commodity hardware ( Vertica, Aster Data, Greenplum )  - very fast and can provide random, SQL  access to big data; their management features and general feature sets are immature
  • proprietary MPP databases running on specialized hardware ( Teradata ) - fairly expensive ( don't run on commodity hardware )
  • new platforms that will or are trying to emulate Google Percolator, Dremel  ( latest Google technologies dealing with big data ACID compliant transactions and reporting ), similarly to how Hadoop originated from  Google GFS and MapReduce.

We would say that there is no single, generic product or platform available today that can handle this task. Depending on your needs you have to deploy  and combinne quite a few of technologies to bring you closer to achieving end-to-end efficient, comprehensive processing of Big Data. You will quite likely have to custom build solutions that will fit your particular needs as off-the-shelf solutions are still immature, incomplete or not available.

Big Data is an area of growth and innovation, so current picture is bound to change as new products and technologies appear, bringing us closer to the ultimate goal of routine, efficient processing of Big Data.

More Stories By Ranko Mosic

Ranko Mosic, BScEng, is specializing in Big Data/Data Architecture consulting services ( database/data architecture, machine learning ). His clients are in finance, retail, telecommunications industries. Ranko is welcoming inquiries about his availability for consulting engagements and can be reached at 408-757-0053 or [email protected]

Comments (0)

Share your thoughts on this story.

Add your comment
You must be signed in to add a comment. Sign-in | Register

In accordance with our Comment Policy, we encourage comments that are on topic, relevant and to-the-point. We will remove comments that include profanity, personal attacks, racial slurs, threats of violence, or other inappropriate material that violates our Terms and Conditions, and will block users who make repeated violations. We ask all readers to expect diversity of opinion and to treat one another with dignity and respect.


IoT & Smart Cities Stories
SYS-CON Events announced today that Silicon India has been named “Media Sponsor” of SYS-CON's 21st International Cloud Expo, which will take place on Oct 31 – Nov 2, 2017, at the Santa Clara Convention Center in Santa Clara, CA. Published in Silicon Valley, Silicon India magazine is the premiere platform for CIOs to discuss their innovative enterprise solutions and allows IT vendors to learn about new solutions that can help grow their business.
We are seeing a major migration of enterprises applications to the cloud. As cloud and business use of real time applications accelerate, legacy networks are no longer able to architecturally support cloud adoption and deliver the performance and security required by highly distributed enterprises. These outdated solutions have become more costly and complicated to implement, install, manage, and maintain.SD-WAN offers unlimited capabilities for accessing the benefits of the cloud and Internet. ...
Founded in 2000, Chetu Inc. is a global provider of customized software development solutions and IT staff augmentation services for software technology providers. By providing clients with unparalleled niche technology expertise and industry experience, Chetu has become the premiere long-term, back-end software development partner for start-ups, SMBs, and Fortune 500 companies. Chetu is headquartered in Plantation, Florida, with thirteen offices throughout the U.S. and abroad.
SYS-CON Events announced today that CrowdReviews.com has been named “Media Sponsor” of SYS-CON's 22nd International Cloud Expo, which will take place on June 5–7, 2018, at the Javits Center in New York City, NY. CrowdReviews.com is a transparent online platform for determining which products and services are the best based on the opinion of the crowd. The crowd consists of Internet users that have experienced products and services first-hand and have an interest in letting other potential buye...
Business professionals no longer wonder if they'll migrate to the cloud; it's now a matter of when. The cloud environment has proved to be a major force in transitioning to an agile business model that enables quick decisions and fast implementation that solidify customer relationships. And when the cloud is combined with the power of cognitive computing, it drives innovation and transformation that achieves astounding competitive advantage.
DXWorldEXPO LLC announced today that "IoT Now" was named media sponsor of CloudEXPO | DXWorldEXPO 2018 New York, which will take place on November 11-13, 2018 in New York City, NY. IoT Now explores the evolving opportunities and challenges facing CSPs, and it passes on some lessons learned from those who have taken the first steps in next-gen IoT services.
Cloud-enabled transformation has evolved from cost saving measure to business innovation strategy -- one that combines the cloud with cognitive capabilities to drive market disruption. Learn how you can achieve the insight and agility you need to gain a competitive advantage. Industry-acclaimed CTO and cloud expert, Shankar Kalyana presents. Only the most exceptional IBMers are appointed with the rare distinction of IBM Fellow, the highest technical honor in the company. Shankar has also receive...
DXWorldEXPO LLC announced today that ICOHOLDER named "Media Sponsor" of Miami Blockchain Event by FinTechEXPO. ICOHOLDER gives detailed information and help the community to invest in the trusty projects. Miami Blockchain Event by FinTechEXPO has opened its Call for Papers. The two-day event will present 20 top Blockchain experts. All speaking inquiries which covers the following information can be submitted by email to [email protected] Miami Blockchain Event by FinTechEXPOalso offers sp...
DXWordEXPO New York 2018, colocated with CloudEXPO New York 2018 will be held November 11-13, 2018, in New York City and will bring together Cloud Computing, FinTech and Blockchain, Digital Transformation, Big Data, Internet of Things, DevOps, AI, Machine Learning and WebRTC to one location.
@DevOpsSummit at Cloud Expo, taking place November 12-13 in New York City, NY, is co-located with 22nd international CloudEXPO | first international DXWorldEXPO and will feature technical sessions from a rock star conference faculty and the leading industry players in the world. The widespread success of cloud computing is driving the DevOps revolution in enterprise IT. Now as never before, development teams must communicate and collaborate in a dynamic, 24/7/365 environment. There is no time t...