Welcome!

Agile Computing Authors: Elizabeth White, Yeshim Deniz, Liz McMillan, Pat Romanski, Andy Thurai

Related Topics: Java IoT, Microsoft Cloud, Perl, Python

Java IoT: Article

Big Data Kills 30-Year-Old Market

Applications need to go to “Big Data,” not the other way around

Data Services Journal

If you’ve got simply scads of data – and why wouldn’t you? – it’s doubling every 18 months – and are shuttling it to an application for analysis, you’re doing it wrong.

That’s so…so, well, 1980.

According to Aster Data, applications need to go to “Big Data,” not the other way around.

And to do that the company’s got a massively parallel data-application server that can embed applications inside a massively scalable MPP data warehouse and analyze petabytes of data – or terabytes, if that’s all you’ve got – ultra-fast.

Apps are automatically parallelized for scale; users can take their existing Java, C, C++, C#, .NET, Perl and Python applications, MapReduce-enable them and push them down into the data.

The widgetry runs on a cluster of commodity boxes. Figure five servers to start although parallelized applications can utilize terabytes of memory and thousands of CPU cores.

This is not the data warehouses, DBMSes and data analytics solutions of the last three decades that have separated data from applications, a technique Aster says results in massive data movement, latency and restricted analysis.

Traditional systems weren’t built to process billions of rows of data in seconds or handle chi-chi stuff like real-time fraud detection, customer behavior modeling, merchandising optimization, affinity marketing, trending and simulations, trading surveillance and customer calling patterns.

They were built for data sampling, an inexact science. They simply fail in today’s big data, analytics-intensive environments, Aster says.

The company’s Aster Data 4.0 brings data and applications together in one system, fully parallelizing both, to deliver ultra-fast analysis on massive data scales. And it’s got customers like comScore, Full Tilt Poker, Telefonica I+D, SAS and MySpace, with the big clutch of data of all, saying it’s right.

Aster’s Massively Parallel Data-Application Server 4.0, based on research done at Stanford University before commercialization started a couple of years ago, lets companies embed application logic in Aster’s MPP database, which includes MapReduce. It was Aster that brought MapReduce to SQL, a trick it’s now building on.

In Aster’s system data management lives independent of the application processing but – and this is important – the data and applications execute as first-class citizens, with their own respective data and application management services.

The Data-Application Server is responsible for managing and coordinating the cluster’s activities and resource sharing. It acts as a host for the application processing and the data managed inside the cluster.

As a data host, it manages incremental scaling, fault tolerance and heterogeneous hardware for application processing and it manages workloads via Aster’s new Dynamic Workload Management (WLM) capability.

Aster says WLM, described as the first dynamic workload management capability available on a MPP system to run on commodity hardware, can support hundreds of concurrent mixed workloads. It manages data storage, transactional correctness, online backups and information lifecycles (ILM).

The separation of data management and application processing is supposed to provide maximum application portability so a wide range of applications can be pushed down into the system.

Aster says this data analysis architecture distinguishes its solution from lightweight implementations of MapReduce, including what some vendors refer to as ‘In-Database MapReduce.

Richard Zwicky, president of Enquisite, the company that provides search optimization software and solutions, says that with Aster Data, response times for large queries has dropped from five minutes to five-10 seconds, and queries that previously weren’t possible now can be executed in 20-30 seconds.

Aster Data is backed by Sequoia Capital, Jafco Ventures, IVP and Cambrian Ventures, as well as Google’s first investor David Cheriton, Ron Conway and Rajeev Motwani.

More Stories By Maureen O'Gara

Maureen O'Gara the most read technology reporter for the past 20 years, is the Cloud Computing and Virtualization News Desk editor of SYS-CON Media. She is the publisher of famous "Billygrams" and the editor-in-chief of "Client/Server News" for more than a decade. One of the most respected technology reporters in the business, Maureen can be reached by email at maureen(at)sys-con.com or paperboy(at)g2news.com, and by phone at 516 759-7025. Twitter: @MaureenOGara

Comments (1) View Comments

Share your thoughts on this story.

Add your comment
You must be signed in to add a comment. Sign-in | Register

In accordance with our Comment Policy, we encourage comments that are on topic, relevant and to-the-point. We will remove comments that include profanity, personal attacks, racial slurs, threats of violence, or other inappropriate material that violates our Terms and Conditions, and will block users who make repeated violations. We ask all readers to expect diversity of opinion and to treat one another with dignity and respect.


Most Recent Comments
MarlenaFernandezBerkowitz 11/09/09 12:49:00 PM EST

Interesting post from Enquisite CEO Mark Hoffman (former
CEO and founder of Sybase) on how they're using Aster
Data to meet their pretty demanding scalability, always-
on needs...

IoT & Smart Cities Stories
DXWorldEXPO LLC announced today that Ed Featherston has been named the "Tech Chair" of "FinTechEXPO - New York Blockchain Event" of CloudEXPO's 10-Year Anniversary Event which will take place on November 12-13, 2018 in New York City. CloudEXPO | DXWorldEXPO New York will present keynotes, general sessions, and more than 20 blockchain sessions by leading FinTech experts.
Apps and devices shouldn't stop working when there's limited or no network connectivity. Learn how to bring data stored in a cloud database to the edge of the network (and back again) whenever an Internet connection is available. In his session at 17th Cloud Expo, Ben Perlmutter, a Sales Engineer with IBM Cloudant, demonstrated techniques for replicating cloud databases with devices in order to build offline-first mobile or Internet of Things (IoT) apps that can provide a better, faster user e...
Bill Schmarzo, author of "Big Data: Understanding How Data Powers Big Business" and "Big Data MBA: Driving Business Strategies with Data Science" is responsible for guiding the technology strategy within Hitachi Vantara for IoT and Analytics. Bill brings a balanced business-technology approach that focuses on business outcomes to drive data, analytics and technology decisions that underpin an organization's digital transformation strategy.
Charles Araujo is an industry analyst, internationally recognized authority on the Digital Enterprise and author of The Quantum Age of IT: Why Everything You Know About IT is About to Change. As Principal Analyst with Intellyx, he writes, speaks and advises organizations on how to navigate through this time of disruption. He is also the founder of The Institute for Digital Transformation and a sought after keynote speaker. He has been a regular contributor to both InformationWeek and CIO Insight...
Rodrigo Coutinho is part of OutSystems' founders' team and currently the Head of Product Design. He provides a cross-functional role where he supports Product Management in defining the positioning and direction of the Agile Platform, while at the same time promoting model-based development and new techniques to deliver applications in the cloud.
Andrew Keys is Co-Founder of ConsenSys Enterprise. He comes to ConsenSys Enterprise with capital markets, technology and entrepreneurial experience. Previously, he worked for UBS investment bank in equities analysis. Later, he was responsible for the creation and distribution of life settlement products to hedge funds and investment banks. After, he co-founded a revenue cycle management company where he learned about Bitcoin and eventually Ethereal. Andrew's role at ConsenSys Enterprise is a mul...
In his session at 21st Cloud Expo, Raju Shreewastava, founder of Big Data Trunk, provided a fun and simple way to introduce Machine Leaning to anyone and everyone. He solved a machine learning problem and demonstrated an easy way to be able to do machine learning without even coding. Raju Shreewastava is the founder of Big Data Trunk (www.BigDataTrunk.com), a Big Data Training and consulting firm with offices in the United States. He previously led the data warehouse/business intelligence and Bi...
Cell networks have the advantage of long-range communications, reaching an estimated 90% of the world. But cell networks such as 2G, 3G and LTE consume lots of power and were designed for connecting people. They are not optimized for low- or battery-powered devices or for IoT applications with infrequently transmitted data. Cell IoT modules that support narrow-band IoT and 4G cell networks will enable cell connectivity, device management, and app enablement for low-power wide-area network IoT. B...
The Internet of Things will challenge the status quo of how IT and development organizations operate. Or will it? Certainly the fog layer of IoT requires special insights about data ontology, security and transactional integrity. But the developmental challenges are the same: People, Process and Platform and how we integrate our thinking to solve complicated problems. In his session at 19th Cloud Expo, Craig Sproule, CEO of Metavine, demonstrated how to move beyond today's coding paradigm and sh...
What are the new priorities for the connected business? First: businesses need to think differently about the types of connections they will need to make – these span well beyond the traditional app to app into more modern forms of integration including SaaS integrations, mobile integrations, APIs, device integration and Big Data integration. It’s important these are unified together vs. doing them all piecemeal. Second, these types of connections need to be simple to design, adapt and configure...