Welcome!

Virtualization Authors: Carmen Gonzalez, Elizabeth White, Victoria Livschitz, Pat Romanski, Lori MacVittie

Related Topics: Websphere, Cloud Expo

Websphere: Blog Post

Databases in the Cloud

Cloud computing is such a fascinating topic because ...

Cloud computing is such a fascinating topic because, to anybody who can get enough distance from the marketing hype, it really represents a significant departure from the way the software and the hardware industry have been operating in the last decades. In fact, cloud computing is to the traditional IT industry what the Internet has been to the music and film industry: an unwelcome and very threatening development dictating completely different business models.

The basic technical premises behind cloud computing are known since many years. Already in the 80's, as the number of computers around constantly increased, it did not take long until people realized many of them were idle for most of time. The result was the first cluster management tools capable of moving simple jobs between machines. From there, the advances in networking during the 90's brought Grid Computing and then the large computer farms behind the Web brought what we now call Cloud Computing. On the software side, a similar development took place over the years going from monolithic designs to multi-tier architectures and from tightly coupled systems to service oriented architectures. Add virtualization, and all the basic tools for building computing clouds are in place.

Whether Software as a Service (SaaS) or Platform as a Service (PaaS), there are many examples out there of services that work well; make economic sense to all parties involved; and are destined to grow both in the number of users as well as in the functionality they provide. This, however, does not seem to be the case for databases. At least not for databases containing important data.

This is an interesting observation since, technically, there is nothing that prevents databases from residing in the cloud. To understand the complex relation between databases and the cloud, one needs to understand the complex chain of problems that need to be solved before a database with important data resides in the cloud. These problems are: 

- legal aspects of where the data resides

- long term custody warranties

- trust in the cloud

If the data residing in a database is of any real value to anybody except a small group of individuals, it is likely that there are many regulations imposing a wide variety of constraints on where the data can be, who can look at it, and whether it can be moved anywhere. For instance, in federal countries, local governments often have legislation imposing that the data must be stored within the region. At larger scale, it seems unlikely that a country would agree to have government data stored in a different country. In the private sector, many software development outsourcing efforts have failed because of the difficulty to provide realistic data for testing without giving any confidential information away. And if the database stays where it is, it is unlikely that the software stack built on top of it will move to the cloud.

Assuming there is a cloud in the vicinity that fulfills all the locality requirements, the next hurdle is the legal custodian warranties imposed on data. Important and relevant data must be by law available and searchable for long periods of time, often many decades. Clouds cannot provide such guarantees today and, in this matter, the IT industry has never dealt with such time horizons before.

Finally, even there is a cloud in the proper place that guarantees that it will stay there for the nest 50 years, the question that remains is whether it can be trusted to do so. What happens to the data if the cloud simply disappears? Replication makes the location problem even more difficult and it certainly does not help to reduce the cost of the cloud. It also does not solve the problem of a company simply shutting down the service. Without very strong, enforceable guarantees -as it happens in other branches of industry that are critical to the economy - there will be not enough trust to move databases with important data into a commercial cloud.

Does this imply that we will never see databases in the cloud? Not at all. However, the clouds were important databases will reside might be different from the commercial ones that are attracting so much attention these days.

First, the clouds where databases may live very comfortably will be private clouds. Governments, for instance, are likely to own (or contract) such clouds to offer cloud services to the public sector. Second, community clouds linking the private clouds of partner companies are also likely to be common since they spread the costs among several participants while still giving access to more resources that anyone of them directly owns. Being a federation of private clouds, they are easier to protect, organize under well defined contractual agreements, and to tailor to the particular application by using, e.g., application aware networks. Third, public clouds will be used not necessarily for storing the data but for scalability and processing purposes in all those cases where parts of the data can be safely brought into the open. For instance, a company can keep the confidential data within the private cloud but place databases with copies of the publicly available data on a public cloud. By keeping the master copy of the data, the company takes advantage of the public cloud for scalability but can make sure all regulations are followed in house using conventional solutions.

The challenges to bring databases into the cloud are both technical and regulatory. Until the regulatory problems are solved -and that may take a long time- the key to putting databases into the cloud will be to have infrastructures that give users flexibility and complete control over the databases and the data inside, regardless of the type of cloud used. If users can easily create copies of their databases or part of their databases and place those into the cloud; can guarantee that they have within their premises a consistent copy of the data at all times; and can take advantage of the cloud to reduce the costs of provisioning, scaling out, and adding functionality to their data management systems, then databases will move to the cloud. If all these chores are not provided through automatic tools, the overhead, costs, and risks involved will be too high to justify moving enterprise class databases to the cloud.

More Stories By Maximilian Ahrens

Ahrens is an expert and frequent speaker on international conferences for service oriented architecture and virtualization. Before co-founding Zimory, he served as a project manager and research scientist at the innovation development entity of Deutsche Telekom Laboratories. Responsible for infrastructure and enterprise IT projects spanning multiple divisions of the Deutsche Telekom group -- Ahrens is an expert on enterprise IT and business processes. Before Deutsche Telekom, he led several business process reengineering projects for major German companies. Ahrens received his degree in computer science and business administration from Technische Universität Berlin.

@ThingsExpo Stories
Cultural, regulatory, environmental, political and economic (CREPE) conditions over the past decade are creating cross-industry solution spaces that require processes and technologies from both the Internet of Things (IoT), and Data Management and Analytics (DMA). These solution spaces are evolving into Sensor Analytics Ecosystems (SAE) that represent significant new opportunities for organizations of all types. Public Utilities throughout the world, providing electricity, natural gas and water, are pursuing SmartGrid initiatives that represent one of the more mature examples of SAE. We have s...
The security devil is always in the details of the attack: the ones you've endured, the ones you prepare yourself to fend off, and the ones that, you fear, will catch you completely unaware and defenseless. The Internet of Things (IoT) is nothing if not an endless proliferation of details. It's the vision of a world in which continuous Internet connectivity and addressability is embedded into a growing range of human artifacts, into the natural world, and even into our smartphones, appliances, and physical persons. In the IoT vision, every new "thing" - sensor, actuator, data source, data con...
The Internet of Things is tied together with a thin strand that is known as time. Coincidentally, at the core of nearly all data analytics is a timestamp. When working with time series data there are a few core principles that everyone should consider, especially across datasets where time is the common boundary. In his session at Internet of @ThingsExpo, Jim Scott, Director of Enterprise Strategy & Architecture at MapR Technologies, discussed single-value, geo-spatial, and log time series data. By focusing on enterprise applications and the data center, he will use OpenTSDB as an example t...
How do APIs and IoT relate? The answer is not as simple as merely adding an API on top of a dumb device, but rather about understanding the architectural patterns for implementing an IoT fabric. There are typically two or three trends: Exposing the device to a management framework Exposing that management framework to a business centric logic Exposing that business layer and data to end users. This last trend is the IoT stack, which involves a new shift in the separation of what stuff happens, where data lives and where the interface lies. For instance, it's a mix of architectural styles ...
The 3rd International Internet of @ThingsExpo, co-located with the 16th International Cloud Expo - to be held June 9-11, 2015, at the Javits Center in New York City, NY - announces that its Call for Papers is now open. The Internet of Things (IoT) is the biggest idea since the creation of the Worldwide Web more than 20 years ago.
An entirely new security model is needed for the Internet of Things, or is it? Can we save some old and tested controls for this new and different environment? In his session at @ThingsExpo, New York's at the Javits Center, Davi Ottenheimer, EMC Senior Director of Trust, reviewed hands-on lessons with IoT devices and reveal a new risk balance you might not expect. Davi Ottenheimer, EMC Senior Director of Trust, has more than nineteen years' experience managing global security operations and assessments, including a decade of leading incident response and digital forensics. He is co-author of t...
The Internet of Things will greatly expand the opportunities for data collection and new business models driven off of that data. In her session at @ThingsExpo, Esmeralda Swartz, CMO of MetraTech, discussed how for this to be effective you not only need to have infrastructure and operational models capable of utilizing this new phenomenon, but increasingly service providers will need to convince a skeptical public to participate. Get ready to show them the money!
The Internet of Things will put IT to its ultimate test by creating infinite new opportunities to digitize products and services, generate and analyze new data to improve customer satisfaction, and discover new ways to gain a competitive advantage across nearly every industry. In order to help corporate business units to capitalize on the rapidly evolving IoT opportunities, IT must stand up to a new set of challenges. In his session at @ThingsExpo, Jeff Kaplan, Managing Director of THINKstrategies, will examine why IT must finally fulfill its role in support of its SBUs or face a new round of...
One of the biggest challenges when developing connected devices is identifying user value and delivering it through successful user experiences. In his session at Internet of @ThingsExpo, Mike Kuniavsky, Principal Scientist, Innovation Services at PARC, described an IoT-specific approach to user experience design that combines approaches from interaction design, industrial design and service design to create experiences that go beyond simple connected gadgets to create lasting, multi-device experiences grounded in people's real needs and desires.
Enthusiasm for the Internet of Things has reached an all-time high. In 2013 alone, venture capitalists spent more than $1 billion dollars investing in the IoT space. With "smart" appliances and devices, IoT covers wearable smart devices, cloud services to hardware companies. Nest, a Google company, detects temperatures inside homes and automatically adjusts it by tracking its user's habit. These technologies are quickly developing and with it come challenges such as bridging infrastructure gaps, abiding by privacy concerns and making the concept a reality. These challenges can't be addressed w...
The Domain Name Service (DNS) is one of the most important components in networking infrastructure, enabling users and services to access applications by translating URLs (names) into IP addresses (numbers). Because every icon and URL and all embedded content on a website requires a DNS lookup loading complex sites necessitates hundreds of DNS queries. In addition, as more internet-enabled ‘Things' get connected, people will rely on DNS to name and find their fridges, toasters and toilets. According to a recent IDG Research Services Survey this rate of traffic will only grow. What's driving t...
Scott Jenson leads a project called The Physical Web within the Chrome team at Google. Project members are working to take the scalability and openness of the web and use it to talk to the exponentially exploding range of smart devices. Nearly every company today working on the IoT comes up with the same basic solution: use my server and you'll be fine. But if we really believe there will be trillions of these devices, that just can't scale. We need a system that is open a scalable and by using the URL as a basic building block, we open this up and get the same resilience that the web enjoys.
Connected devices and the Internet of Things are getting significant momentum in 2014. In his session at Internet of @ThingsExpo, Jim Hunter, Chief Scientist & Technology Evangelist at Greenwave Systems, examined three key elements that together will drive mass adoption of the IoT before the end of 2015. The first element is the recent advent of robust open source protocols (like AllJoyn and WebRTC) that facilitate M2M communication. The second is broad availability of flexible, cost-effective storage designed to handle the massive surge in back-end data in a world where timely analytics is e...
We are reaching the end of the beginning with WebRTC, and real systems using this technology have begun to appear. One challenge that faces every WebRTC deployment (in some form or another) is identity management. For example, if you have an existing service – possibly built on a variety of different PaaS/SaaS offerings – and you want to add real-time communications you are faced with a challenge relating to user management, authentication, authorization, and validation. Service providers will want to use their existing identities, but these will have credentials already that are (hopefully) i...
"Matrix is an ambitious open standard and implementation that's set up to break down the fragmentation problems that exist in IP messaging and VoIP communication," explained John Woolf, Technical Evangelist at Matrix, in this SYS-CON.tv interview at @ThingsExpo, held Nov 4–6, 2014, at the Santa Clara Convention Center in Santa Clara, CA.
P2P RTC will impact the landscape of communications, shifting from traditional telephony style communications models to OTT (Over-The-Top) cloud assisted & PaaS (Platform as a Service) communication services. The P2P shift will impact many areas of our lives, from mobile communication, human interactive web services, RTC and telephony infrastructure, user federation, security and privacy implications, business costs, and scalability. In his session at @ThingsExpo, Robin Raymond, Chief Architect at Hookflash, will walk through the shifting landscape of traditional telephone and voice services ...
Explosive growth in connected devices. Enormous amounts of data for collection and analysis. Critical use of data for split-second decision making and actionable information. All three are factors in making the Internet of Things a reality. Yet, any one factor would have an IT organization pondering its infrastructure strategy. How should your organization enhance its IT framework to enable an Internet of Things implementation? In his session at Internet of @ThingsExpo, James Kirkland, Chief Architect for the Internet of Things and Intelligent Systems at Red Hat, described how to revolutioniz...
Bit6 today issued a challenge to the technology community implementing Web Real Time Communication (WebRTC). To leap beyond WebRTC’s significant limitations and fully leverage its underlying value to accelerate innovation, application developers need to consider the entire communications ecosystem.
The definition of IoT is not new, in fact it’s been around for over a decade. What has changed is the public's awareness that the technology we use on a daily basis has caught up on the vision of an always on, always connected world. If you look into the details of what comprises the IoT, you’ll see that it includes everything from cloud computing, Big Data analytics, “Things,” Web communication, applications, network, storage, etc. It is essentially including everything connected online from hardware to software, or as we like to say, it’s an Internet of many different things. The difference ...
Cloud Expo 2014 TV commercials will feature @ThingsExpo, which was launched in June, 2014 at New York City's Javits Center as the largest 'Internet of Things' event in the world.