Welcome!

Cloud Expo Authors: Maureen O'Gara, Jeremy Geelan, Elizabeth White, Pat Romanski, Greg Schulz

Related Topics: Cloud Expo, Java, SOA & WOA, Virtualization, GovIT

Cloud Expo: Blog Feed Post

MaaS – The Solution to Design, Map, Integrate and Publish Open Data

Data models can be shared, off-line tested and verified to define data designing requirements, data topology, performance, place

Open Data is data that can be freely used, reused and redistributed by anyone – subject only, at the most, to the requirement for attributes and sharealikes (Open Software Service Definition – OSSD). As a consequence, Open Data should create value and might have a positive impact in many different areas such as government (tax money expenditure), health (medical research, hospital acceptance by pathology), quality of life (air breathed in our city, pollution) or might influence public decisions like investments, public economy and expenditure. We are talking about services, so open data are services needed to connect the community with the public bodies. However, the required open data should be part of a design and then integrated, mapped, updated and published in a form, which is easy to use. MaaS is the Open Data driver and enables Open Data portability into the Cloud.

Introduction
Data models used as a service mainly provide the following topics:

  • Implementing and sharing data structure models;
  • Verifying data model properties according to private and public cloud requirements;
  • Designing and testing new query types. Specific query classes need to support heterogeneous data;
  • Designing of the data storage model. The model should enable query processing directly against databases to ensure privacy and secure changes from data updates and review;
  • Modeling data to predict usage “early”;
  • Portability, a central property when data is shared among fields of application;
  • Sharing, redistribution and participation of data among datasets and applications.

As a consequence, the data should be available as a whole and at a reasonable fee, preferably by finding, navigating and downloading over the Cloud. It should also be available in a usable and changeable form. This means modeling Open Data and then using the models to map location and usage, configuration, integration and changes along the Open Data lifecycle.

What is MaaS
Data models can be shared, off-line tested and verified to define data designing requirements, data topology, performance, placement and deployment. This means models themselves can be supplied as a service to allow providers to verify how and where data has to be designed to meet the Cloud service’s requisites: this is MaaS. As a consequence by using MaaS, Open Data designers can verify “on-premise” how and why datasets meet Open Data requirements. With this approach, Open Data models can be tuned on real usage and then mapped “on-premise” to the public body’s service. Further, MaaS inherits all the defined service’s properties and so the data model can be reused, shared and classified for new Open Data design and publication.

Open Data implementation is MaaS (Model as a Service) driven
Open Data is completely supported by data modeling and then MaaS completely supports Open Data. MaaS should be the first practice, helping to tune analysis and Open Data design. Furthermore, data models govern design, deployment, storage, changes, resources allocation, hence MaaS supports:

  • Applying Best Practice for Open Data design;
  • Classifying Open Data field of application;
  • Designing Open Data taxonomy and integration;
  • Guiding Open Data implementation;
  • Documenting data maturity and evolution by applying DaaS lifecycle.

Accordingly, Maas provides “on-premise” properties supporting Open Data design and publication:

  1. AnalysisWhat data are you planning to make open? When working with MaaS, a data model is used to perform data analysis. This means the Open Data designer might return to this step to correct, update and improve the incoming analysis: he always works on an “on-premise” data model. Analysis performed by model helps in identifying data integration and interoperability. The latter assists in choosing what data has to be published and in defining open datasets;
  2. DesignDuring the analysis step, the design is carried out too. The design can be changed and traced along the Open Data lifecycle. Remember that with MaaS the model is a service, and the data opened offers the designed service;
  3. Data securityData security becomes the key property to rule data access and navigation. MaaS plays a crucial role in data security: in fact, the models contain all the infrastructure properties and include information to classify accesses, classes of users, perimeters and risk mitigation assets. Models are the central way to enable data protection within the Open Data device;
  4. Participation - Because the goal is “everyone must be able to use Open Data”, participation is comprehensive of people and groups without any discrimination or restriction. Models contain data access rules and accreditations (open licensing).
  5. Mapping – The MaaS mapping property is important because many people can obtain the data after long navigation and several “bridges” connecting different fields of applications. Looking at this aspect, MaaS helps the Open Data designer to define the best initial “route” between transformation and aggregation linking different areas. Then continually engaging citizens, developers, sector’s expert, managers … helps in modifying the model to better update and scale Open Data contents: the easier it is for outsiders to discover data, the faster new and useful Open Data services will be built.
  6. OntologyDefining metadata vocabulary for describing ontologies. Starting from standard naming definition, data models provide grouping and reorganizing vocabulary for further metadata re-use, integration, maintenance, mapping and versioning;
  7. Portability – Models contain all the properties belonging to data in order that MaaS can enable Open Data service’s portability to the Cloud. The model is portable by definition and it can be generated to different database and infrastructures;
  8. Availability – The DaaS lifecycle assures structure validation in terms of MaaS accessibility;
  9. Reuse and distribution – Open Data can include merging with additional datasets belonging to other fields of application (for example, medical research vs. air pollution). Open Data built by MaaS has this advantage. Merging open datasets means merging models by comparing and synchronizing, old and new versions, if needed;
  10. Change Management and History – Data models are organized in libraries to preserve Open Data changes and history. Changes are traced and maintained to restore, if necessary, model and/or datasets;
  11. Redesign – Redesigning Open Data, means redesigning the model it belongs to: the  model drives the history of the changes;
  12. Fast BI – Publishing Open Data is an action strictly related to the BI process. Redesigning and publishing Open Data are two automated steps starting from the design of the data model and from its successive updates.

Conclusion
MaaS is the emerging solution for Open Data implementation. Open Data is public and private accessible data, designed to connect the social community with the public bodies. This data should be made available without restriction although it is placed under security and open licensing. In addition, Open Data is always up-to-date and transformation and aggregation have to be simple and time saving for inesperienced users. To achieve these goals, the Open Data service has to be model driven designed and providing data integration, interoperability, mapping, portability, availability, security, distribution, all properties assured by applying MaaS.

References
[1] N. Piscopo - ERwin® in the Cloud: How Data Modeling Supports Database as a Service (DaaS) Implementations
[2] N. Piscopo - CA ERwin® Data Modeler’s Role in the Relational Cloud
[3] N. Piscopo - DaaS Contract templates: main constraints and examples, in press
[4] D. Burbank, S. Hoberman - Data Modeling Made Simple with CA ERwin® Data Modeler r8
[7] N. Piscopo – Best Practices for Moving to the Cloud using Data Models in theDaaS Life Cycle
[8] N. Piscopo – Using CA ERwin® Data Modeler and Microsoft SQL Azure to Move Data to the Cloud within the DaaS Life Cycle
[9] The Open Software Service Definition (OSSD) at opendefinition.org

Read the original blog entry...

More Stories By Cloud Ventures

The Cloud Ventures Network is an expert community of leading Cloud pioneers. Follow our best practice blogs at http://CloudBestPractices.net

Cloud Expo Breaking News
"Since Cloud Expo is running the week of June 10, we thought it'd be a great idea to schedule our Meetup this week. That way, if you have colleagues, friends, or family in town that week for the Expo, you can invite them to join you!" With those words, the OpenStack New York Meetup Group's organizer's launched a landing page this week where anyone interested can register for the June 12 evening event.
In an ideal developer/systems administrator’s world, most applications would deploy seamlessly to multiple platforms and scale elastically with minimal effort bringing the unprecedented agility of the cloud within immediate reach of developer teams and IT organizations. OpenStack, a RackSpace and NASA initiative, is now managed by an independent foundation and is supported by multiple vendors. It defines APIs for compute, storage, networking, services, monitoring, and additional infrastructure...
In his session at 12th Cloud Expo | Cloud Expo New York [June 10-13, 2013], Intel's Chris Black will review the background of Apache Hadoop, its application, and methods to accelerate data system clusters with Intel SSD technology. The session will overview the genius of Hadoop and provide an overview of the ecosystem landscape. Cloud Expo/Big Data Expo delegates will learn how the Hadoop framework and SSD technology augment cloud data systems ranging from analytics to on-line transaction pro...
Cloud computing is transforming the way businesses think about and leverage technology. As a result, the general understanding of cloud computing has come a long way in a short time. However, there are still many misconceptions about what cloud computing is and what it can do for businesses that adopt this game-changing computing model. In his General Session at the 12th International Cloud Expo, Gene Eun, Senior Director, Oracle Cloud at Oracle, will discuss and dispel some of the common myth...
SYS-CON Events announced today that Wowrack will exhibit at SYS-CON's 12th International Cloud Expo, which will take place on June 10–13, 2013, at the Javits Center in New York City, New York. Wowrack’s core expertise lies in high-availability Private and Public Cloud IaaS Hosting Solutions. Wowrack provides a true Hybrid service – where business release all IT management and hardware provisioning – taking the data center and server system administrative headaches off our customer’s shoulders. ...
SYS-CON Events announced today that nfina Technologies, a provider of highly reliable cloud server products, will exhibit at SYS-CON's 12th International Cloud Expo, which will take place on June 10–13, 2013, at the Javits Center in New York City, New York. nfina Technologies develops, manufactures, and markets highly reliable cloud server products, designed to solve the most demanding data center requirements in mission-critical cloud applications. Nfina’s staff has decades of experience in co...
SYS-CON Events announced today that OpenStack will exhibit at SYS-CON's 12th International Cloud Expo, which will take place on June 10–13, 2013, at the Javits Center in New York City, New York. OpenStack software controls large pools of compute, storage, and networking resources throughout a datacenter, all managed by a dashboard that gives administrators control while empowering their users to provision resources through a web interface. OpenStack powers some of the most widely-used SaaS app...
“Cloud has everything to do with what has happened with Big Data,” explained Jason Deck, Director of Strategic Alliances at Logicworks, in this exclusive Q&A with Cloud Expo Conference Chair Jeremy Geelan. “Big Data doesn’t exist in its easily accessible way without cloud. From reduced startup costs, to cheap storage, to fast processing, to adequate security, to the easy incorporation of third-party analytics tools, cloud made Big Data accessible to customers of all sizes, with all different bud...
As enterprises deploy private IaaS clouds into production they are reevaluating their future application delivery models. SUSE and WSO2 believe that private PaaS will leverage the automation and scalability of Private IaaS solutions, such as OpenStack-based SUSE Cloud, to deliver the secure, standardized development environments that will make migrating to an agile, serviceoriented delivery model possible. In their session at the 12th International Cloud Expo, Chris Haddad, VP of Technology Ev...
Organizations across the world are increasingly starting to see the benefits of moving more and more services to the cloud. The focus on the cost-saving potential of cloud is rapidly shifting to completely transforming the business with cloud. As organizations are investing enormous sums on technology they are starting to realize that in order to maximize the return on investment and accelerate the business transformation process the first area of focus should be people. By ensuring the organiza...