Welcome!

SDN Journal Authors: Pat Romanski, Patrick Hubbard, Elizabeth White, Sven Olav Lund, Liz McMillan

Related Topics: @BigDataExpo, Java IoT, Linux Containers, Open Source Cloud, SDN Journal

@BigDataExpo: Article

Planet @Hadoop By @ABridgwater | @BigDataExpo [#BigData]

Hadoop has been called a foundational technology, rather than ‘just’ a database by some commentators

What Kinds of People Live on Planet Hadoop?

Labor market analytics firm Wanted Analytics recently assessed the market for technology professionals and found that demand for people with proficient levels of Hadoop expertise had skyrocketed by around 33% since last year - it is true, Hadoop is hard technology to master and the labor market is not exactly flooded with an over-abundance of skilled practitioners.

Hadoop has been called a foundational technology, rather than ‘just' a database by some commentators - this almost pushes it towards being an ‘environment' rather than it being a single software product... and this all goes towards making Hadoop even harder to master, many will agree.

Interestingly, we can essentially define Hadoop as a free Java-based ‘programming framework' that supports the processing of large data sets in a distributed computing environment with robust data processing and storage power.

Hadoop confusion
Is Hadoop a foundational technology, a data environment or programming framework? It is all of these things of course, but this may be where some of the challenge (and, specifically, the skills challenge) lies.

What kinds of people live on Planet Hadoop? Well, quite apart from being able to breath a noxious mixture of data, storage and Java-flavored oxygen, Hadoop native beings (we'll call them Hadoopers) are deeply innovative.

Hadoopers are the kind of people that are driven by an inherent desire to push the ‘analytics envelope' so far forward that it becomes an operational business process that in itself drives innovation - and therefore greater profits.

These are the sort of people who are prepared to help ‘industrialize the analytical process' and make it part of the way a firm operates from first principles. As Adrian Jones of SAS has put it, the analytics factory provides an approach that offers flexibility for analysts and a structured framework for governance.

"Your data preparation process should make it easy to quickly move between the two states. This step requires the creativity of the business combined with the process and operational efficiency of IT. Taking a factory approach to data preparation will remove this gray area from many organizations, reducing the conflict that comes through duplications and inefficiencies," writes Jones.

Hard-wired into Hadoop
As operational processes throughout the business start to get hard-wired into Hadoop, we see analytics deployed in a production state where it becomes part of the core workflow of all operations across the business - the truly analytical company starts to flourish from this point onward.

This Hadoop planet (or company, or environment, or biosphere) is populated by people with ability and aptitude of course; Hadoop is no place for dummies unfortunately... but what other traits to Hadoopers exhibit?

"Resourcefully creative with an ability to think on their feet," - would be a good way of describing these folk, i.e., they are capable of surveying the analytics landscape and knowing instinctively which datasets to combine, splice, analyze and focus on - but more than anything, they can perceive the kind of outcome that ‘might' be useful without actually knowing how it should initially be quantified and qualified.

A steely stoicism
Being able to convert the analytic process into an operational process requires an aptitude for practicality and a steely stoicism with the patience to wait and look for the right outcomes.

Really successful Hadoopers are corporate beasts with an appreciation for their company's regulatory processes and systems, i.e., there's not much point in producing great Hadoop analytics if it remains untamed analytics. Good corporate Hadoop citizens are capable of making Hadoop insights ‘universally accessible' so that every stakeholder in the organization gets the appropriate level of access.

As Fiona McNeill, Product Marketing, SAS, puts it, "The difference with analytically mature, innovative organizations, is that insights are universally accessible - whether you work in HR, finance, sales, logistics, marketing or services. The data is recognized as a corporate asset and analytical methods become intellectual property."

Nimble feet, fingers and foreheads
This discussion could go on and on. Good Hadoop pros exhibit many, many attributes, but one that is appropriate to finish with is nimbleness. Although this term ‘nimble' has become hackneyed, overused and overtired across the IT industry, we will bid to use it this one last time.

Hadoop is all about trial and error in terms of finding out what works best in relation to any particular analytics job. Nimble Hadoopers are unyielding in their ability to keep trying out analytics scenarios until they get what they need - or, perhaps more important, until they get what they want for the business.

These are some of the many elements currently populating the prosperous parts of Hadoop, whether we survive in this brave new world or not is dependent on our ability to be this kind of Hadooper - welcome to the new world.

This post is brought to you by SAS.

SAS is a leader in business analytics software and services and the largest independent vendor in the business intelligence market.

More Stories By Adrian Bridgwater

Adrian Bridgwater is a freelance journalist and corporate content creation specialist focusing on cross platform software application development as well as all related aspects software engineering, project management and technology as a whole.

Comments (0)

Share your thoughts on this story.

Add your comment
You must be signed in to add a comment. Sign-in | Register

In accordance with our Comment Policy, we encourage comments that are on topic, relevant and to-the-point. We will remove comments that include profanity, personal attacks, racial slurs, threats of violence, or other inappropriate material that violates our Terms and Conditions, and will block users who make repeated violations. We ask all readers to expect diversity of opinion and to treat one another with dignity and respect.


@CloudExpo Stories
SYS-CON Events announced today that Daiya Industry will exhibit at the Japanese Pavilion at SYS-CON's 21st International Cloud Expo®, which will take place on Oct 31 – Nov 2, 2017, at the Santa Clara Convention Center in Santa Clara, CA. Ruby Development Inc. builds new services in short period of time and provides a continuous support of those services based on Ruby on Rails. For more information, please visit https://github.com/RubyDevInc.
When it comes to cloud computing, the ability to turn massive amounts of compute cores on and off on demand sounds attractive to IT staff, who need to manage peaks and valleys in user activity. With cloud bursting, the majority of the data can stay on premises while tapping into compute from public cloud providers, reducing risk and minimizing need to move large files. In his session at 18th Cloud Expo, Scott Jeschonek, Director of Product Management at Avere Systems, discussed the IT and busine...
As businesses evolve, they need technology that is simple to help them succeed today and flexible enough to help them build for tomorrow. Chrome is fit for the workplace of the future — providing a secure, consistent user experience across a range of devices that can be used anywhere. In her session at 21st Cloud Expo, Vidya Nagarajan, a Senior Product Manager at Google, will take a look at various options as to how ChromeOS can be leveraged to interact with people on the devices, and formats th...
First generation hyperconverged solutions have taken the data center by storm, rapidly proliferating in pockets everywhere to provide further consolidation of floor space and workloads. These first generation solutions are not without challenges, however. In his session at 21st Cloud Expo, Wes Talbert, a Principal Architect and results-driven enterprise sales leader at NetApp, will discuss how the HCI solution of tomorrow will integrate with the public cloud to deliver a quality hybrid cloud e...
Is advanced scheduling in Kubernetes achievable? Yes, however, how do you properly accommodate every real-life scenario that a Kubernetes user might encounter? How do you leverage advanced scheduling techniques to shape and describe each scenario in easy-to-use rules and configurations? In his session at @DevOpsSummit at 21st Cloud Expo, Oleg Chunikhin, CTO at Kublr, will answer these questions and demonstrate techniques for implementing advanced scheduling. For example, using spot instances ...
The next XaaS is CICDaaS. Why? Because CICD saves developers a huge amount of time. CD is an especially great option for projects that require multiple and frequent contributions to be integrated. But… securing CICD best practices is an emerging, essential, yet little understood practice for DevOps teams and their Cloud Service Providers. The only way to get CICD to work in a highly secure environment takes collaboration, patience and persistence. Building CICD in the cloud requires rigorous ar...
SYS-CON Events announced today that Yuasa System will exhibit at the Japan External Trade Organization (JETRO) Pavilion at SYS-CON's 21st International Cloud Expo®, which will take place on Oct 31 – Nov 2, 2017, at the Santa Clara Convention Center in Santa Clara, CA. Yuasa System is introducing a multi-purpose endurance testing system for flexible displays, OLED devices, flexible substrates, flat cables, and films in smartphones, wearables, automobiles, and healthcare.
Companies are harnessing data in ways we once associated with science fiction. Analysts have access to a plethora of visualization and reporting tools, but considering the vast amount of data businesses collect and limitations of CPUs, end users are forced to design their structures and systems with limitations. Until now. As the cloud toolkit to analyze data has evolved, GPUs have stepped in to massively parallel SQL, visualization and machine learning.
The session is centered around the tracing of systems on cloud using technologies like ebpf. The goal is to talk about what this technology is all about and what purpose it serves. In his session at 21st Cloud Expo, Shashank Jain, Development Architect at SAP, will touch upon concepts of observability in the cloud and also some of the challenges we have. Generally most cloud-based monitoring tools capture details at a very granular level. To troubleshoot problems this might not be good enough.
Organizations do not need a Big Data strategy; they need a business strategy that incorporates Big Data. Most organizations lack a road map for using Big Data to optimize key business processes, deliver a differentiated customer experience, or uncover new business opportunities. They do not understand what’s possible with respect to integrating Big Data into the business model.
When it comes to cloud computing, the ability to turn massive amounts of compute cores on and off on demand sounds attractive to IT staff, who need to manage peaks and valleys in user activity. With cloud bursting, the majority of the data can stay on premises while tapping into compute from public cloud providers, reducing risk and minimizing need to move large files. In his session at 18th Cloud Expo, Scott Jeschonek, Director of Product Management at Avere Systems, discussed the IT and busine...
Enterprises have taken advantage of IoT to achieve important revenue and cost advantages. What is less apparent is how incumbent enterprises operating at scale have, following success with IoT, built analytic, operations management and software development capabilities – ranging from autonomous vehicles to manageable robotics installations. They have embraced these capabilities as if they were Silicon Valley startups. As a result, many firms employ new business models that place enormous impor...
SYS-CON Events announced today that Taica will exhibit at the Japan External Trade Organization (JETRO) Pavilion at SYS-CON's 21st International Cloud Expo®, which will take place on Oct 31 – Nov 2, 2017, at the Santa Clara Convention Center in Santa Clara, CA. Taica manufacturers Alpha-GEL brand silicone components and materials, which maintain outstanding performance over a wide temperature range -40C to +200C. For more information, visit http://www.taica.co.jp/english/.
SYS-CON Events announced today that Dasher Technologies will exhibit at SYS-CON's 21st International Cloud Expo®, which will take place on Oct 31 - Nov 2, 2017, at the Santa Clara Convention Center in Santa Clara, CA. Dasher Technologies, Inc. ® is a premier IT solution provider that delivers expert technical resources along with trusted account executives to architect and deliver complete IT solutions and services to help our clients execute their goals, plans and objectives. Since 1999, we'v...
Recently, REAN Cloud built a digital concierge for a North Carolina hospital that had observed that most patient call button questions were repetitive. In addition, the paper-based process used to measure patient health metrics was laborious, not in real-time and sometimes error-prone. In their session at 21st Cloud Expo, Sean Finnerty, Executive Director, Practice Lead, Health Care & Life Science at REAN Cloud, and Dr. S.P.T. Krishnan, Principal Architect at REAN Cloud, will discuss how they b...
We all know that end users experience the Internet primarily with mobile devices. From an app development perspective, we know that successfully responding to the needs of mobile customers depends on rapid DevOps – failing fast, in short, until the right solution evolves in your customers' relationship to your business. Whether you’re decomposing an SOA monolith, or developing a new application cloud natively, it’s not a question of using microservices – not doing so will be a path to eventual b...
SYS-CON Events announced today that MIRAI Inc. will exhibit at the Japan External Trade Organization (JETRO) Pavilion at SYS-CON's 21st International Cloud Expo®, which will take place on Oct 31 – Nov 2, 2017, at the Santa Clara Convention Center in Santa Clara, CA. MIRAI Inc. are IT consultants from the public sector whose mission is to solve social issues by technology and innovation and to create a meaningful future for people.
SYS-CON Events announced today that TidalScale, a leading provider of systems and services, will exhibit at SYS-CON's 21st International Cloud Expo®, which will take place on Oct 31 - Nov 2, 2017, at the Santa Clara Convention Center in Santa Clara, CA. TidalScale has been involved in shaping the computing landscape. They've designed, developed and deployed some of the most important and successful systems and services in the history of the computing industry - internet, Ethernet, operating s...
SYS-CON Events announced today that IBM has been named “Diamond Sponsor” of SYS-CON's 21st Cloud Expo, which will take place on October 31 through November 2nd 2017 at the Santa Clara Convention Center in Santa Clara, California.
Join IBM November 1 at 21st Cloud Expo at the Santa Clara Convention Center in Santa Clara, CA, and learn how IBM Watson can bring cognitive services and AI to intelligent, unmanned systems. Cognitive analysis impacts today’s systems with unparalleled ability that were previously available only to manned, back-end operations. Thanks to cloud processing, IBM Watson can bring cognitive services and AI to intelligent, unmanned systems. Imagine a robot vacuum that becomes your personal assistant tha...