Welcome!

Cognitive Computing Authors: Pat Romanski, Yeshim Deniz, Elizabeth White, Zakia Bouachraoui, Liz McMillan

Blog Feed Post

Seven Thoughts on Hadoop’s Seventh Birthday

By

Editor’s note: It has only been seven years since Hadoop’s first release, and think of the amazing things it has already empowered us to accomplish. Doug Cutting (Hadoop’s founder, Apache Software Foundation chair and Cloudera’s chief architect) published seven thoughts on Hadoop’s seventh birthday at the Cloudera blog, and we sought his permission to repost those here. 

Seven Thoughts on Hadoop’s Seventh Birthday:

  1. Open-source accelerates adoption.If Hadoop had been created as proprietary software it would not have spread as rapidly. We’ve seen incredible growth in the use of Hadoop. Partly that’s because it’s useful. But many would have been cautious to make a vendor-controlled platform part of their infrastructure, useful or not.
  2.  Apache builds collaborative communities.The Hadoop ecosystem has hundreds of developers working for tens of organizations. Competitors productively collaborate on a daily basis, improving the software we all share. The Apache Software Foundation gives us the methodology that enables this. (Thanks, Apache!)
  3. The timing is right.Folks flock to Hadoop not just because it is open-source and works, but also because it fills a need. Moore’s law provides us with a bounty of affordable hardware. This has led to computing devices spreading through our world. Cars and tractors have computers. Phones, and cash registers and more have become computers. Data flows through each of these. Hadoop gives us tools to save and analyze more of this data, improving our understanding of the world.
  4. Random names are good names.When we started out, I had no idea what Hadoop would become, so I proposed a name for it that didn’t have any connotation. The project has grown, giving that name meaning. Not everyone may pronounce the word “Hadoop” the same, but we all know what it is. A whimsical name also helps remind us to have fun.
  5. People love a story.People love to hear me tell the story of my son’s toy elephant. They love to see that toy. The story has become the mythological prehistory of Hadoop. I guess every movement needs its origin story!
  6. Hadoop is a phase transition in computing.Hadoop’s developers did not invent distributed computing, nor is Hadoop its most advanced form, but Hadoop has brought distributed computing to the mainstream. Hadoop gets thousands of unreliable computers to work together reliably. As the number of computers grows, we no longer think about them individually but instead as parts of a whole. This is a fundamentally different way of using computers that’s rapidly becoming commonplace.
  7. The sky’s the limit.Hadoop was originally created to help build search engines. It’s still used for that, but its uses have grown far beyond that. Its core features have advanced tremendously, but even more dramatic is the range of systems being built on top of Hadoop. From machine learning to real-time queries, Hadoop is becoming a great platform for nearly any task folks imagine involving large amounts of data. The trends that gave rise to Hadoop continue, and Hadoop is evolving and growing to meet new challenges. We are still in the early days of this revolution.

Here’s to the next seven years!

 

DougBWSquare1 150x150 Seven Thoughts on Hadoop’s Seventh Birthday

Doug (@cutting) is the creator of numerous successful open source projects, including Lucene, Nutch and Hadoop. Doug joined Cloudera in 2009 from Yahoo!, where he was a key member of the team that built and deployed a production Hadoop storage and analysis cluster for mission-critical business analytics. Doug holds a Bachelor’s degree from Stanford University and sits on the Board (and is currently chairman) of the Apache Software Foundation.

 

 Seven Thoughts on Hadoop’s Seventh Birthday

Read the original blog entry...

More Stories By Bob Gourley

Bob Gourley writes on enterprise IT. He is a founder of Crucial Point and publisher of CTOvision.com

IoT & Smart Cities Stories
With 10 simultaneous tracks, keynotes, general sessions and targeted breakout classes, @CloudEXPO and DXWorldEXPO are two of the most important technology events of the year. Since its launch over eight years ago, @CloudEXPO and DXWorldEXPO have presented a rock star faculty as well as showcased hundreds of sponsors and exhibitors! In this blog post, we provide 7 tips on how, as part of our world-class faculty, you can deliver one of the most popular sessions at our events. But before reading...
If a machine can invent, does this mean the end of the patent system as we know it? The patent system, both in the US and Europe, allows companies to protect their inventions and helps foster innovation. However, Artificial Intelligence (AI) could be set to disrupt the patent system as we know it. This talk will examine how AI may change the patent landscape in the years to come. Furthermore, ways in which companies can best protect their AI related inventions will be examined from both a US and...
Poor data quality and analytics drive down business value. In fact, Gartner estimated that the average financial impact of poor data quality on organizations is $9.7 million per year. But bad data is much more than a cost center. By eroding trust in information, analytics and the business decisions based on these, it is a serious impediment to digital transformation.
Digital Transformation: Preparing Cloud & IoT Security for the Age of Artificial Intelligence. As automation and artificial intelligence (AI) power solution development and delivery, many businesses need to build backend cloud capabilities. Well-poised organizations, marketing smart devices with AI and BlockChain capabilities prepare to refine compliance and regulatory capabilities in 2018. Volumes of health, financial, technical and privacy data, along with tightening compliance requirements by...
DXWorldEXPO LLC, the producer of the world's most influential technology conferences and trade shows has announced the 22nd International CloudEXPO | DXWorldEXPO "Early Bird Registration" is now open. Register for Full Conference "Gold Pass" ▸ Here (Expo Hall ▸ Here)
@DevOpsSummit at Cloud Expo, taking place November 12-13 in New York City, NY, is co-located with 22nd international CloudEXPO | first international DXWorldEXPO and will feature technical sessions from a rock star conference faculty and the leading industry players in the world. The widespread success of cloud computing is driving the DevOps revolution in enterprise IT. Now as never before, development teams must communicate and collaborate in a dynamic, 24/7/365 environment. There is no time t...
CloudEXPO New York 2018, colocated with DXWorldEXPO New York 2018 will be held November 11-13, 2018, in New York City and will bring together Cloud Computing, FinTech and Blockchain, Digital Transformation, Big Data, Internet of Things, DevOps, AI, Machine Learning and WebRTC to one location.
The best way to leverage your Cloud Expo presence as a sponsor and exhibitor is to plan your news announcements around our events. The press covering Cloud Expo and @ThingsExpo will have access to these releases and will amplify your news announcements. More than two dozen Cloud companies either set deals at our shows or have announced their mergers and acquisitions at Cloud Expo. Product announcements during our show provide your company with the most reach through our targeted audiences.
Machine Learning helps make complex systems more efficient. By applying advanced Machine Learning techniques such as Cognitive Fingerprinting, wind project operators can utilize these tools to learn from collected data, detect regular patterns, and optimize their own operations. In his session at 18th Cloud Expo, Stuart Gillen, Director of Business Development at SparkCognition, discussed how research has demonstrated the value of Machine Learning in delivering next generation analytics to impr...
The challenges of aggregating data from consumer-oriented devices, such as wearable technologies and smart thermostats, are fairly well-understood. However, there are a new set of challenges for IoT devices that generate megabytes or gigabytes of data per second. Certainly, the infrastructure will have to change, as those volumes of data will likely overwhelm the available bandwidth for aggregating the data into a central repository. Ochandarena discusses a whole new way to think about your next...