Click here to close now.

Welcome!

Apache Authors: Carmen Gonzalez, Liz McMillan, Plutora Blog, Pat Romanski, Trevor Parsons

News Feed Item

Cloudera Qualifies Data Scientists With New Certification Program

Hands-On Certification Prepares Data Scientists for Success With Real-World Data; Data Science Challenge Begins March 31

PALO ALTO, CA -- (Marketwired) -- 03/26/14 -- Cloudera, the leader in enterprise analytic data management powered by Apache Hadoop™, today announced the industry's first hands-on data science certification, called Cloudera Certified Professional: Data Scientist (CCP: DS). Comprised of a Data Science Essentials exam, a twice-annual Data Science Challenge, and several preparatory and enablement resources, Cloudera's data scientist certification program helps developers, analysts, statisticians, and engineers get experience with relevant big data tools and techniques and validate their abilities while helping prospective employers identify elite, highly skilled data scientists. The next Cloudera Data Science Challenge begins March 31, 2014.

Industry Faces Shortage of Qualified Data Scientists

Enterprises are increasingly storing massive amounts of data in Hadoop to streamline the path to actionable insights, develop advanced analytics models, and build big data tools that were previously unattainable for most organizations. As a result, the demand for data scientists is at an all-time high. Data scientists possess a rare combination of engineering capabilities, statistical skills, and subject matter expertise that is difficult to find. Job openings for data scientists far outpace the limited supply of these highly in-demand workers, and the skills gap is widening. The situation is complicated by the fact that there has historically not been a clearly established skill set or university degree that an individual could acquire to qualify as a data scientist. Companies seeking to hire their first data scientists often have little idea what credentials to look for in a candidate.

Cloudera Addresses Demand for Data Scientists through Training and Certification

As the global leader in Hadoop training and professional certification, Cloudera is addressing the widespread industry need for data scientists with its new CCP:DS certification. Designed and led by Cloudera's own elite team of data scientists, the CCP:DS program helps aspiring data scientists develop and prove out the skills they need to succeed with real-world enterprise data.

In addition to the certification exam, the program includes an optional three-day Introduction to Data Science course focused on teaching data professionals to build machine learning models and implement complex recommender systems with Hadoop as a platform using industry-standard tools like Python and Apache Mahout. Cloudera also offers a 60-question Data Science Essentials Practice Test for candidates to self-assess their exam-readiness, and a free Data Science Challenge Solution Kit consisting of a live data set, a step-by-step tutorial, and a detailed explanation of the processes required to arrive at the correct outcomes for real-world data science questions focused on classification, clustering, and collaborative filtering of web analytics.

Once candidates have passed the Data Science Essentials exam, they must then successfully complete a Cloudera Data Science Challenge, offered twice annually. By passing Cloudera's examination and live-data challenge, CCP:DS-credentialed individuals have demonstrated their ability to work with big data and build market-relevant data science models under real-world conditions at the very highest level. Cloudera Certified Professional: Data Scientist is the world's only certification that provides evidence of true experience and expertise developing a production-ready data science solution that is peer-evaluated for accuracy, scalability, and robustness.

Introducing the Data Science Challenge: Detecting Anomalies in Medicare Claims
Cloudera's second Data Science Challenge opens on March 31, 2014. Participants will have three months to complete the challenge. Designed by Cloudera's Director of Data Science, Sean Owen, the Data Science Challenge asks aspiring data scientists to detect possible errors and anomalies in Medicare claims using a massive set of anonymized healthcare data. Successful challengers will be able to answer questions, including:

  • Which medical procedures have the highest relative variance in cost?
  • Which three providers had the highest average amount claimed for the largest number of procedures?
  • Based on amount and type of procedures claimed, which three providers and regions are least like the others?
  • Identify 10,000 patients that seem most likely to need review for possible errors or anomalies. Describe some common features in these patients.

To learn more about the Data Science Challenge or to register, please visit: http://cloudera.com/content/cloudera/en/training/certification/ccp-ds/challenge/register.html

Join us for a webinar about the current Data Science Challenge on April 10: http://go.cloudera.com/LP=385

What Data Scientists Say about CCP:DS:
"The certification program that Cloudera has put together goes beyond the written test, including a challenge that is designed to assess the data scientist skills in much greater depth than could be achieved in a multiple choice questionnaire. From my perspective, this makes the exercise much more compelling, valuable, and meaningful than any other certification available today. You are actually solving problems through data analysis in a full simulation of situations data scientists face in the field."
- Luis Quintela, Samsung SDS, Cloudera Certified Professional: Data Scientist

"CCP:DS goes a long way towards removing ambiguity about who and what a data scientist is. Being associated with Cloudera earns instant respect, as well. Because the exam is based on real-world challenges and is fully vetted by some of the world's top experts, the certification does the hard work of pre-evaluating candidates against the multiple highly technical areas that would otherwise be difficult to qualify."
- David F. McCoy, confidential employer, Cloudera Certified Professional: Data Scientist

"I'm pumped to earn the CCP:DS credential! It holds true weight in the market because it replicates a real, sufficiently difficult big data scenario I would see on the job and requires a professional-level approach to solving problems. The exam captured all the relevant elements of data science and machine learning, and the challenge made the experience completely non-trivial."
- Stuart Horsman, Cloudera, Cloudera Certified Professional: Data Scientist

Learn More About Cloudera's Training and Professional Certification Programs
To learn more about Cloudera's comprehensive offering of big data training programs and professional certifications, including the new CCP: Data Scientist program, please visit:
http://university.cloudera.com.

About Cloudera
Cloudera is revolutionizing enterprise data management by offering the first unified Platform for Big Data, an enterprise data hub built on Apache Hadoop™. Cloudera offers enterprises one place to store, process and analyze all their data, empowering them to extend the value of existing investments while enabling fundamental new ways to derive value from their data. Only Cloudera offers everything needed on a journey to an enterprise data hub, including software for business critical data challenges such as storage, access, management, analysis, security and search. As the leading educator of Hadoop professionals, Cloudera has trained over 20,000 individuals worldwide. Over 900 partners and a seasoned professional services team help deliver greater time to value. Finally, only Cloudera provides proactive and predictive support to run an enterprise data hub with confidence. Leading organizations in every industry plus top public sector organizations globally run Cloudera in production.
www.cloudera.com

Connect with Cloudera
Read our Vision blog: http://vision.cloudera.com
Follow Cloudera on Twitter: http://twitter.com/cloudera
Follow Cloudera University on Twitter: http://twitter.com/ClouderaU
Visit us on Facebook: http://www.facebook.com/cloudera

Cloudera, Cloudera Platform for Big Data, Cloudera Enterprise Basic Edition, Cloudera Enterprise Flex Edition, Cloudera Enterprise Data Hub Edition and CDH are trademarks or registered trademarks of Cloudera in the United States and in jurisdictions throughout the world. All other company and product names may be trade names or trademarks of their respective owners.

Add to Digg Bookmark with del.icio.us Add to Newsvine

More Stories By Marketwired .

Copyright © 2009 Marketwired. All rights reserved. All the news releases provided by Marketwired are copyrighted. Any forms of copying other than an individual user's personal reference without express written permission is prohibited. Further distribution of these materials is strictly forbidden, including but not limited to, posting, emailing, faxing, archiving in a public database, redistributing via a computer network or in a printed form.

@ThingsExpo Stories
The cloud is now a fact of life but generating recurring revenues that are driven by solutions and services on a consumption model have been hard to implement, until now. In their session at 16th Cloud Expo, Ermanno Bonifazi, CEO & Founder of Solgenia, and Ian Khan, Global Strategic Positioning & Brand Manager at Solgenia, will discuss how a top European telco has leveraged the innovative recurring revenue generating capability of the consumption cloud to enable a unique cloud monetization model to drive results.
As organizations shift toward IT-as-a-service models, the need for managing and protecting data residing across physical, virtual, and now cloud environments grows with it. CommVault can ensure protection &E-Discovery of your data – whether in a private cloud, a Service Provider delivered public cloud, or a hybrid cloud environment – across the heterogeneous enterprise. In his session at 16th Cloud Expo, Randy De Meno, Chief Technologist - Windows Products and Microsoft Partnerships, will discuss how to cut costs, scale easily, and unleash insight with CommVault Simpana software, the only si...
Analytics is the foundation of smart data and now, with the ability to run Hadoop directly on smart storage systems like Cloudian HyperStore, enterprises will gain huge business advantages in terms of scalability, efficiency and cost savings as they move closer to realizing the potential of the Internet of Things. In his session at 16th Cloud Expo, Paul Turner, technology evangelist and CMO at Cloudian, Inc., will discuss the revolutionary notion that the storage world is transitioning from mere Big Data to smart data. He will argue that today’s hybrid cloud storage solutions, with commodity...
Cloud data governance was previously an avoided function when cloud deployments were relatively small. With the rapid adoption in public cloud – both rogue and sanctioned, it’s not uncommon to find regulated data dumped into public cloud and unprotected. This is why enterprises and cloud providers alike need to embrace a cloud data governance function and map policies, processes and technology controls accordingly. In her session at 15th Cloud Expo, Evelyn de Souza, Data Privacy and Compliance Strategy Leader at Cisco Systems, will focus on how to set up a cloud data governance program and s...
Roberto Medrano, Executive Vice President at SOA Software, had reached 30,000 page views on his home page - http://RobertoMedrano.SYS-CON.com/ - on the SYS-CON family of online magazines, which includes Cloud Computing Journal, Internet of Things Journal, Big Data Journal, and SOA World Magazine. He is a recognized executive in the information technology fields of SOA, internet security, governance, and compliance. He has extensive experience with both start-ups and large companies, having been involved at the beginning of four IT industries: EDA, Open Systems, Computer Security and now SOA.
The industrial software market has treated data with the mentality of “collect everything now, worry about how to use it later.” We now find ourselves buried in data, with the pervasive connectivity of the (Industrial) Internet of Things only piling on more numbers. There’s too much data and not enough information. In his session at @ThingsExpo, Bob Gates, Global Marketing Director, GE’s Intelligent Platforms business, to discuss how realizing the power of IoT, software developers are now focused on understanding how industrial data can create intelligence for industrial operations. Imagine ...
We certainly live in interesting technological times. And no more interesting than the current competing IoT standards for connectivity. Various standards bodies, approaches, and ecosystems are vying for mindshare and positioning for a competitive edge. It is clear that when the dust settles, we will have new protocols, evolved protocols, that will change the way we interact with devices and infrastructure. We will also have evolved web protocols, like HTTP/2, that will be changing the very core of our infrastructures. At the same time, we have old approaches made new again like micro-services...
Every innovation or invention was originally a daydream. You like to imagine a “what-if” scenario. And with all the attention being paid to the so-called Internet of Things (IoT) you don’t have to stretch the imagination too much to see how this may impact commercial and homeowners insurance. We’re beyond the point of accepting this as a leap of faith. The groundwork is laid. Now it’s just a matter of time. We can thank the inventors of smart thermostats for developing a practical business application that everyone can relate to. Gone are the salad days of smart home apps, the early chalkb...
Operational Hadoop and the Lambda Architecture for Streaming Data Apache Hadoop is emerging as a distributed platform for handling large and fast incoming streams of data. Predictive maintenance, supply chain optimization, and Internet-of-Things analysis are examples where Hadoop provides the scalable storage, processing, and analytics platform to gain meaningful insights from granular data that is typically only valuable from a large-scale, aggregate view. One architecture useful for capturing and analyzing streaming data is the Lambda Architecture, representing a model of how to analyze rea...
SYS-CON Events announced today that Vitria Technology, Inc. will exhibit at SYS-CON’s @ThingsExpo, which will take place on June 9-11, 2015, at the Javits Center in New York City, NY. Vitria will showcase the company’s new IoT Analytics Platform through live demonstrations at booth #330. Vitria’s IoT Analytics Platform, fully integrated and powered by an operational intelligence engine, enables customers to rapidly build and operationalize advanced analytics to deliver timely business outcomes for use cases across the industrial, enterprise, and consumer segments.
Today’s enterprise is being driven by disruptive competitive and human capital requirements to provide enterprise application access through not only desktops, but also mobile devices. To retrofit existing programs across all these devices using traditional programming methods is very costly and time consuming – often prohibitively so. In his session at @ThingsExpo, Jesse Shiah, CEO, President, and Co-Founder of AgilePoint Inc., discussed how you can create applications that run on all mobile devices as well as laptops and desktops using a visual drag-and-drop application – and eForms-buildi...
Containers and microservices have become topics of intense interest throughout the cloud developer and enterprise IT communities. Accordingly, attendees at the upcoming 16th Cloud Expo at the Javits Center in New York June 9-11 will find fresh new content in a new track called PaaS | Containers & Microservices Containers are not being considered for the first time by the cloud community, but a current era of re-consideration has pushed them to the top of the cloud agenda. With the launch of Docker's initial release in March of 2013, interest was revved up several notches. Then late last...
SYS-CON Events announced today that Dyn, the worldwide leader in Internet Performance, will exhibit at SYS-CON's 16th International Cloud Expo®, which will take place on June 9-11, 2015, at the Javits Center in New York City, NY. Dyn is a cloud-based Internet Performance company. Dyn helps companies monitor, control, and optimize online infrastructure for an exceptional end-user experience. Through a world-class network and unrivaled, objective intelligence into Internet conditions, Dyn ensures traffic gets delivered faster, safer, and more reliably than ever.
In their session at @ThingsExpo, Shyam Varan Nath, Principal Architect at GE, and Ibrahim Gokcen, who leads GE's advanced IoT analytics, focused on the Internet of Things / Industrial Internet and how to make it operational for business end-users. Learn about the challenges posed by machine and sensor data and how to marry it with enterprise data. They also discussed the tips and tricks to provide the Industrial Internet as an end-user consumable service using Big Data Analytics and Industrial Cloud.
The explosion of connected devices / sensors is creating an ever-expanding set of new and valuable data. In parallel the emerging capability of Big Data technologies to store, access, analyze, and react to this data is producing changes in business models under the umbrella of the Internet of Things (IoT). In particular within the Insurance industry, IoT appears positioned to enable deep changes by altering relationships between insurers, distributors, and the insured. In his session at @ThingsExpo, Michael Sick, a Senior Manager and Big Data Architect within Ernst and Young's Financial Servi...
Performance is the intersection of power, agility, control, and choice. If you value performance, and more specifically consistent performance, you need to look beyond simple virtualized compute. Many factors need to be considered to create a truly performant environment. In his General Session at 15th Cloud Expo, Harold Hannon, Sr. Software Architect at SoftLayer, discussed how to take advantage of a multitude of compute options and platform features to make cloud the cornerstone of your online presence.
Even as cloud and managed services grow increasingly central to business strategy and performance, challenges remain. The biggest sticking point for companies seeking to capitalize on the cloud is data security. Keeping data safe is an issue in any computing environment, and it has been a focus since the earliest days of the cloud revolution. Understandably so: a lot can go wrong when you allow valuable information to live outside the firewall. Recent revelations about government snooping, along with a steady stream of well-publicized data breaches, only add to the uncertainty
The explosion of connected devices / sensors is creating an ever-expanding set of new and valuable data. In parallel the emerging capability of Big Data technologies to store, access, analyze, and react to this data is producing changes in business models under the umbrella of the Internet of Things (IoT). In particular within the Insurance industry, IoT appears positioned to enable deep changes by altering relationships between insurers, distributors, and the insured. In his session at @ThingsExpo, Michael Sick, a Senior Manager and Big Data Architect within Ernst and Young's Financial Servi...
PubNub on Monday has announced that it is partnering with IBM to bring its sophisticated real-time data streaming and messaging capabilities to Bluemix, IBM’s cloud development platform. “Today’s app and connected devices require an always-on connection, but building a secure, scalable solution from the ground up is time consuming, resource intensive, and error-prone,” said Todd Greene, CEO of PubNub. “PubNub enables web, mobile and IoT developers building apps on IBM Bluemix to quickly add scalable realtime functionality with minimal effort and cost.”
Docker is an excellent platform for organizations interested in running microservices. It offers portability and consistency between development and production environments, quick provisioning times, and a simple way to isolate services. In his session at DevOps Summit at 16th Cloud Expo, Shannon Williams, co-founder of Rancher Labs, will walk through these and other benefits of using Docker to run microservices, and provide an overview of RancherOS, a minimalist distribution of Linux designed expressly to run Docker. He will also discuss Rancher, an orchestration and service discovery platf...