Click here to close now.

Welcome!

Open Source Authors: Irit Gillath, Pat Romanski, Carmen Gonzalez, XebiaLabs Blog, Glenn Rossman

News Feed Item

Two Science Translational Medicine Reports: DREAM and Sage Bionetworks Tap into the Wisdom of the Crowd to Fight the Challenge of Breast Cancer Prognosis and Treatment

Two new reports issuing in Science Translational Medicine (STM) today showcase the potential of teams of scientists working together to solve increasingly complex medical problems.

The results demonstrate that better predictors of breast cancer progression than those currently available can be rapidly evolved by running open Big Data Challenges such as The Sage Bionetworks/DREAM Breast Cancer Prognosis Challenge (BCC).

In breast cancer, a key undertaking is determining those patients whose disease is most likely to progress rapidly and therefore tailor the best course of treatment for them. Currently oncologists are using gene-expression based assays such as MammaPrint and Oncotype Dx, that are based on 10 year old science, and both do better with breast cancer risk prediction than models based only on clinical data.

Dr. Stephen Friend, the Founder of Sage Bionetworks and one of the organizers of the BCC reflects, “Ten years ago, members of our research group used gene expression profiling to build one of the first breast cancer predictors. Mammaprint and Oncotype Dx were developed off of that but further improvement seems to have stalled. We wondered if running a Challenge like BCC would motivate lots of different groups to tackle this problem, some working collaboratively, and if that might be more fruitful than the current 'go it alone' single researcher approach.”

To push the envelope on all the innovations that could be incorporated into the BCC, Sage partnered with the DREAM Project, a visionary distributed systems biology group that has run 24 successful open computational challenges over the last five years.

DREAM’s founder and leader, Dr. Gustavo Stolovitzky saw the BCC as an opportunity to, “… refocus our efforts to create a collaborative research environment that fosters a complementary way of doing science, which accelerates the pace of discovery with the goal of contributing to a faster reduction of suffering due to disease. This seems to me like an ethical imperative.”

The goal of the BCC was to build a computational model that accurately predicts breast cancer survival. To do this, participants of the Challenge used genomic and clinical information from 2000 women diagnosed with breast cancer (the METABRIC data set). They accessed this data on Synapse, Sage Bionetworks’ open compute platform for data sharing and analysis: Google donated cloud-based standardized virtual machines that each participant used to train their models against the data. Individual participants and/or teams submitted their computational models to Synapse as open source code made viewable to all: their models were assessed against a hidden dataset and their scores were reported on a real-time leaderboard. The combination of immediate feedback and code-sharing allowed participants to improve their leaderboard ranking by adjusting their own models or by borrowing the code of others to forge new models.

Throughout the July-October 2012 model-training phase, a crowd of 350 players from 35 countries across the globe joined the Challenge and submitted a total of 1700 computational models for scoring. The winning model was determined by scoring the predictive accuracy of players’ models against a newly generated data set: for this, the Avon Foundation For Women funded the generation of gene expression and copy number data as well as collection of corresponding clinical information from 180 breast cancer patients. Finally, the BCC organizers recognized that the basic science community might be most energized to participate if the Challenge prize were not money but the invitation to publish an article about the winning model in a top tier journal. The editors of STM saw the unique opportunity to run their own experiment on how to structure the peer-review process for competition-based crowdsourcing studies such as the BCC. Today’s issue of STM features not only the winner’s article (the BCC Challenge prize) and a report from the BCC organizers on the Challenge’s conception, execution and insights -- STM also chose to highlight the BCC with an Editorial Summary and an iconic cover of “Rosie the Riveter,” intended to symbolize the power of women and their data to transform health.

Quipped Challenge participant Richard Savage (MRC Fellow in Biostatistics at the University of Warwick) on the prospect of winning the opportunity to publish in STM, “This is huge and a genuinely new way to do some great science. I really think the organizers are onto something with this.”

The winner turned out not to be a breast cancer doctor, or even a breast cancer researcher: the winning team (“Attractor Metagenes”) hails from Professor Dimitris Anastassiou’s laboratory at Columbia University's School of Engineering and Applied Science. Anastassiou, now a member of the Columbia Initiative in Systems Biology, funded this research from his own inventor’s research allocation of patent royalties related to his previous work on digital television, which is now used in all DVDs and TV broadcasting systems worldwide. Working with two of his Ph.D. students, they developed the winning model underpinned by so-called “attractor metagenes,” gene signatures that they had identified as behaving similarly in multiple cancer types. They refer to attractor metagenes as “bioinformatic hallmarks of cancer.” Remarks Professor Anastassiou, “We had discovered these ‘pan-cancer’ gene signatures previously, and so we hypothesized that they play important roles in cancer in general. The BCC allowed us to prove that they are indeed highly prognostic at least in breast cancer.” Indeed, the winning model’s predictive accuracy for breast cancer survival outperformed the best 60 models of a pre-competition group of expert programmers and bested current clinical standards. He is now excited with the prospect of collaborating with medical researchers to make good use of these signatures of cancer for potential use in diagnostic, prognostic and eventually therapeutic products applicable in multiple cancer types.

Based on the success of the BCC, Sage Bionetworks and DREAM announced earlier this year that they would merge to run open science computational Challenges which foster the broader collaboration of the research community and provide a meaningful impact to both discovery and clinical research. Their merger provides a collaborative framework that will bring the ideals of open science one step closer to reality.

The BCC demonstrated the wisdom of the crowd to develop predictive models but also highlighted that the value of those models is limited by the questions being posed and by the data being utilized. Even as the BCC reports in this week’s issue of STM, Sage Bionetworks and DREAM are announcing five DREAM8 Challenges at Sage’s 4th Commons Congress taking place in San Francisco and working with the Avon Foundation For Women, Susan G. Komen, the Breast Cancer Research Foundation to develop the next BCC which will start by mobilizing breast cancer patients to donate their data to drive the solving of a clinically relevant question in breast cancer with the potential to transform patient treatment.

ABOUT THE DREAM PROJECT

The Dialogue on Reverse Engineering Assessment and Methods Project (DREAM Project), founded in 2006 by Andrea Califano (Columbia University) and Gustavo Stolovitzky (IBM), was originally conceived as an initiative to advance the nascent field of network biology through the organization of Challenges on network reconstruction and pathway inference. Since the first set of network inference challenges of 2007 (DREAM2) the concept of using collaborative-competitions as a vehicle to carry on a meaningful dialogue in the computational biology community has evolved significantly. In 2012, the last DREAM7 project featured four powerful challenges of which one was on network biology and the other three dealt with three important problems in translational medicine. With the experience gathered by the launching of 24 successful challenges over the past five years, the “Challenge” concept has reached a status of legitimacy and maturity. The DREAM Challenges have brought rigor in the process of verification of computational methods, have enabled the democratization of different kinds of biological data, and have facilitated the collaboration of dozens of research teams. This success has triggered considerable interest by different government institutions and private organizations in working with DREAM to engage distributed teams to solve tough computational problems in biomedical research.

ABOUT SAGE BIONETWORKS

Sage Bionetworks is a nonprofit biomedical research organization, founded in 2009, with a vision to promote innovations in personalized medicine by enabling a community-based approach to scientific inquiries and discoveries. Sage Bionetworks strives to activate patients and to incentivize scientists, funders and researchers to work in fundamentally new ways in order to shape research, accelerate access to knowledge and transform human health. It is located on the campus of the Fred Hutchinson Cancer Research Center in Seattle, Washington and is supported through a portfolio of philanthropic donations, competitive research grants, and commercial partnerships. More information is available at www.sagebase.org.

More Stories By Business Wire

Copyright © 2009 Business Wire. All rights reserved. Republication or redistribution of Business Wire content is expressly prohibited without the prior written consent of Business Wire. Business Wire shall not be liable for any errors or delays in the content, or for any actions taken in reliance thereon.

@ThingsExpo Stories
The cloud is now a fact of life but generating recurring revenues that are driven by solutions and services on a consumption model have been hard to implement, until now. In their session at 16th Cloud Expo, Ermanno Bonifazi, CEO & Founder of Solgenia, and Ian Khan, Global Strategic Positioning & Brand Manager at Solgenia, will discuss how a top European telco has leveraged the innovative recurring revenue generating capability of the consumption cloud to enable a unique cloud monetization model to drive results.
As organizations shift toward IT-as-a-service models, the need for managing and protecting data residing across physical, virtual, and now cloud environments grows with it. CommVault can ensure protection &E-Discovery of your data – whether in a private cloud, a Service Provider delivered public cloud, or a hybrid cloud environment – across the heterogeneous enterprise. In his session at 16th Cloud Expo, Randy De Meno, Chief Technologist - Windows Products and Microsoft Partnerships, will discuss how to cut costs, scale easily, and unleash insight with CommVault Simpana software, the only si...
Analytics is the foundation of smart data and now, with the ability to run Hadoop directly on smart storage systems like Cloudian HyperStore, enterprises will gain huge business advantages in terms of scalability, efficiency and cost savings as they move closer to realizing the potential of the Internet of Things. In his session at 16th Cloud Expo, Paul Turner, technology evangelist and CMO at Cloudian, Inc., will discuss the revolutionary notion that the storage world is transitioning from mere Big Data to smart data. He will argue that today’s hybrid cloud storage solutions, with commodity...
Cloud data governance was previously an avoided function when cloud deployments were relatively small. With the rapid adoption in public cloud – both rogue and sanctioned, it’s not uncommon to find regulated data dumped into public cloud and unprotected. This is why enterprises and cloud providers alike need to embrace a cloud data governance function and map policies, processes and technology controls accordingly. In her session at 15th Cloud Expo, Evelyn de Souza, Data Privacy and Compliance Strategy Leader at Cisco Systems, will focus on how to set up a cloud data governance program and s...
Roberto Medrano, Executive Vice President at SOA Software, had reached 30,000 page views on his home page - http://RobertoMedrano.SYS-CON.com/ - on the SYS-CON family of online magazines, which includes Cloud Computing Journal, Internet of Things Journal, Big Data Journal, and SOA World Magazine. He is a recognized executive in the information technology fields of SOA, internet security, governance, and compliance. He has extensive experience with both start-ups and large companies, having been involved at the beginning of four IT industries: EDA, Open Systems, Computer Security and now SOA.
The industrial software market has treated data with the mentality of “collect everything now, worry about how to use it later.” We now find ourselves buried in data, with the pervasive connectivity of the (Industrial) Internet of Things only piling on more numbers. There’s too much data and not enough information. In his session at @ThingsExpo, Bob Gates, Global Marketing Director, GE’s Intelligent Platforms business, to discuss how realizing the power of IoT, software developers are now focused on understanding how industrial data can create intelligence for industrial operations. Imagine ...
We certainly live in interesting technological times. And no more interesting than the current competing IoT standards for connectivity. Various standards bodies, approaches, and ecosystems are vying for mindshare and positioning for a competitive edge. It is clear that when the dust settles, we will have new protocols, evolved protocols, that will change the way we interact with devices and infrastructure. We will also have evolved web protocols, like HTTP/2, that will be changing the very core of our infrastructures. At the same time, we have old approaches made new again like micro-services...
Every innovation or invention was originally a daydream. You like to imagine a “what-if” scenario. And with all the attention being paid to the so-called Internet of Things (IoT) you don’t have to stretch the imagination too much to see how this may impact commercial and homeowners insurance. We’re beyond the point of accepting this as a leap of faith. The groundwork is laid. Now it’s just a matter of time. We can thank the inventors of smart thermostats for developing a practical business application that everyone can relate to. Gone are the salad days of smart home apps, the early chalkb...
Operational Hadoop and the Lambda Architecture for Streaming Data Apache Hadoop is emerging as a distributed platform for handling large and fast incoming streams of data. Predictive maintenance, supply chain optimization, and Internet-of-Things analysis are examples where Hadoop provides the scalable storage, processing, and analytics platform to gain meaningful insights from granular data that is typically only valuable from a large-scale, aggregate view. One architecture useful for capturing and analyzing streaming data is the Lambda Architecture, representing a model of how to analyze rea...
Today’s enterprise is being driven by disruptive competitive and human capital requirements to provide enterprise application access through not only desktops, but also mobile devices. To retrofit existing programs across all these devices using traditional programming methods is very costly and time consuming – often prohibitively so. In his session at @ThingsExpo, Jesse Shiah, CEO, President, and Co-Founder of AgilePoint Inc., discussed how you can create applications that run on all mobile devices as well as laptops and desktops using a visual drag-and-drop application – and eForms-buildi...
SYS-CON Events announced today that Vitria Technology, Inc. will exhibit at SYS-CON’s @ThingsExpo, which will take place on June 9-11, 2015, at the Javits Center in New York City, NY. Vitria will showcase the company’s new IoT Analytics Platform through live demonstrations at booth #330. Vitria’s IoT Analytics Platform, fully integrated and powered by an operational intelligence engine, enables customers to rapidly build and operationalize advanced analytics to deliver timely business outcomes for use cases across the industrial, enterprise, and consumer segments.
Containers and microservices have become topics of intense interest throughout the cloud developer and enterprise IT communities. Accordingly, attendees at the upcoming 16th Cloud Expo at the Javits Center in New York June 9-11 will find fresh new content in a new track called PaaS | Containers & Microservices Containers are not being considered for the first time by the cloud community, but a current era of re-consideration has pushed them to the top of the cloud agenda. With the launch of Docker's initial release in March of 2013, interest was revved up several notches. Then late last...
SYS-CON Events announced today that Dyn, the worldwide leader in Internet Performance, will exhibit at SYS-CON's 16th International Cloud Expo®, which will take place on June 9-11, 2015, at the Javits Center in New York City, NY. Dyn is a cloud-based Internet Performance company. Dyn helps companies monitor, control, and optimize online infrastructure for an exceptional end-user experience. Through a world-class network and unrivaled, objective intelligence into Internet conditions, Dyn ensures traffic gets delivered faster, safer, and more reliably than ever.
In their session at @ThingsExpo, Shyam Varan Nath, Principal Architect at GE, and Ibrahim Gokcen, who leads GE's advanced IoT analytics, focused on the Internet of Things / Industrial Internet and how to make it operational for business end-users. Learn about the challenges posed by machine and sensor data and how to marry it with enterprise data. They also discussed the tips and tricks to provide the Industrial Internet as an end-user consumable service using Big Data Analytics and Industrial Cloud.
Performance is the intersection of power, agility, control, and choice. If you value performance, and more specifically consistent performance, you need to look beyond simple virtualized compute. Many factors need to be considered to create a truly performant environment. In his General Session at 15th Cloud Expo, Harold Hannon, Sr. Software Architect at SoftLayer, discussed how to take advantage of a multitude of compute options and platform features to make cloud the cornerstone of your online presence.
The explosion of connected devices / sensors is creating an ever-expanding set of new and valuable data. In parallel the emerging capability of Big Data technologies to store, access, analyze, and react to this data is producing changes in business models under the umbrella of the Internet of Things (IoT). In particular within the Insurance industry, IoT appears positioned to enable deep changes by altering relationships between insurers, distributors, and the insured. In his session at @ThingsExpo, Michael Sick, a Senior Manager and Big Data Architect within Ernst and Young's Financial Servi...
Even as cloud and managed services grow increasingly central to business strategy and performance, challenges remain. The biggest sticking point for companies seeking to capitalize on the cloud is data security. Keeping data safe is an issue in any computing environment, and it has been a focus since the earliest days of the cloud revolution. Understandably so: a lot can go wrong when you allow valuable information to live outside the firewall. Recent revelations about government snooping, along with a steady stream of well-publicized data breaches, only add to the uncertainty
The explosion of connected devices / sensors is creating an ever-expanding set of new and valuable data. In parallel the emerging capability of Big Data technologies to store, access, analyze, and react to this data is producing changes in business models under the umbrella of the Internet of Things (IoT). In particular within the Insurance industry, IoT appears positioned to enable deep changes by altering relationships between insurers, distributors, and the insured. In his session at @ThingsExpo, Michael Sick, a Senior Manager and Big Data Architect within Ernst and Young's Financial Servi...
PubNub on Monday has announced that it is partnering with IBM to bring its sophisticated real-time data streaming and messaging capabilities to Bluemix, IBM’s cloud development platform. “Today’s app and connected devices require an always-on connection, but building a secure, scalable solution from the ground up is time consuming, resource intensive, and error-prone,” said Todd Greene, CEO of PubNub. “PubNub enables web, mobile and IoT developers building apps on IBM Bluemix to quickly add scalable realtime functionality with minimal effort and cost.”
Docker is an excellent platform for organizations interested in running microservices. It offers portability and consistency between development and production environments, quick provisioning times, and a simple way to isolate services. In his session at DevOps Summit at 16th Cloud Expo, Shannon Williams, co-founder of Rancher Labs, will walk through these and other benefits of using Docker to run microservices, and provide an overview of RancherOS, a minimalist distribution of Linux designed expressly to run Docker. He will also discuss Rancher, an orchestration and service discovery platf...