| By Liz McMillan | Article Rating: |
|
| March 22, 2012 05:00 AM EDT | Reads: |
3,649 |
Talend and MapR Technologies announced on Wednesday that Talend Open Studio for Big Data, an open source big data integration solution, has been certified to use with MapR’s Hadoop Distribution, providing users of Hadoop in the enterprise with high performance and scalable integration options.
Offered under the Apache Software License, Talend Open Studio for Big Data is a powerful and versatile open source solution for big data integration that natively supports Apache Hadoop. By leveraging Apache Hadoop’s MapReduce architecture for highly distributed data processing, Talend Open Studio for Big Data generates native Hadoop code and runs data transformations directly inside Hadoop for maximum scalability. Its easy-to-use graphical development environment dramatically improves the efficiency of data integration job design.

According to a recent report from Forrester Research, “Big data is all about processing and analyzing very large amounts of structured, unstructured, or semi-structured data very quickly. Although Hadoop offers a scalable data platform to process very large amounts of data for analytics purposes, getting data into the Hadoop platform is often not easy. Most leading ETL vendors are extending their solutions to integrate with Hadoop to offload large amounts of data from databases and data warehouses.” 1
Talend Open Studio for Big Data bridges the gap between Hadoop and the rest of the information system, thanks to the broadest palette of connectors currently available on the market – both for Hadoop and many other data platforms and IT systems. Hadoop connectors included in Talend Open Studio for Big Data, and which are certified to use with MapR’s Distribution for Hadoop, include:
- Hadoop Distributed File System (HDFS), to load and extract data from Hadoop, in batch or streaming mode.
- NFS, to load, extract and stream data using MapR’s high performance standard file-base access.
- HBase, to load, extract and transform data into Hadoop’s column-oriented database.
- Pig (Pig Latin generation) and Hive (HiveQL generation), to process Hadoop data in place, leveraging the power of the MapR cluster.
- Sqoop, for building direct Hadoop-to-database links without coding.
MapR’s Distribution is also available as the EMC Greenplum MR Edition and available with Cisco UCS.
To ensure connectivity of Hadoop with the rest of the information system, Talend Open Studio for Big Data also provides an unmatched set of connectors for:
- Databases, such as Oracle, MS SQL Server, MySQL, DB2, PostgreSQL, Teradata, Vertica, EMC Greenplum, Infobright, etc.
- ERP/CRM systems, such as SAP, MS Dynamics, Salesforce.com, etc.
- SaaS, Cloud and Social platforms, including Amazon Web Services, NetSuite, Marketo, Google, Twitter, etc.
- Files – structured, poly-structured and semi-structured, flat, XML, etc.
- Mainframes & midrange systems, Cobol files, etc.
Published March 22, 2012 Reads 3,649
Copyright © 2012 SYS-CON Media, Inc. — All Rights Reserved.
Syndicated stories and blog feeds, all rights reserved by the author.
More Stories By Liz McMillan
Liz is Associate Online Editor at Ulitzer.com, where she covers emerging technologies including Cloud Computing and Virtualization, as well as mergers and acquisitions and "new-media" strategies as described under the Ulitzer Live! umbrella. You can forward your press releases by email lizmcmillan.ulitzer.com.
- Cloud People: A Who's Who of Cloud Computing
- Cloud Expo New York: Cloud Is Changing the Economics of Business
- Windows Azure IaaS Reaches General Availability
- Portable Experimenter’s Platform, Powered by Raspberry Pi
- Cloudant to Exhibit at Cloud Expo & Big Data Expo New York
- Learn How To Use Google Apps Script
- Cloud Expo New York: Basics of SSD Technology and Its Use in Cloud
- Cloud Computing Is Simplifying Things
- Session Topics: 12th Cloud Expo / Cloud Expo New York
- Cloud Expo New York: The Big Challenge of Big Data & Hadoop Integration
- Overview of the OpenStack Cloud
- The Flexible Cloud
- Cloud People: A Who's Who of Cloud Computing
- Cloud Expo New York: Cloud Is Changing the Economics of Business
- Cloud Expo New York: How to Use Google Apps Script
- Windows Azure IaaS Reaches General Availability
- Rackspace Hosting Named “Platinum Plus Sponsor” of Cloud Expo New York
- Portable Experimenter’s Platform, Powered by Raspberry Pi
- Small Cancers, Big Data, and a Life Examined
- SUSE Receives Common Criteria Security Certifications
- Cloudant to Exhibit at Cloud Expo & Big Data Expo New York
- Basho Announces Open Source Riak CS and General Availability of Riak CS Enterprise v1.3
- Learn How To Use Google Apps Script
- VMware Sets Up New Hybrid Cloud Unit
- After Ubuntu, Windows Looks Increasingly Bad, Increasingly Archaic, Increasingly Unfriendly
- SCO CEO Posts Open Letter to the Open Source Community
- Simula Labs Launches Hosted Delivery Platform To Enable Enterprise Open Source Adoption
- Where Are RIA Technologies Headed in 2008?
- Source Claims SCO Will Sue Google
- How Open Is "Open"? – Industry Luminaries Join the Debate
- Latest SCO News is Plain Weird
- SCO Claims Linux Lifted ELF
- IBM Tells SCO Court It Can't Find AIX-on-Power Code
- Developing an Application Using the Eclipse BIRT Report Engine API
- Should RIM BlackBerries Be Rented?
- Flashback: Investing in 'Professional Open Source' - Exclusive 2004 Interview with David Skok, Matrix Partners



















