Inside the Briefcase

How to Transform Your Website into a Lead Generating Machine

How to Transform Your Website into a Lead Generating Machine

Responsive customer service has become of special importance, as...

Ironclad SaaS Security for Cloud-Forward Enterprises

Ironclad SaaS Security for Cloud-Forward Enterprises

The 2015 Anthem data breach was the result of...

The Key Benefits of Using Social Media for Business

The Key Benefits of Using Social Media for Business

Worldwide, there are more than 2.6 billion social media...

Infographic: The Three Pillars of Digital Identity: Trust, Consent, Knowledge

Infographic: The Three Pillars of Digital Identity: Trust, Consent, Knowledge

8,434 adults were surveyed to gauge consumer awareness of...

FICO Scales with Oracle Cloud

FICO Scales with Oracle Cloud

Doug Clare, Vice President at FICO, describes how Oracle...

Talend and MapR Announce Certification of Big Data Integration and Big Data Quality

March 21, 2012 No Comments

SOURCE:  Talend

Provides Broad Capabilities for Integrating, Streaming and Cleansing Data into Hadoop

Los Altos, CA and San Jose, CA – March 21, 2012 – Talend, a global open source software leader, and MapR Technologies, Inc., the provider of the industry’s most advanced distribution for Apache Hadoop, announced today that Talend Open Studio for Big Data, the leading open source big data integration solution, has been certified to use with MapR’s Hadoop Distribution, providing users of Hadoop in the enterprise with high performance and scalable integration options.

Offered under the Apache Software License, Talend Open Studio for Big Data is a powerful and versatile open source solution for big data integration that natively supports Apache Hadoop. By leveraging Apache Hadoop’s MapReduce architecture for highly distributed data processing, Talend Open Studio for Big Data generates native Hadoop code and runs data transformations directly inside Hadoop for maximum scalability. Its easy-to-use graphical development environment dramatically improves the efficiency of data integration job design.

According to a recent report from Forrester Research, “Big data is all about processing and analyzing very large amounts of structured, unstructured, or semi-structured data very quickly. Although Hadoop offers a scalable data platform to process very large amounts of data for analytics purposes, getting data into the Hadoop platform is often not easy. Most leading ETL vendors are extending their solutions to integrate with Hadoop to offload large amounts of data from databases and data warehouses.” [1]

Talend Open Studio for Big Data bridges the gap between Hadoop and the rest of the information system, thanks to the broadest palette of connectors currently available on the market – both for Hadoop and many other data platforms and IT systems. Hadoop connectors included in Talend Open Studio for Big Data, and which are certified to use with MapR’s Distribution for Hadoop, include:

  • Hadoop Distributed File System (HDFS), to load and extract data from Hadoop, in batch or streaming mode.
  • NFS, to load, extract and stream data using MapR’s high performance standard file-base access.
  • HBase, to load, extract and transform data into Hadoop’s column-oriented database.
  • Pig (Pig Latin generation) and Hive (HiveQL generation), to process Hadoop data in place, leveraging the power of the MapR cluster.
  • Sqoop, for building direct Hadoop-to-database links without coding.

MapR’s Distribution is also available as the EMC Greenplum MR Edition and available with Cisco UCS.

To ensure connectivity of Hadoop with the rest of the information system, Talend Open Studio for Big Data also provides an unmatched set of connectors for:

  • Databases, such as Oracle, MS SQL Server, MySQL, DB2, PostgreSQL, Teradata, Vertica, EMC Greenplum, Infobright, etc.
  • ERP/CRM systems, such as SAP, MS Dynamics,, etc.
  • SaaS, Cloud and Social platforms, including Amazon Web Services, NetSuite, Marketo, Google, Twitter, etc.
  • Files – structured, poly-structured and semi-structured, flat, XML, etc.
  • Mainframes & midrange systems, Cobol files, etc.

“We are thrilled to announce the certification of Talend’s technology with MapR’s Distribution,” said Alan Geary, senior director of business development, MapR. “Organizations must be able to integrate Hadoop with the rest of their information systems in order to take advantage of big data for improved profitability and competitive advantage. Talend provides this ability.”

Talend Open Studio for Big Data is a core component of the Talend Platform for Big Data, which enables organizations to increase their productivity by deploying big data solutions in hours instead of weeks or months. The Talend Platform for Big Data easily integrates data of all types – structured, semi-structured and un-structured – and maximizes an organization’s resources by abstracting the technical complexity of big data tools and technologies. The Talend Platform for Big Data is compatible with all Apache Hadoop distributions.

“MapR provides one of the leading Hadoop distributions, and we’re pleased to be working with them,” said Fabrice Bonan, co-founder and chief operating officer, Talend. “Talend Open Studio for Big Data is set to democratize the deployment of Hadoop, helping organizations gather meaningful insight from their big data. By providing certified support for major implementations of Hadoop, Talend is making Hadoop more accessible and more integrated within the enterprise, providing scalable options for managing and analyzing big data.”

About MapR Technologies
MapR delivers on the promise of Hadoop, making managing and analyzing Big Data a reality for more business users. The award-winning MapR Distribution brings unprecedented dependability, speed and ease-of-use to Hadoop combined with data protection and business continuity, enabling customers to harness the power of Big Data analytics. The company is headquartered in San Jose, CA. Investors include Lightspeed Venture Partners, NEA and Redpoint Ventures. To download the latest MapR Distribution for Apache Hadoop, please visit

About Talend

Talend is the recognized market leader in open source integration solutions.  The company’s holistic integration platform helps organizations minimize costs and maximize the value of data integration, ETL, data quality, master data management, application integration and business process management, while supporting their shift toward the Cloud and Big Data.  More than 3,500 paying customers worldwide, including eBay, ING, The Weather Channel, Deutsche Post and Allianz, subscribe to Talend’s solutions and services. With over 20 million downloads, Talend’s products are the most trusted integration solutions in the world. The company has major offices in North America, Europe and Asia, and a global network of technical and services partners.  For more information, please visit

PR Contacts:

Juliet McGinnis

Talend, Inc.


Kim Leadley

PAN Communications



Leave a Reply