DataTorrent tackles complexity of Hadoop data ingestion

It designed dtIngest to streamline the collection, aggregation and transfer of data to and from a Hadoop cluster

While the buzz around big data analysis is at a peak, there is less discussion about how to get the necessary data into the systems in the first place, which can involve the cumbersome task of setting up and maintaining a number of data processing pipelines.

To help solve this problem, Santa Clara, California start-up DataTorrent has released what it calls the first enterprise-grade ingestion application for Hadoop, DataTorrent dtIngest.

The application is designed to streamline the process of collecting, aggregating, and moving data onto and off of a Hadoop cluster.

The software is based on Project Apex, an open source software package available under the Apache 2.0 license.

Working as a component within a Hadoop platform, dtIngest can work with both streaming and batch data. It can exchange data across a variety of file systems and protocols, including NFS, FTP, the Hadoop File System, Amazon Web Service's Simple Storage Service (S3), Kafka, and the Java Message Service.

The software is fault tolerant, in that it can resume a file transfer automatically after disruption. It comes with a point-and-click interface, as well as monitoring logs.

The company has released dtIngest for free, hoping that users will upgrade to DataTorrent's enterprise Hadoop data ingestion pipeline software, DataTorrent RTS 3, which is based on dtIngest/Project Apex and includes additional capabilities for operational management, easy development and data visualization.

DataTorrent was co-founded by Amol Kekre and Phu Hoang, a pair of engineers who used to work at Hadoop pioneer Yahoo. The company has formed partnerships with Hadoop distributors Hortonworks and Pivotal, and has drummed up nearly $24 million in early stage funding from investors.

Joab Jackson covers enterprise software and general technology breaking news for The IDG News Service. Follow Joab on Twitter at @Joab_Jackson. Joab's e-mail address is Joab_Jackson@idg.com

Join the newsletter!

Or

Sign up to gain exclusive access to email subscriptions, event invitations, competitions, giveaways, and much more.

Membership is free, and your security and privacy remain protected. View our privacy policy before signing up.

Error: Please check your email address.

Tags Data managementsoftwareapplicationsdata miningDataTorrent

Keep up with the latest tech news, reviews and previews by subscribing to the Good Gear Guide newsletter.

Joab Jackson

IDG News Service
Show Comments

Brand Post

Most Popular Reviews

Latest Articles

Resources

PCW Evaluation Team

Andrew Teoh

Brother MFC-L9570CDW Multifunction Printer

Touch screen visibility and operation was great and easy to navigate. Each menu and sub-menu was in an understandable order and category

Louise Coady

Brother MFC-L9570CDW Multifunction Printer

The printer was convenient, produced clear and vibrant images and was very easy to use

Edwina Hargreaves

WD My Cloud Home

I would recommend this device for families and small businesses who want one safe place to store all their important digital content and a way to easily share it with friends, family, business partners, or customers.

Walid Mikhael

Brother QL-820NWB Professional Label Printer

It’s easy to set up, it’s compact and quiet when printing and to top if off, the print quality is excellent. This is hands down the best printer I’ve used for printing labels.

Ben Ramsden

Sharp PN-40TC1 Huddle Board

Brainstorming, innovation, problem solving, and negotiation have all become much more productive and valuable if people can easily collaborate in real time with minimal friction.

Sarah Ieroianni

Brother QL-820NWB Professional Label Printer

The print quality also does not disappoint, it’s clear, bold, doesn’t smudge and the text is perfectly sized.

Featured Content

Product Launch Showcase

Latest Jobs

Don’t have an account? Sign up here

Don't have an account? Sign up now

Forgot password?