Try our new research platform with insights from 80,000+ expert users

IBM InfoSphere Information Server vs StreamSets comparison

 

Comparison Buyer's Guide

Executive SummaryUpdated on Dec 19, 2024

Review summaries and opinions

We asked business professionals to review the solutions they use. Here are some excerpts of what they said:
 

Categories and Ranking

IBM InfoSphere Information ...
Ranking in Data Integration
29th
Average Rating
8.2
Reviews Sentiment
5.8
Number of Reviews
8
Ranking in other categories
Metadata Management (6th)
StreamSets
Ranking in Data Integration
22nd
Average Rating
8.4
Reviews Sentiment
7.0
Number of Reviews
21
Ranking in other categories
No ranking in other categories
 

Mindshare comparison

As of January 2026, in the Data Integration category, the mindshare of IBM InfoSphere Information Server is 1.0%, up from 0.8% compared to the previous year. The mindshare of StreamSets is 1.2%, down from 1.6% compared to the previous year. It is calculated based on PeerSpot user engagement data.
Data Integration Market Share Distribution
ProductMarket Share (%)
StreamSets1.2%
IBM InfoSphere Information Server1.0%
Other97.8%
Data Integration
 

Featured Reviews

MI
Senior Data Engineer at Mohammed Mansour Alrumiah
Faced challenges with customer support and documentation but have benefited from reliable data integration over the years
As for utilizing the platform's metadata management feature, I have not worked on that feature yet, but personally, I have done that. To evaluate the effectiveness of IBM InfoSphere Information Server's data integration capabilities, if IBM is providing all the solutions we are using, then it is definitely a helpful thing. Mostly, the other thing is that it is a big area including data governance, data lineage, data management, and metadata, but every customer is not putting that much effort and money on that. They mostly migrate the data, use it, and forget it, but slowly things are changing. I am working in Saudi Arabia, so here also data governance, data management, and those kinds of things are getting attention. Regarding how scalable IBM InfoSphere Information Server is, I need to learn how to tune performance and scalability on the cloud. I am familiar with localized hardware, but on the cloud, I still have to do the work around it. In the beginning, we estimate the load and based on that, we put the hardware, but if there is continuous increase, I believe IBM also faces problems. Scalability needs to be improved because once the demand comes, you should be able to improve it, but for that, documentation on how to add hardware or resources to the software needs to be proper. I do not have much hands-on experience with that.
SS
Enterprise Solutions Architect at a energy/utilities company with 1,001-5,000 employees
Enables effective batch loading with visual interface and enterprise support
One issue I observed with StreamSets is that the memory runs out quickly when processing large volumes of data. Because of this memory issue, we have to upgrade our EC2 boxes in the Amazon AWS infrastructure. I had to switch to a new EC2 box, even though the processor was not fully utilized. It would be beneficial if StreamSets addressed any potential memory leak issues to prevent unnecessary upgrades. Additionally, it would be a great enhancement if StreamSets could produce a lineage graph to visualize how the data has passed through the system.

Quotes from Members

We asked business professionals to review the solutions they use. Here are some excerpts of what they said:
 

Pros

"This solution is extremely flexible and scalable."
"The integration with different technologies is the most valuable feature."
"IBM InfoSphere Information Server is stable."
"Stability-wise, I rate the solution a ten out of ten."
"Over the years of working with IBM InfoSphere Information Server, I see basically the strength of the tool, capability, and load balancing, which I see is really good."
"StreamSets Transformer is a good feature because it helps you when you are developing applications and when you don't want to write a lot of code. That is the best feature overall."
"Also, the intuitive canvas for designing all the streams in the pipeline, along with the simplicity of the entire product are very big pluses for me. The software is very simple and straightforward. That is something that is needed right now."
"The ability to have a good bifurcation rate and fewer mistakes is valuable."
"The most valuable feature is the pipelines because they enable us to pull in and push out data from different sources and to manipulate and clean things up within them."
"One of the things I like is the data pipelines. They have a very good design. Implementing pipelines is very straightforward. It doesn't require any technical skill."
"StreamSets’ data drift resilience has reduced the time it takes us to fix data drift breakages. For example, in our previous Hadoop scenario, when we were creating the Sqoop-based processes to move data from source to destinations, we were getting the job done. That took approximately an hour to an hour and a half when we did it with Hadoop. However, with the StreamSets, since it works on a data collector-based mechanism, it completes the same process in 15 minutes of time. Therefore, it has saved us around 45 minutes per data pipeline or table that we migrate. Thus, it reduced the data transfer, including the drift part, by 45 minutes."
"The most valuable features are the option of integration with a variety of protocols, languages, and origins."
"The ETL capabilities are very useful for us. We extract and transform data from multiple data sources, into a single, consistent data store, and then we put it in our systems. We typically use it to connect our Apache Kafka with data lakes. That process is smooth and saves us a lot of time in our production systems."
 

Cons

"This solution would benefit from the engine being made more lightweight."
"There are certain shortcomings in the cloud side of the solution, where improvements are required."
"Their technical support needs improvement."
"IBM InfoSphere Information Server should be more scalable. It should have the option to change the configuration to run on a single, non-multiple node, or multi-threading processing."
"Unlike other tools, IBM tools do not provide much help from the internet, so additional support should be available."
"They need to improve their customer care services. Sometimes it has taken more than 48 hours to resolve an issue. That should be reduced. They are aware of small or generic issues, but not the more technical or deep issues. For those, they require some time, generally 48 to 72 hours to respond. That should be improved."
"One thing that I would like to add is the ability to manually enter data. The way the solution currently works is we don't have the option to manually change the data at any point in time. Being able to do that will allow us to do everything that we want to do with our data. Sometimes, we need to manually manipulate the data to make it more accurate in case our prior bifurcation filters are not good. If we have the option to manually enter the data or make the exact iterations on the data set, that would be a good thing."
"Visualization and monitoring need to be improved and refined."
"If you use JDBC Lookup, for example, it generally takes a long time to process data."
"We often faced problems, especially with SAP ERP. We struggled because many columns weren't integers or primary keys, which StreamSets couldn't handle. We had to restructure our data tables, which was painful. Also, pipeline failures were common, and data drifting wasn't addressed, which made things worse. Licensing was another issue we encountered."
"One issue I observed with StreamSets is that the memory runs out quickly when processing large volumes of data. Because of this memory issue, we have to upgrade our EC2 boxes in the Amazon AWS infrastructure."
"In terms of the product, I don't think there is any room for improvement because it is very good. One small area of improvement that is very much needed is on the knowledge base side. Sometimes, it is not very clear how to set up a certain process or a certain node for a person who's using the platform for the first time."
"Sometimes, it is not clear at first how to set up nodes. A site with an explanation of how each node works would be very helpful."
 

Pricing and Cost Advice

"The licensing cost of IBM InfoSphere Information Server depends on how many users there are."
"The pricing is affordable for any business."
"The licensing is expensive, and there are other costs involved too. I know from using the software that you have to buy new features whenever there are new updates, which I don't really like. But initially, it was very good."
"There are two editions, Professional and Enterprise, and there is a free trial. We're using the Professional edition and it is competitively priced."
"We use the free version. It's great for a public, free release. Our stance is that the paid support model is too expensive to get into. They should honestly reevaluate that."
"The pricing is too fixed. It should be based on how much data you need to process. Some businesses are not so big that they process a lot of data."
"Its pricing is pretty much up to the mark. For smaller enterprises, it could be a big price to pay at the initial stage of operations, but the moment you have the Seed B or Seed C funding and you want to scale up your operations and aren't much worried about the funds, at that point in time, you would need a solution that could be scaled."
"It has a CPU core-based licensing, which works for us and is quite good."
"We are running the community version right now, which can be used free of charge."
report
Use our free recommendation engine to learn which Data Integration solutions are best for your needs.
881,082 professionals have used our research since 2012.
 

Top Industries

By visitors reading reviews
Government
18%
Financial Services Firm
16%
Insurance Company
8%
Retailer
7%
Insurance Company
8%
Financial Services Firm
8%
Manufacturing Company
8%
Computer Software Company
8%
 

Company Size

By reviewers
Large Enterprise
Midsize Enterprise
Small Business
By reviewers
Company SizeCount
Small Business5
Midsize Enterprise1
Large Enterprise3
By reviewers
Company SizeCount
Small Business9
Midsize Enterprise2
Large Enterprise11
 

Questions from the Community

What needs improvement with IBM InfoSphere Information Server?
As for utilizing the platform's metadata management feature, I have not worked on that feature yet, but personally, I have done that. To evaluate the effectiveness of IBM InfoSphere Information Ser...
What is your primary use case for IBM InfoSphere Information Server?
My usual use case for IBM InfoSphere Information Server is ETL, where we take data from one source to another data warehouse solution.
What advice do you have for others considering IBM InfoSphere Information Server?
Currently, IBM InfoSphere Information Server is deployed on-premises in my organization. Mostly it is on-premises only, but slowly things are changing towards pro-cloud. It will not be a public clo...
What do you like most about StreamSets?
The best thing about StreamSets is its plugins, which are very useful and work well with almost every data source. It's also easy to use, especially if you're comfortable with SQL. You can customiz...
What needs improvement with StreamSets?
One issue I observed with StreamSets is that the memory runs out quickly when processing large volumes of data. Because of this memory issue, we have to upgrade our EC2 boxes in the Amazon AWS infr...
What is your primary use case for StreamSets?
We are using StreamSets for batch loading.
 

Also Known As

InfoSphere Information Server, IBM Information Server
No data available
 

Overview

 

Sample Customers

Canadian National Railway Company, Chickasaw Nation Division of Commerce, Swedish Armed Forces, BG RCI, Janata Sahakari Bank Ltd., University of Arizona, Biogrid Australia
Availity, BT Group, Humana, Deluxe, GSK, RingCentral, IBM, Shell, SamTrans, State of Ohio, TalentFulfilled, TechBridge
Find out what your peers are saying about IBM InfoSphere Information Server vs. StreamSets and other solutions. Updated: January 2026.
881,082 professionals have used our research since 2012.