Starburst Galaxy logo

    Starburst Galaxy

    Sold by
    Starburst Galaxy offers a full-featured data lake analytics platform that allows you to discover, manage, and consume the data in and around your data lake.

    Ratings and reviews

    4.3
    126 ratings
    55%
    43%
    1%
    0%
    1%
    8 AWS reviews
    |
    118 external reviews
    External reviews are from G2  and PeerSpot .

    Filters

    Review type

    AWS Marketplace reviews
    External reviews
    Reviews (126)
    Pedromachado Ventura

    Unified SQL layer has streamlined access to distributed historical data for analytics and reporting

    Reviewed on Sep 22, 2026
    Review provided by PeerSpot

    What is our primary use case?

    I have mainly used Starburst Galaxy as a query and access layer for analytical data, as my current work involves dealing with data stored across different platforms, including Hive and SingleStore. Starburst gives us a more convenient way to query that data and expose it to reporting tools such as Power BI, so it is more for accessing and querying data used for analytics and reporting. One of the main use cases I utilize is working with historical data stored in Hive and making that data available for reporting and monitoring purposes. We also use it as part of a data access layer between the underlying platforms and Power BI.

    The main benefit of using Starburst Galaxy as a data access layer between my storage platforms and Power BI is easier access to the data for analytics and reporting. It also helps create a more consistent access layer, particularly when working with historical data sets, instead of building separate connections and logic for every source.

    I would recommend Starburst Galaxy, particularly for organizations that already have data distributed across different platforms and want or need a common SQL layer for analytics. It makes the most sense when there is a real need to query data where it already lives instead of moving everything into one system. I mainly use this for an analytics and reporting perspective and would probably give it an eight out of ten. It works well for analytical use cases, at least the ones I have been involved with, and makes access to distributed data easier. There is still room for improved usability and diagnostics, especially for users who are not platform specialists. I am mainly a data consumer rather than a Starburst Galaxy administrator, so my experience is focused on querying, reporting, performance, and usability.

    What is most valuable?

    One of the best features of Starburst Galaxy is that it can sit between different data sources and the reporting layer. This separation is useful because the reporting tool does not necessarily need to know all the complexity of where the data is physically stored. I think it helps create a more consistent access layer, particularly when working with historical data sets, instead of building separate connections and logic for every point of access.

    The main time-saving for me is having a common SQL access layer to the data. Instead of working directly with the different underlying systems, I can query what I need directly through Starburst Galaxy. In my case, this is particularly useful for reporting. For example, when building Power BI dashboards on top of historical data in Hive, Starburst Galaxy makes that data much easier to access and consume. Overall, it reduces the amount of work needed to get the data into a usable form and lets me spend more time on actual analysis and reporting. A good example is for historical reporting with data stored in Hive—rather than building a complete separate reporting process around that storage, we can access it through Starburst Galaxy for Power BI reporting. This simplifies the architecture from my perspective as an analyst and saves time when accessing and analyzing historical data. I have not measured the time saving precisely, but it definitely reduces the manual effort and complexity involved in accessing data from these different sources.

    The main positive impact Starburst Galaxy has made is making data more accessible for analytics and reporting. We have data across different platforms, including historical data in Hive, and Starburst Galaxy provides a common layer to access that information, making it easier to build reporting solutions without creating completely separate data access processes for every source. From an organizational perspective, that means less complexity, faster access to information, and better use of the data we already possess. In my use case, it also helps teams focus more on analysis and monitoring rather than on how to retrieve the data itself. The biggest impact is really easier access to the data.

    What needs improvement?

    One area that I think could be improved is the experience when performance issues occur. When a query is slow, it is not always immediately obvious to me whether the bottleneck comes from Starburst Galaxy itself, the underlying data source, the query design, or the reporting tool. Better visibility into query performance and easier diagnostics for non-administrators would be useful.

    Another potential improvement would be enhancing the experience with BI tools to make it more seamless. I work a lot with Power BI, and when you are working with larger data sets, performance can sometimes depend on several different layers. Having more visibility into what is happening between the BI tool, Starburst Galaxy, and the underlying source would be helpful. I also think onboarding could be a little more accessible for analysts. There is good technical documentation, but sometimes I just need to understand the best way to approach a common use case without diving too deep into the platform architecture. The main improvement would be troubleshooting.

    I have not used the AI capabilities extensively, so I cannot give a detailed assessment. I am not sure if my organization has the full capabilities of Starburst Galaxy, but I think adding AI on top of the data layer is interesting, especially if it can help users discover data, understand data sets, and interact with them more naturally.

    For governance and security, one of the strengths of Starburst Galaxy is that you can centralize access to data while still controlling what different users are allowed to see. Role-based access, fine-grained permissions, and data masking are important because giving people easier access to data should not mean giving everyone access to everything. I think that is even more important than any AI capabilities that are introduced. If you do introduce AI, I think it should respect exactly the same data permissions and governance rules as the user that is accessing the data.

    For how long have I used the solution?

    I have been using Starburst Galaxy for the past one and a half to two years.

    What do I think about the stability of the solution?

    Starburst Galaxy has been stable so far.

    What do I think about the scalability of the solution?

    My experience with Starburst Galaxy's scalability has been good. We work with large volumes of data, particularly historical data sets in Hive, and it allows us to query that data without moving everything into a separate system first. This is one of the advantages—as the amount of data grows, we can continue accessing it through the same SQL layer. I do not manage infrastructure directly, so I cannot comment on the technical scaling configuration, but from a user's perspective, it has handled our analytical and reporting use cases well.

    Which solution did I use previously and why did I switch?

    We previously used SingleStore before switching to Starburst Galaxy because we needed access to different data layers in different sources. That need led us to change to Starburst Galaxy directly, to have a unified central layer that can connect to all external data sources.

    What was our ROI?

    I do not have a specific percentage or cost-saving figure that I can confidently attribute to Starburst Galaxy alone. The impact I can see directly is more operational. For example, we have been able to use historical data stored in Hive for Power BI reporting through a common SQL access layer rather than creating separate extraction processes for each use case. The measurable outcome from my perspective includes reduced complexity in data access and less manual work for me, making historical data available for reporting with faster deliveries when dealing with monitoring and analytical use cases. From my day-to-day experience, it clearly reduces the number of steps required to access and consume data for reporting.

    Which other solutions did I evaluate?

    I cannot share whether other options were evaluated because I was not the one who decided that.

    What other advice do I have?

    My advice would be to first be very clear about the use case. Starburst Galaxy makes a lot of sense if you have data distributed across different systems and want a common SQL layer without constantly moving or duplicating it. I would recommend starting with a clear use case rather than just implementing the technology. If you have data in different platforms and want access through a common SQL layer, then Starburst Galaxy can be very useful. I would give this product an eight out of ten.

    shivam s.

    Simplifies Data Access and Analytics Across Sources

    Reviewed on Sep 16, 2026
    Review provided by G2
    What do you like best about the product?
    I like that Starburst makes it easy to access and analyze data across different sources as it's fast and reliable, simplifying our data workflows without adding complexity. The federated query capabilities, data source connectors, and fast SQL-based analytics are features I especially value, as they make it easy to work with data across multiple systems without having to move everything into one place. I also appreciate how Starburst connects data from different sources, allowing us to query it in one location, making data access simple, fast, and efficient. Plus, its ability to integrate seamlessly with our existing databases and BI/analytics tools using SQL has made it a perfect fit in our data stack.
    What do you dislike about the product?
    The overall experience is good, but the initial setup and configuration can take some time. More beginner-friendly documentation and simpler configuration options would make it easier for new users to get started.
    What problems is the product solving and how is that benefiting you?
    Starburst simplifies our data workflows by providing fast, reliable access and analytics across different sources, reducing data movement and eliminating silos. It helps us query data efficiently in one place, making analytics faster and more effective.
    Sainy t.

    Fast, Streamlined Multi-Source Queries with Starburst

    Reviewed on Sep 08, 2026
    Review provided by G2
    What do you like best about the product?
    The best features of Starburst that I like most are its multi-source data access and the performance of the Trino engine. Without needing data migration, queries run super fast and workflows get streamlined. The AI-driven query optimization is an unexpected bonus that helps reduce my workload.
    What do you dislike about the product?
    Starburst’s pricing feels steep for smaller teams, the advanced setup can be complicated for non-technical users, and the dashboard customization options are fairly limited, which reduces overall flexibility.
    What problems is the product solving and how is that benefiting you?
    Starburst helped us solve our scattered data problem by enabling centralized queries across cloud, on‑prem, and lakehouses. It reduced our ETL costs, improved query speed, and made our workflows more streamlined and consistent.
    Nadeem Ahmad K.

    Enhanced Data Integration & Analysis with Stellar Flexibility

    Reviewed on Sep 08, 2026
    Review provided by G2
    What do you like best about the product?
    I like that Starburst lets us easily access and query data from multiple sources, simplifying data analysis and improving accessibility. It speeds up insights without having to move data around. I also appreciate its flexibility and performance, as it integrates well with different data sources, which makes it easier for our team to work with data efficiently. The initial setup was fairly straightforward, and I find the documentation and configuration options helpful for getting started. Overall, it’s a reliable and flexible solution that makes working with data from multiple sources much easier.
    What do you dislike about the product?
    The initial learning curve can be a bit steep for new users. The interface and some advanced configurations could be more intuitive. The interface could be more intuitive by simplifying navigation and making common tasks easier to find. Clearer setup guidance, better documentation, and more straightforward configuration options would also help new users get started faster.
    What problems is the product solving and how is that benefiting you?
    Starburst lets us access and analyze data from multiple sources in one place, simplifying queries and making analytics faster and efficient. It solves data integration challenges, improving accessibility and enabling insights without data duplication.
    Computer Software

    Starburst Makes Cross-Source Queries Fast and Easy

    Reviewed on Sep 07, 2026
    Review provided by G2
    What do you like best about the product?
    What I like most about Starburst is being able to query data across multiple sources from a single place, without having to move or duplicate anything. It delivers solid query performance and makes it easier for data and analytics teams to access and work with large datasets. I also appreciate its integrations with other data and BI tools, along with the governance and access-control capabilities it provides.
    What do you dislike about the product?
    The main downside is that Starburst can come with a learning curve, particularly when you’re dealing with more complex queries and configurations. Performance can also be inconsistent with very large datasets or highly complex queries, so some optimization may be necessary. I also feel the documentation could include more practical, hands-on examples for advanced use cases.
    What problems is the product solving and how is that benefiting you?
    Starburst helps solve the challenge of having data spread across multiple systems, which can make it hard to access and analyze. It lets us query data from different sources in one place without needing to move or duplicate it. That saves time, simplifies analysis, and makes it easier to get the information we need to support faster decision-making.
    Alternative Dispute Resolution

    Effortless Data Querying & Analysis

    Reviewed on Sep 04, 2026
    Review provided by G2
    What do you like best about the product?
    useless never used it I only did the review because g2.com asked and they was paying 20 or 25 euros through an ad they sent through linked in I never did any querying on its interface or ui
    What do you dislike about the product?
    all good I guessdddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddddd
    What problems is the product solving and how is that benefiting you?
    I use Starburst to simplify data access and avoid duplicating large amounts of data, making it easier to run queries and analyze information efficiently.
    Information Technology and Services

    One Query, Zero Data Movement - Starburst Nailed it!

    Reviewed on Sep 02, 2026
    Review provided by G2
    What do you like best about the product?
    What I like most about Starburst is the freedom to query from anywhere. There’s no vendor lock-in, no need to duplicate data, and no waiting around for pipelines to finish. I can just query what I need, right where it already lives. Overall, it feels fast, clean, and simple.
    What do you dislike about the product?
    If I had to pick worst thing, it would be the combination of cost and complexity. Starburst is a powerful tool, but you don’t really get the best out of it without some setup and tuning. And once people start running lots of heavy federated queries, the usage-based pricing can climb quickly unless you actively monitor it and keep it under control.
    What problems is the product solving and how is that benefiting you?
    The main issue for us was data fragmentation. Our data was spread across AWS, Snowflake, and on-prem systems, and Starburst brought it together under a single query layer. The benefits were clear: we save hours every week, our storage costs are lower, and we’re not locked into any one vendor. It basically solved our overall data headache in one shot.
    Recommendations to others considering the product:
    Recommendations not provided in the input.
    Puty Y.

    Clean, Unified Console for Visibility, Workload Monitoring, and Access Control

    Reviewed on Sep 02, 2026
    Review provided by G2
    What do you like best about the product?
    The web console and management interface provide clean visibility into cluster health, query execution paths, and resource usage. Navigating the UI makes it straightforward to monitor active workloads, administer fine-grained access control, and manage data products from a single pane of glass without jumping across disparate cloud environments. Good project.
    What do you dislike about the product?
    Cost and Resource Tuning While the performance is exceptional, optimizing resource consumption and memory limits for massive, highly concurrent federated queries requires deep tuning expertise. Misconfigured queries or missing pushdowns can inadvertently trigger high cloud infrastructure costs. Additionally, enterprise licensing and managed service pricing can scale steeply for smaller teams, making it a platform best suited for large-scale data environments.
    What problems is the product solving and how is that benefiting you?
    Data Silos and Expensive ETL Pipelines Enterprise data is naturally fragmented across multi-cloud object storage (such as Amazon S3 or Google Cloud Storage), traditional relational databases, and modern data lakehouses (like Apache Iceberg). Starburst solves the core problem of data isolation by serving as a distributed SQL query engine that allows us to query data directly where it lives—eliminating the need to build and maintain complex, brittle ETL pipelines just to run a join.
    Telecommunications

    Effortlessly Manages Diverse Data Sources

    Reviewed on Sep 01, 2026
    Review provided by G2
    What do you like best about the product?
    I like how Starburst lets me query data from different sources all in one place, making data management much simpler. I also appreciate the flexibility and performance it offers, especially in handling large datasets efficiently. Another thing I find valuable is that it allows working with various data without needing custom integrations, which is neat.
    What do you dislike about the product?
    I think the learning curve for Starburst is a bit steep; it takes a while to learn.
    What problems is the product solving and how is that benefiting you?
    I use Starburst to manage and analyze data from different sources in one place, solving the challenge of working with data from different systems. It lets me query everything in one place, handles large datasets well, and allows flexibility without custom integrations.
    paresh j.

    Starburst Galaxy: Reliable, Scalable SQL Queries Across Data Silos

    Reviewed on Aug 20, 2026
    Review provided by G2
    What do you like best about the product?
    What I like most is being able to query disparate data silos with standard SQL without having to move the data. This has significantly reduced data duplication and improved time-to-insight for our analytics team. The managed setup in Starburst Galaxy has been reliable for us, is straightforward to monitor, and scales smoothly with our query workloads.
    What do you dislike about the product?
    Cost tracking and resource consumption can ramp up quickly if cluster sizing and auto-suspension policies aren’t carefully managed, especially during peak query periods. Dynamic scaling is definitely convenient, but I’d like to see more granular, out-of-the-box cost-allocation and budgeting dashboards at the workspace level to make administrative oversight simpler.
    What problems is the product solving and how is that benefiting you?
    It addresses the challenge of managing fragmented data architecture and the significant maintenance overhead that often comes with traditional data warehouse migrations.

    How it benefits us: By querying data where it resides, we save valuable engineering time, strengthen data governance through unified access controls, and give our team immediate access to live data across all our storage systems.