Apache Spark
An engine for processing data across many machines, using the same code whether the data is small or large.
ExploreApache Kafka
A distributed log where records are appended and kept, so many consumers can read them at their own pace.
ExploreAWS
Amazon's cloud platform, offering a large range of services for computing, storage, databases and more.
ExploreAlation
A data catalogue that records what data exists and observes how it is actually queried to inform what it shows.
ExploreApache Airflow
An open source platform for building, scheduling and monitoring batch data workflows written in Python.
ExploreApache Superset
An open source platform for exploring data, building charts and publishing dashboards over any database it can connect to.
ExploreAtlan
A data workspace that brings cataloguing, lineage and collaboration together where the data team already works.
ExploreAzure Synapse Analytics
An Azure service bringing warehouse queries, Spark processing and data integration together in one workspace.
ExploreApache Atlas
An open source metadata and governance framework built for Hadoop platforms, with a type system you extend.
ExploreNot sure which tool fits your problem?
We will tell you when the answer is the one you already have.