Skip to main content
BRILLIQS

Data & AI

The platforms, languages and governance tooling we use to move data from raw source to confident decision.

All Data & AI tools

Apache SparkData Engineering

Apache Spark

An engine for processing data across many machines, using the same code whether the data is small or large.

Explore
Apache KafkaData Engineering

Apache Kafka

A distributed log where records are appended and kept, so many consumers can read them at their own pace.

Explore
AlationData Management

Alation

A data catalogue that records what data exists and observes how it is actually queried to inform what it shows.

Explore
Apache AirflowData Engineering

Apache Airflow

An open source platform for building, scheduling and monitoring batch data workflows written in Python.

Explore
Apache SupersetData Visualization

Apache Superset

An open source platform for exploring data, building charts and publishing dashboards over any database it can connect to.

Explore
AtlanData Management

Atlan

A data workspace that brings cataloguing, lineage and collaboration together where the data team already works.

Explore
Azure Synapse AnalyticsData Modernization

Azure Synapse Analytics

An Azure service bringing warehouse queries, Spark processing and data integration together in one workspace.

Explore
Apache AtlasData Management

Apache Atlas

An open source metadata and governance framework built for Hadoop platforms, with a type system you extend.

Explore
Apache FlinkData Engineering

Apache Flink

A framework for processing continuous data streams with local state, event time and exactly once recovery.

Explore
AmundsenData Management

Amundsen

An open source data discovery tool built around search, ranking results by how heavily each asset is used.

Explore
Apache NiFiData Engineering

Apache NiFi

A platform for automating the flow of data between systems, built and controlled from a visual interface.

Explore
Amazon RedshiftData Modernization

Amazon Redshift

A data warehouse service on AWS that stores data by column and distributes it across the nodes of a cluster.

Explore
Apache BeamData Engineering

Apache Beam

One programming model for batch and streaming pipelines that can run on several different processing engines.

Explore
AtaccamaData Management

Ataccama

A data management platform combining cataloguing, quality rules and master data handling in one product.

Explore
Apache HadoopData Engineering

Apache Hadoop

A framework for storing and processing large datasets across a cluster of machines.

Explore
Apache HiveData Engineering

Apache Hive

A data warehouse system that lets you query very large files in distributed storage using SQL.

Explore
Apache IcebergData Engineering

Apache Iceberg

An open table format that lets several engines read and write the same large analytic tables safely.

Explore
CatBoostData Analytics

CatBoost

A gradient boosting library built around handling categorical columns without the leakage that naive encoding causes.

Explore
Apache HudiData Engineering

Apache Hudi

A data lake platform built around record keys, so individual rows can be updated and changes can be read incrementally.

Explore
Amazon QuickSightData Visualization

Amazon QuickSight

An AWS business intelligence service that builds dashboards over data either queried directly or held in its own memory engine.

Explore
Apache PulsarData Engineering

Apache Pulsar

A messaging and streaming platform that keeps its serving layer and its storage layer separate.

Explore
Amazon KinesisData Engineering

Amazon Kinesis

An AWS service for collecting streaming records continuously and making them available to consumer applications.

Explore
Amazon S3Data Modernization

Amazon S3

An object storage service from AWS where files are stored under keys in buckets rather than in a filesystem.

Explore
Amazon SageMakerData Analytics

Amazon SageMaker

AWS machine learning service covering preparation, training, tuning and hosting, with infrastructure provisioned per job.

Explore
Azure Data Lake StorageData Modernization

Azure Data Lake Storage

Azure object storage with a hierarchical namespace, so directories are real rather than implied by naming.

Explore
Azure Machine LearningData Analytics

Azure Machine Learning

The Microsoft Azure service for machine learning, organised around a workspace that holds assets and compute.

Explore
AirbyteData Engineering

Airbyte

An open source platform for moving data between sources and destinations, with connectors you can also build yourself.

Explore
AWS Database Migration ServiceData Modernization

AWS Database Migration Service

An AWS service that copies data from a source database to a target and can keep applying changes as they occur.

Explore
Azure MigrateData Modernization

Azure Migrate

An Azure service for discovering existing servers and databases, assessing what they need, and migrating them.

Explore
AlteryxData Analytics

Alteryx

A tool where data preparation and analysis are built by dragging tools onto a canvas and connecting them into a workflow.

Explore
AWS GlueData Engineering

AWS Glue

A serverless AWS service that catalogues your data and runs the jobs that prepare it, without a cluster to manage.

Explore
Azure Data FactoryData Engineering

Azure Data Factory

A cloud service for building pipelines that move data between systems and orchestrate the steps around that movement.

Explore
ClouderaData Modernization

Cloudera

A commercial data platform assembling Hadoop ecosystem projects with management, security and support around them.

Explore
Apache CassandraData Management

Apache Cassandra

A distributed database in which every node is equal, designed to keep accepting writes while parts of the cluster are unavailable.

Explore
AWS Lake FormationData Modernization

AWS Lake Formation

An AWS service for granting access to data lake tables and columns rather than to the underlying storage paths.

Explore
BokehData Visualization

Bokeh

A Python library that produces interactive plots which run in a browser, with an optional server for Python driven behaviour.

Explore
Amazon AthenaData Modernization

Amazon Athena

A query service that runs SQL against files in Amazon S3 without any infrastructure being provisioned.

Explore
Apache EChartsData Visualization

Apache ECharts

An open source JavaScript charting library where a whole chart is described by one configuration object.

Explore
Amazon EMRData Engineering

Amazon EMR

An AWS platform for running open source big data frameworks such as Spark, Hive and Presto on managed clusters.

Explore
Azure Data ExplorerData Modernization

Azure Data Explorer

An Azure service for exploring large volumes of log and telemetry data, queried with the Kusto Query Language.

Explore
Apache DruidData Engineering

Apache Druid

A database built for fast analytical queries over event data, where time is treated as a first class column.

Explore
Chart.jsData Visualization

Chart.js

A small open source JavaScript charting library covering the common chart types with little configuration.

Explore
Amazon DynamoDBData Management

Amazon DynamoDB

A managed key value and document database from AWS where servers are not provisioned and capacity is a setting.

Explore
AWS Schema Conversion ToolData Modernization

AWS Schema Conversion Tool

A tool that converts database schemas and code from one engine to another and reports what it could not convert.

Explore
ClickHouseData Engineering

ClickHouse

A column oriented database management system that answers analytical SQL queries over very large tables.

Explore
Apache StormData Engineering

Apache Storm

A distributed system for processing unbounded streams of records as they arrive, one record at a time.

Explore
CockroachDBData Management

CockroachDB

A distributed SQL database that spreads data across nodes while keeping transactions and consistency across all of them.

Explore
BIRTData Visualization

BIRT

An Eclipse project providing a report designer and a Java reporting engine that applications can embed.

Explore
Apache HBaseData Management

Apache HBase

A column family store that runs on Hadoop storage and provides record level reads and writes over very large tables.

Explore
BigIDData Management

BigID

A discovery platform that finds personal and sensitive data across systems and works out whose data it is.

Explore
Apache ArrowData Engineering

Apache Arrow

A standard way of laying out columnar data in memory so different tools and languages can share it without converting.

Explore
Apache SqoopData Engineering

Apache Sqoop

A command line tool for bulk transfer of data between relational databases and Hadoop storage.

Explore
Apache FlumeData Engineering

Apache Flume

A service for collecting log data from many machines and delivering it to central storage through configured agents.

Explore
Apache RangerData Management

Apache Ranger

An open source framework for defining and enforcing access policies across Hadoop and related data services.

Explore
AmplitudeData Analytics

Amplitude

A product analytics service for examining how people use a product, built around events and the users who produce them.

Explore
CARTOData Visualization

CARTO

A location analysis platform that runs spatial queries inside your cloud data warehouse rather than moving data out of it.

Explore
AcceldataData Management

Acceldata

An observability platform covering data quality, pipeline behaviour and the performance of the systems underneath.

Explore
amChartsData Visualization

amCharts

A commercial JavaScript charting library covering charts, maps and stock charts, built around composable objects.

Explore
Apache OozieData Engineering

Apache Oozie

A workflow scheduler for Hadoop jobs, where workflows are defined in XML and can wait for data before running.

Explore
BigeyeData Management

Bigeye

A data quality monitoring platform that learns normal behaviour for a table and alerts when readings depart from it.

Explore
Apache ZooKeeperData Engineering

Apache ZooKeeper

A coordination service that distributed systems use to agree on shared state such as configuration and leadership.

Explore
BigQuery MLData Analytics

BigQuery ML

A capability that trains and applies machine learning models using SQL statements inside a data warehouse.

Explore
Amazon RDSData Management

Amazon RDS

A managed service from AWS that runs standard database engines and handles patching, backups and failover.

Explore
AnacondaData Analytics

Anaconda

A Python distribution and package manager that installs libraries along with the compiled components they depend on.

Explore
Azure SQL DatabaseData Management

Azure SQL Database

A managed SQL Server database service on Azure where the database is provided without a server being managed.

Explore
Apache ZeppelinData Analytics

Apache Zeppelin

A notebook where different paragraphs can use different languages, sharing results between them in one note.

Explore
Amazon AuroraData Management

Amazon Aurora

A database built by AWS with PostgreSQL and MySQL compatibility and a storage layer designed for the cloud.

Explore
All tools

Planning a data & ai initiative?

We will help you choose the stack before you commit to it.

Book a Consultation