Amundsen
A data discovery and metadata catalog platform, originally built at Lyft, for finding and understanding datasets across an organization.
Declarative grammar for interactive visualizations, providing the Vega and Vega-Lite runtimes plus a CLI that renders specifications to SVG, PNG or PDF. Built and maintained by WIEWAVE for Azure Marketplace, on Ubuntu and Debian.
| Built & maintained by | WIEWAVE |
|---|---|
| Category | Analytics & Big Data |
| Operating systems | Ubuntu, Debian |
| Marketplaces | Azure Marketplace |
| Offer types | Public listing, with private offers on request |
| Marketplace | Status |
|---|---|
| Azure Marketplace | Published |
| AWS Marketplace | Available on request |
| Google Cloud Marketplace | Available on request |
Need it on another marketplace, or as a private offer for your organisation? cloud@wiewave.com
Each image follows its distribution's own provisioning model, package manager and security tooling — not one build relabelled several times.
LTS and interim releases, Minimal and Pro variants, built to Canonical's cloud-image conventions.
Stable and oldstable, with backports where a workload needs a newer runtime than the release ships.
The same four steps behind every offer we've published, including the hardening and CIS Benchmark checks every build goes through.
The distribution, licensing model and target marketplaces are agreed before anything is built.
Packer templates, Ansible provisioning and a pinned package set — then hardened, scanned and checked against the CIS Benchmark for its distribution.
Taken through each cloud's own certification pipeline before it goes live on the marketplace.
Rebuilt on the upstream security cadence and re-published, with old versions retired without breaking deployments.
If yours isn't here, ask our marketplace team directly.
Batch and streaming engines, orchestration, query federation and the BI layer that sits in front of them.
A data discovery and metadata catalog platform, originally built at Lyft, for finding and understanding datasets across an organization.
Airflow with the scheduler, webserver and a Celery or Local executor already chosen, backed by PostgreSQL.
A cross-language, in-memory columnar data format and library for fast analytics and data interchange between systems.
Druid with its service roles laid out for a real cluster rather than the single-machine quickstart.
Flink with JobManager and TaskManager roles split, checkpointing configured and state backend pointed at durable storage.
Hadoop with HDFS and YARN configured for single-node or cluster start-up, and the classic port map documented on the box.
Tell us the distribution, the marketplace and the commercial model — we'll build, certify and publish it as a public listing or a private offer.