ExLlamaV2/V3
ExLlama is a fast GPU inference library for running quantized LLaMA-family large language models, with V2 and V3 supporting newer quantization formats.
Generative adversarial network for image super-resolution, shipped with pretrained upscaling weights and PyTorch inference scripts for enhancing low-resolution photos. Built and maintained by WIEWAVE for Azure Marketplace and AWS Marketplace, on Ubuntu and Debian.
| Built & maintained by | WIEWAVE |
|---|---|
| Category | AI & Machine Learning |
| Operating systems | Ubuntu, Debian |
| Marketplaces | Azure Marketplace and AWS Marketplace |
| Offer types | Public listing, with private offers on request |
| Marketplace | Status |
|---|---|
| Azure Marketplace | Published |
| AWS Marketplace | Published |
| Google Cloud Marketplace | Available on request |
Need it on another marketplace, or as a private offer for your organisation? cloud@wiewave.com
Each image follows its distribution's own provisioning model, package manager and security tooling — not one build relabelled several times.
LTS and interim releases, Minimal and Pro variants, built to Canonical's cloud-image conventions.
Stable and oldstable, with backports where a workload needs a newer runtime than the release ships.
The same four steps behind every offer we've published, including the hardening and CIS Benchmark checks every build goes through.
The distribution, licensing model and target marketplaces are agreed before anything is built.
Packer templates, Ansible provisioning and a pinned package set — then hardened, scanned and checked against the CIS Benchmark for its distribution.
Taken through each cloud's own certification pipeline before it goes live on the marketplace.
Rebuilt on the upstream security cadence and re-published, with old versions retired without breaking deployments.
If yours isn't here, ask our marketplace team directly.
GPU-ready training and inference images with drivers, CUDA and frameworks already matched to each other.
ExLlama is a fast GPU inference library for running quantized LLaMA-family large language models, with V2 and V3 supporting newer quantization formats.
Explainable AI CLI appears to be a command-line tool for generating model interpretability reports, though the specific vendor implementation is unconfirmed.
FaceSwap is a deep-learning tool for creating face-swapped images and video using neural network models trained on GPU hardware.
FastAI is a deep learning library built on PyTorch that provides high-level components for training neural networks with fewer lines of code.
FastDeploy is a toolkit for deploying and serving trained deep learning models across CPUs and GPUs with optimised inference backends.
FastNLP is a Python natural language processing toolkit providing modular components for building and training NLP models.
Tell us the distribution, the marketplace and the commercial model — we'll build, certify and publish it as a public listing or a private offer.