Cargo Registry
148 cargos available on the Archipelag.io network
148 cargo(s)
Audio Convert
audio-convert
Convert audio between formats (mp3, wav, flac, ogg, aac)
Audio Merge
audio-merge
Merge multiple audio files into one
Audio Normalize
normalize-audio
Normalize audio volume levels
Audio Split
audio-split
Split audio files by silence detection or timestamps
Barcode Generator
barcode
Generate barcodes and QR codes as images
BART-CNN Summarization
onnx-summarize
Summarize long text using BART
BART-MNLI Classification
onnx-classify
Classify text into arbitrary categories
Base64 (WASM)
wasm-base64
Encode and decode Base64 strings
Batch Template
batch-template
Render templates with batch data (Jinja2, Mustache)
BERT Base Fill-Mask
onnx-fill-mask
Predict masked tokens in text using BERT
BERT-NER
onnx-ner
Identify named entities (persons, orgs, locations)
BGE Small English
onnx-bge-small
Compact English text embeddings
BLIP Captioning
onnx-caption
Generate captions for images using BLIP
Chatterbox TTS
onnx-chatterbox-tts
ResembleAI expressive TTS
Code Formatter
format-code
Format source code (Python, JS, Go, Rust, etc.)
CodeGen-350M
onnx-code-complete
Autocomplete code using CodeGen
CogVideoX-2b
diffusers-cogvideo-x-2b
CogVideo text-to-video 2B
Compress (WASM)
wasm-compress
Compress and decompress data (gzip, zstd, brotli)
Cron Parser (WASM)
wasm-cron
Parse and evaluate cron expressions
CSM TTS
onnx-csm-tts
Sesame conversational speech model
CSV Parser (WASM)
wasm-csv
Parse and transform CSV data
Depth Anything V2 (WebGPU)
webgpu-depth-anything
Monocular depth estimation in-browser via WebGPU
DistilBERT QA
onnx-qa
Answer questions given a context passage
DistilBERT Sentiment
onnx-sentiment
Classify text sentiment (positive/negative)
DOCX Extract
docx-extract
Extract text and metadata from DOCX files
DPT-Hybrid Depth
onnx-depth
Estimate depth from monocular images
E5 Large Multilingual
onnx-e5-large-multilingual
Multilingual text embeddings
Echo (WASM)
wasm-echo
Simple echo workload for testing WASM runtime
Face Blur
blur-face
Detect and blur faces in images for privacy
FireRed OCR
onnx-firered-ocr
High-accuracy OCR engine
Fish Audio S2
onnx-fish-audio-s2
Fish Audio speech synthesis
Florence-2 (WebGPU)
webgpu-florence2
Image captioning and visual understanding in-browser via WebGPU
FLUX.1-schnell
diffusers-flux1-schnell
FLUX.1 fast text-to-image
Gemma 3 12B
gguf-gemma-3-12b
Google Gemma 3 12B — high-quality mid-range
Gemma 3 1B
gguf-gemma-3-1b
Google Gemma 3 1B — ultra-light, 128K context
Gemma 3 27B
gguf-gemma-3-27b
Google Gemma 3 27B — large multimodal
Gemma 3 4B
gguf-gemma-3-4b
Google Gemma 3 4B — multimodal, 128K context
Gemma 3n E2B
gguf-gemma-3n-e2b-mobile
Google Gemma 3n — efficient on-device model, multimodal
GIF Maker
gif-make
Create animated GIFs from images or video clips
GLM-4-32B
gguf-glm-4-32b
GLM-4 32B — strong coding and reasoning
GLM OCR
onnx-glm-ocr
Vision-language OCR with GLM
GPT-OSS-20B
gguf-gpt-oss-20b
OpenAI GPT-OSS 20B
Hash (WASM)
wasm-hash
Compute cryptographic hashes (SHA256, SHA512, MD5, BLAKE3)
HTML Sanitize (WASM)
wasm-sanitize
Sanitize HTML to prevent XSS attacks
HTML to PDF
html-to-pdf
Convert HTML content to PDF documents
Image Colorize
colorize
Colorize grayscale images using deep learning
Image Compress
compress-image
Compress images with quality control (JPEG, PNG, WebP)
Image Convert
convert-image
Convert images between formats (PNG, JPEG, WebP, AVIF, TIFF)
Image Diff
image-diff
Compare two images and highlight differences
Image Generation
image-gen
Generate images from text prompts using Stable Diffusion
Image Metadata
image-metadata
Extract EXIF, IPTC, and XMP metadata from images
Image Resize
resize
Resize images with various interpolation methods
IndexTTS-2
onnx-indextts-2
IndexTeam TTS model
Jina Embeddings v5
onnx-jina-embed-v5
Small text embeddings from Jina
JSON to Types
json-to-types
Generate TypeScript/Python/Go types from JSON samples
JSON Transform (WASM)
wasm-json
Validate, format, and transform JSON (jq-like queries)
JWT (WASM)
wasm-jwt
Decode, validate, and inspect JWT tokens
KeyBERT Extraction
onnx-keyword-extract
Extract keywords from text using embeddings
Kokoro TTS
onnx-kokoro-tts
Lightweight high-quality TTS
Kokoro TTS (CoreML)
coreml-kokoro-tts
On-device text-to-speech on iPhone via CoreML
Kokoro TTS (WebGPU)
webgpu-kokoro-tts
Text-to-speech synthesis in-browser via WebGPU
Language Detect
language-detect
Detect the language of input text
Llama 3.1 8B (Exo)
exo-llama-8b
Llama 3.1 8B via Exo runtime — single Island or distributed
Llama-3.1-8B-Instruct
gguf-llama-3.1-8b
Meta Llama 3.1 8B Instruct
LLM Chat (Mistral 7B)
llm-chat
Real LLM inference using Mistral 7B with llama.cpp
LLM Chat (Mock)
llm-chat-mock
Mock LLM for testing - instant responses
Markdown Render (WASM)
wasm-markdown
Render Markdown to HTML with syntax highlighting
Markdown TOC (WASM)
wasm-markdown-toc
Generate table of contents from Markdown headings
Markdown to PDF
markdown-to-pdf
Convert Markdown to styled PDF documents
Minify (WASM)
wasm-minify
Minify HTML, CSS, and JavaScript
MiniLM-L6 Embeddings
onnx-embeddings
Generate text embeddings using MiniLM
Mistral 7B Instruct
gguf-mistral-7b
Mistral 7B Instruct v0.2 (Q4_K_M)
Mistral Small 3.1 24B
gguf-mistral-small-3.1
Mistral Small 3.1 — fast, multimodal, 128K context
MTCNN-FaceNet
onnx-face
Detect faces and generate face embeddings
Neural Style Transfer
onnx-style-transfer
Apply artistic styles to images
Noise Remove
noise-remove
Remove background noise from audio recordings
NuMarkdown
onnx-numarkdown
Convert images/PDFs to Markdown
OCR
ocr
Extract text from images using optical character recognition
OPUS-MT Translation
onnx-translate
Translate text between languages
Parakeet ASR
onnx-parakeet-asr
NVIDIA Parakeet speech recognition
PDF Extract
pdf-extract
Extract text, tables, and images from PDF files
PDF Merge
pdf-merge
Merge multiple PDF files into one document
PDF Sign
pdf-sign
Add digital signatures to PDF documents
Phi-4 Mini
gguf-phi-4-mini
Microsoft Phi-4 Mini 3.8B — reasoning on constrained hardware
QR Code (WASM)
wasm-qrcode
Generate QR codes as SVG or PNG
Qwen3.5-0.8B
gguf-qwen3.5-0.8b
Qwen3.5 0.8B — ultra-light reasoning
Qwen3.5-0.8B (CoreML)
coreml-qwen3.5-0.8b
On-device LLM inference on iPhone via CoreML
Qwen3.5-27B
gguf-qwen3.5-27b
Qwen3.5 27B — top-tier open-source
Qwen3.5-35B-A3B
gguf-qwen3.5-35b-a3b
Qwen3.5 35B MoE (3B active) — efficient large model
Qwen3.5-4B
gguf-qwen3.5-4b
Qwen3.5 4B — compact all-rounder
Qwen3.5-9B
gguf-qwen3.5-9b
Qwen3.5 9B — best mid-range reasoning
Qwen3 8B
gguf-qwen3-8b
Qwen3 8B — strong all-rounder
Qwen3 ASR
onnx-qwen3-asr
Qwen3 automatic speech recognition
Qwen3-Coder-Next
gguf-qwen3-coder
Qwen3 Coder Next — 80B MoE, 3B active, best open-source code model
Qwen3 Embedding 0.6B
onnx-qwen3-embed-0.6b
Small Qwen3 text embeddings
Qwen3 Embedding 8B
onnx-qwen3-embed-8b
Large Qwen3 text embeddings
Qwen3 TTS
onnx-qwen3-tts
Qwen3 text-to-speech with custom voices
Qwen3 VL Embedding
onnx-qwen3-vl-embed
Multimodal vision-language embeddings
Qwen-Image-2512
diffusers-qwen-image
Qwen text-to-image generation
Real-ESRGAN Upscaling
onnx-upscale
Upscale images 4x using Real-ESRGAN
Real-ESRGAN (WebGPU)
webgpu-esrgan
4x image upscaling in-browser via WebGPU
Regex (WASM)
wasm-regex
Test, match, and replace with regular expressions
Screenshot
screenshot
Capture screenshots of web pages
SDXL Turbo (WebGPU)
webgpu-sdxl-turbo
Text-to-image generation running in-browser via WebGPU
Semver (WASM)
wasm-semver
Parse, compare, and validate semantic versions
Slug (WASM)
wasm-slug
Generate URL-safe slugs from text
Smart Crop
crop-smart
Intelligently crop images to focus on subjects
SmolLM2-360M (WebGPU)
webgpu-smollm2
Tiny GPU-accelerated LLM in-browser via WebGPU
Stable Diffusion 1.5 (Diffusers)
diffusers-stable-diffusion
Text-to-image generation (SD 1.5 / SDXL)
Stable Diffusion 3.5 Large
diffusers-sd-3.5-large
Latest SD 3.5 text-to-image
Stable Diffusion (Mock)
sd-mock
Mock image generation for testing
SVG Optimize
svg-optimize
Optimize and minify SVG files
SVG to PNG
svg-to-png
Render SVG files to PNG images at any resolution
Syntax Highlight (WASM)
wasm-highlight
Syntax-highlight source code to HTML
T5 Grammar Correction
onnx-grammar
Correct grammar using T5
T5 Paraphrasing
onnx-paraphrase
Paraphrase text using T5
Tacotron2 TTS
onnx-tts
Classic text-to-speech synthesis
Text Diff (WASM)
wasm-diff
Compute unified diffs between texts
TinyLlama 1.1B
gguf-tinyllama-mobile
On-device LLM inference via llama.cpp
TOML Parser (WASM)
wasm-toml
Parse and convert TOML to JSON
Toxic-BERT
onnx-toxicity
Detect toxic content in text
TrOCR (WebGPU)
webgpu-trocr
Optical character recognition in-browser via WebGPU
U2-Net Background Removal
onnx-remove-bg
Remove backgrounds from images
URL Parser (WASM)
wasm-url
Parse, validate, and manipulate URLs
UUID (WASM)
wasm-uuid
Generate and validate UUIDs (v4, v5, v7)
VibeVoice ASR
onnx-vibevoice-asr
Microsoft VibeVoice speech recognition
VibeVoice Realtime TTS
onnx-vibevoice-realtime
Low-latency streaming TTS
VibeVoice TTS
onnx-vibevoice-tts
Microsoft VibeVoice TTS
Video Clip
video-clip
Extract clips from video files by timestamp
Video Compress
video-compress
Compress video files with quality control
Video Subtitle
video-subtitle
Burn subtitles into video files (SRT, ASS)
Video Thumbnail
video-thumbnail
Generate thumbnail images from video files
Video Transcode
video-transcode
Transcode video between formats (H.264, H.265, VP9, AV1)
Voxtral ASR
onnx-voxtral-asr
Mistral Voxtral speech recognition
Wan2.1-T2V-1.3B
diffusers-wan2.1-t2v-1.3b
Wan-AI text-to-video 1.3B (lightweight)
Wan2.1-T2V-14B
diffusers-wan2.1-t2v-14b
Wan-AI text-to-video 14B
Wan2.2-T2V-A14B
diffusers-wan2.2-t2v-a14b
Wan-AI text-to-video MoE
Watermark
watermark
Add text or image watermarks to images and PDFs
Whisper Base
onnx-whisper
Transcribe audio to text using Whisper
WhisperKit (CoreML)
coreml-whisperkit
On-device speech-to-text on iPhone via CoreML
Whisper Large v3 Turbo
onnx-whisper-large-v3-turbo
Fast large-model transcription
Whisper Small (WebGPU)
webgpu-whisper-small
Speech-to-text transcription in-browser via WebGPU
XLSX Parse
xlsx-parse
Parse Excel spreadsheets to JSON
YAML Parser (WASM)
wasm-yaml
Parse and convert YAML to JSON
YOLOv8 Detection
onnx-detect
Detect objects in images using YOLOv8
YOLOv8 Segmentation
onnx-segment
Segment objects in images using YOLOv8-Seg
YOLOv9 (WebGPU)
webgpu-yolov9
Real-time object detection in-browser via WebGPU
Z-Image-Turbo
diffusers-z-image-turbo
Tongyi Z-Image turbo generation