Compare tools
Side-by-side features, use cases and pricing — because the right pick depends on your job and budget, not just the ranking.
⇄ Comparison dimension — pick the market you're actually shopping in
Free all-in-one desktop AI app to chat with your documents and run RAG and AI agents fully local and private.
End-to-end computer vision platform for teams annotating data, training YOLO models, and deploying them at scale.
Pay-as-you-go API aggregating thousands of image, video, audio and LLM models with custom inference hardware for lower per-request cost.
TypeScript backend-as-a-service with a reactive database, server functions, auth and file storage for full-stack and AI apps.
Wafer-scale AI hardware and inference cloud delivering record-fast, low-latency inference for open and frontier models.
No public pricing
Free trial available
- ✦Chat with your documents (RAG)
- ✦Runs locally and offline for privacy
- ✦Supports any LLM (local or cloud)
- ✦Built-in AI agents
- ✦Handles PDFs, Word, CSV, codebases
- ✦No-code setup
- ✦Smart data annotation with SAM-powered one-click masks across six task types
- ✦Cloud training with 22+ GPU configurations from RTX 2000 Ada to B200
- ✦Support for YOLOv5 through YOLO26 model families
- ✦One-click deployment across 43 global regions with auto-scaling
- ✦Export to 18 formats including ONNX, TensorRT, and CoreML
- ✦Live training metrics and experiment comparison dashboard
- ✦Single API for image, video, audio, 3D and LLM models
- ✦Standardized model addressing across hosted, partner and custom uploads
- ✦Support for LoRAs, ControlNets, VAEs and embeddings on open-source models
- ✦WebSocket and REST access with async webhook delivery
- ✦Pay-per-request billing with no infrastructure to manage
- ✦Raw serverless GPU/CPU compute for custom workloads
- ✦Reactive real-time database
- ✦TypeScript server functions (queries/mutations/actions)
- ✦Built-in authentication
- ✦Cron jobs and backend workflows
- ✦File storage, text and vector search
- ✦ACID transactions; open-source/self-host
- ✦Wafer-Scale Engine AI processor
- ✦High-speed inference API (OpenAI-compatible)
- ✦Cloud, on-prem and on-device deployment
- ✦Support for GLM, Qwen, Llama, GPT-OSS and more
- ✦Fine-tuning and training on one platform
- ✦Partner access via AWS, OpenRouter, HuggingFace, Vercel
- →Privately querying your own documents
- →Running local AI without the cloud
- →Building AI agents over your data
- →Using multiple LLM providers in one app
- →Building and training custom object detection or segmentation models
- →Labeling large image/video datasets for computer vision projects
- →Deploying vision models to edge or mobile devices
- →Running quality control or defect detection in manufacturing
- →Powering retail, logistics, or agriculture vision applications
- →Adding AI image or video generation to an app without managing infra
- →Batching multi-modal generation tasks in one API call
- →Running custom fine-tuned models via Model Upload
- →Cutting inference costs at high generation volume
- →Building real-time reactive apps
- →Backends for AI agents
- →Replacing Firebase or Supabase
- →Full-stack TypeScript development
- →Low-latency inference for agents and copilots
- →Real-time voice and reasoning apps
- →Fine-tuning and serving custom models
More in AI Agent Local LLM Runner▶
No more tools in this category.