toolspool

Compare tools

Side-by-side features, use cases and pricing — because the right pick depends on your job and budget, not just the ranking.

WaveSpeedAI logo
WaveSpeedAI
✓ verifiedFree trial

Pay-per-use API hub aggregating 1000+ image, video, and audio generation models for developers building AI media pipelines.

2.2M visits/mo
Fireworks AI logo
Fireworks AI
✓ verifiedPaid

Developer platform for fast serverless inference and training of open generative models, billed per token or GPU-second.

611K visits/mo1.3K saves
Reka Core logo
Reka Core
✓ verifiedPaid

AI research lab building multimodal 'omni' foundation models and infrastructure aimed at robotics and physical-world applications.

252K visits/mo
Kimi Chat logo
Kimi Chat
✓ verifiedFree

Kimi is Moonshot AI's conversational assistant known for long-context chat, coding help, and agentic tasks.

103K visits/mo
Unsloth AI logo
Unsloth AI
✓ verifiedFreemium

Open-source library and desktop app for fast, memory-efficient local fine-tuning and inference of open LLMs.

1.1M visits/mo29K saves
Pricing
Silver: $100 top-up (higher rate limits)
Gold: $1,000 top-up (higher rate limits)
Ultra: $10,000 top-up (highest rate limits)

Free trial available

On-Demand H100/H200: $7/GPU-hour
On-Demand B200: $10/GPU-hour
On-Demand B300: $12/GPU-hour
Fine-tuning (LoRA SFT, models up to 16B): from $0.50 per 1M training tokens

No public pricing

No public pricing

No public pricing

Core features
  • Unified API access to 1000+ image/video/audio generation models
  • Pay-per-use pricing billed per image or per second of video
  • Includes chat/LLM model access (Claude, GPT, Gemini, etc.) priced per token
  • Account tiers unlock higher GPU limits and concurrency
  • CLI and desktop app for building workflows
  • Enterprise options with dedicated support and custom deployment
  • Serverless per-token inference with OpenAI/Anthropic-compatible APIs
  • On-demand dedicated and reserved GPU deployments
  • Fine-tuning and reinforcement-learning training pipelines
  • Large library of open LLM, vision, image and audio models
  • Optimized inference engine for throughput and latency
  • Omni multimodal model research and development
  • Real-time inference API (Infer) for enterprise use
  • Video tagging, search, and clipping infrastructure
  • Training data generation from egocentric and robotics footage
  • Conversational AI assistant
  • Long-context document understanding
  • Coding assistance
  • Agent and plugin capabilities
  • Web and mobile app access
  • Optimized LoRA/FFT/PT training kernels for 500+ models
  • Local offline model runner for Mac and Windows
  • No-code dataset creation from PDFs, CSVs, and JSON
  • Unlimited tool-calling and web search inside model runs
  • Data Recipes workflow to turn documents into training datasets
  • Export to safetensors or GGUF for llama.cpp, vLLM, Ollama
  • Multi-GPU support on paid tiers
Use cases
  • Integrating AI image/video generation into an app via API
  • Building automated content pipelines needing multiple AI models
  • Testing and comparing many generative models from one account
  • Scaling AI media production with volume-based account tiers
  • Accessing both media-generation and LLM APIs from one platform
  • Serving open models in production apps and agents
  • Fine-tuning models on private data
  • Powering code assistants, chatbots and RAG at scale
  • Powering robotics perception with multimodal AI
  • Running large-scale video search and analysis via API
  • Sourcing specialized training data for frontier AI models
  • Answering questions and research
  • Summarizing long documents
  • Writing and editing help
  • Coding support
  • ML engineers fine-tuning open models on a single GPU for free
  • Teams building custom datasets from unstructured documents
  • Developers wanting to run and compare LLMs fully offline
  • Enterprises needing faster, more accurate multi-node training
Visit
More in Llms Foundation Models