toolspool

Compare tools

Side-by-side features, use cases and pricing — because the right pick depends on your job and budget, not just the ranking.

Soundify logo
Soundify
✓ verifiedFreemium

AI tool that generates short custom sound effects from a text description, aimed at TikTok, meme and AI-video creators.

2.8K visits/mo2.9K saves
PhotoGov logo
PhotoGov
✓ verifiedFreemium

Online tool converting a selfie into a compliant passport, visa, or ID photo meeting official size rules for 900+ document types.

319K visits/mo
Passport Photo Template logo
Passport Photo Template
✓ verifiedFreemium

AiPassportPhotos makes compliant passport and visa photos online, auto-adjusting size, background and framing from a selfie.

63K visits/mo3.5K saves
Voiser logo
Voiser
✓ verified

Text-to-speech and speech-to-text in 75+ languages.

219K visits/mo14K saves
SpeechGen logo
SpeechGen
✓ verifiedFreemium

Browser TTS generator with thousands of voices in 150 languages for marketers, e-learning teams and IVR builders needing fast voiceovers.

585K visits/mo17K saves
Pricing
Free: $0 (3 free sound effect generations, up to 4s each)
Starter: $9.99 (400 sound effects, 200 generation credits, up to 20s each)
Premium: $29.99 (1800 sound effects, 900 generation credits, up to 20s each)

Free trial available

Monthly subscription: $9.90/month (unlimited photo generation, all formats, no watermark)
US digital photo: $5.90 one-time
US printable A4/PDF photo: $9.90 one-time

No public pricing

No public pricing

Free: 1,000 characters (no account required)
Core features
  • Text-to-sound-effect generation
  • Adjustable clip duration and settings
  • Library of pre-defined prompt examples
  • Downloadable and shareable audio clips
  • Option to make generated sound effects private
  • Automatic sizing and cropping to government-specific ID photo standards
  • Coverage of 900+ document types across roughly 200 countries
  • Optional expert human verification of photo compliance
  • Digital and print-ready (A4, 300 DPI PDF) file downloads
  • One free daily photo in supported regions
  • Sizing/cropping only in the US, with no digital face alteration, per US State Department rules
  • AI passport/visa photo generation
  • Country- and document-specific presets
  • Automatic sizing, framing and background
  • Printable photo templates
  • Related tools (background removal, enhancer, colorizer)
  • Compliance guarantee
  • Text-to-speech conversion in 75+ languages
  • Speech-to-text transcription
  • Voice cloning
  • Online dictation
  • YouTube subtitle generation
  • Talking Website feature
  • Over 5,000 AI voices across 150 languages
  • Adjustable speed, pitch, volume and pause timing
  • SSML controls for intonation and pronunciation
  • Background music mixing
  • Bulk conversion of long documents up to 1M characters
  • Upload of DOCX, PDF or SRT source files
  • Commercial usage license included
  • Multiple export formats and bitrates
Use cases
  • Adding custom sound effects to TikTok or meme videos
  • Creating sound to pair with AI-generated video from Sora or Luma
  • Generating royalty-free sound effects for podcasts or games
  • Preparing a passport or visa photo from home without a photo studio
  • Meeting DV Lottery or visa application photo requirements
  • Getting a compliant ID photo quickly while traveling
  • Businesses or agencies needing to process many ID photos to spec
  • Create passport and visa photos at home
  • Generate ID and document photos for many countries
  • Produce printable, compliant photo templates
  • Creating voiceovers for videos
  • Transcribing audio and video files
  • Adding voice to websites
  • Generating subtitles for YouTube videos
  • Creating audiobooks
  • Developing smart guides for museums and exhibitions
  • Producing marketing or product-explainer voiceovers on tight deadlines
  • Creating multilingual e-learning narration
  • Building bilingual phone/IVR prompts for small businesses
  • Generating narration for audio guides and tours
  • Localizing video content into other languages
Visit
More in Text To Speech