AI Tools

318 tools available

AI Detect Lab

AI Detect Lab

AI Detect Lab is a free AI-generated image detection tool for checking whether an image may come from systems such as Midjourney, Stable Diffusion, DALL-E, Flux, or other AI image generators. It is useful for moderation, verification, and editorial review, but results should be treated as signals, not final proof.

AI Detect LabAI image detectorAI generated image detector
2720
Outfit Swap Studio

Outfit Swap Studio

AI virtual try-on / outfit swap for user photos. Generates outfit-changed results while aiming to preserve the original face and background for consistency.

ai-fashionfree
2010
OlmOCR

OlmOCR

A toolkit for training language models to work with PDF documents in the wild.

ocrfree
2560
Umi-OCR

Umi-OCR

Umi-OCR is a free MIT-licensed offline OCR desktop app for Windows and Linux, covering screenshots, batch images, PDFs, QR/barcodes and optional formula plugins. This guide compares Paddle and Rapid builds, privacy, accuracy tests, layout/PDF QA, automation and alternatives.

Umi-OCRoffline OCRfree OCR
2720
AlphaXiv

AlphaXiv

An open academic discussion community based on the arXiv platform that allows users to comment line-by-line, ask questions, and interact in real-time.

ai-researchfree
2690
Chat YouTube

Chat YouTube

Chat YouTube is a lightweight AI tool for summarizing YouTube videos and asking questions about their content. It is useful for students, researchers, and busy viewers who need quick notes from public videos, but output quality depends heavily on transcripts and video clarity.

Chat YouTubeYouTube summaryvideo Q&A
4180
ChatGPT for YouTube

ChatGPT for YouTube

"ChatGPT for YouTube" is a free Chrome Extension that offers instant access to video summaries on YouTube. Quickly grasp video content, save time, and enhance your learning experience.

video-summaryfree
2140
Seamless

Seamless

Seamless is a family of AI models that enable more natural and authentic communication across languages.

speech-translationfree
2710
MuseGen

MuseGen

MuseGen is a credit-based creative studio using models such as Suno for songs, lyrics, vocals, stems, covers, extensions, MIDI, WAV, images, video, and TTS. This review separates marketing claims from legal certainty and covers plans, credit economics, prompting, originality, rights, privacy, mastering, and alternatives.

MuseGenAI music generatorAI song generator
4040
SFX Engine

SFX Engine

Generate unlimited unique sound effects for any project with AI. No experience required. No credit card needed.

ai-musicfree
3490
OptimizerAI

OptimizerAI

Generate unlimited high-quality AI sounds with OptimizerAI

ai-musicfree
2740
Stable Audio

Stable Audio

Stable Audio is Stability AI’s generative audio platform for creating music, loops, stems, ambience, and sound effects from text prompts. The Stable Audio 3.0 family includes models for artistic experimentation, with open-weight options for research and local creative workflows.

Stable AudioStability AI audioAI music generator
2760
AudioCraft

AudioCraft

Open source library for audio/music generation by Meta, which mainly includes two models, MusicGen: text-to-music model, AudioGen: text-generated sound model.

ai-musicfree
2990
Bark

Bark

Bark is a transformer-based text-to-audio model created by Suno. Bark can generate highly realistic, multilingual speech as well as other audio - including music, background noise and simple sound effects.

ai-musicfree
2970
ElevenLabs Sound Effects

ElevenLabs Sound Effects

ElevenLabs Sound Effects is a text-to-sound-effects generator for creating custom SFX, ambience, loops, and cinematic audio from prompts. It supports prompt-based sound design, duration control, prompt influence, looping effects, multiple variations, downloads, and API workflows for video, games, podcasts, and apps.

ElevenLabs Sound EffectsAI sound effectstext to sound effects
3570
Mureka

Mureka

Text to music

ai-musicfree
2910
Udio

Udio

Create music from simple text prompts by specifying topics, genres, and other descriptors which are then transformed into professional quality tracks.

ai-musicfree
2510
Suno AI

Suno AI

Create stunning original music for free in seconds using AI. Make your own masterpieces, share with friends, and discover music from artists worldwide.

ai-musicfree
2460
Lalal.ai

Lalal.ai

Split vocal and instrumental tracks quickly and accurately with LALAL.AI. Upload any audio file and receive high-quality extracted tracks in a few seconds.

voice-processingfree
1810
Vocal Remover

Vocal Remover

Separate voice from music out of a song free with powerful AI algorithms

voice-processingfree
1720
So-VITS-SVC

So-VITS-SVC

So-VITS-SVC 4.1 is an archived open-source research framework for singing voice conversion, not text-to-speech. This independent guide covers consent, licensing, clean datasets, F0 and encoders, training, inference, evaluation, disclosure, security and alternatives.

so-vits-svcsinging voice conversionvoice conversion
3250
Shazam

Shazam

Shazam is Apple’s music recognition app for identifying songs playing nearby or inside other apps. It is useful for listeners, DJs, creators, and marketers who need fast song IDs, lyrics, videos, concert discovery, and Apple Music or playlist follow-up.

Shazammusic recognitionsong identifier
2550
ChatTTS

ChatTTS

ChatTTS is a text-to-speech model designed specifically for dialogue scenario such as LLM assistant. It supports both English and Chinese languages.

text-to-speechfree
2350
Tetos

Tetos

Tetos is an open-source Python and CLI wrapper that provides a unified interface for multiple text-to-speech providers. It is useful for developers who want to compare or switch between Edge TTS, OpenAI, Azure, Google, Volcengine, Baidu, Minimax, Xunfei, Fish Audio, and other engines without rewriting every integration.

Tetostext to speechTTS API
2330
1...678910...14