Artificial Analysis
Artificial Analysis provides independent benchmarks and analysis of AI models and API providers — intelligence, coding and math indices, pricing, latency and throughput performance, plus arena rankings for image, video, speech and music models.
Tools (12)
List Image-Editing Models (Arena)
List image-editing models ranked by the Artificial Analysis arena, with Elo score and 95% confidence interval. Use to find the best-rated image-editing model. Returns all models in one call (unpaginated).
Inputs: limit
List Image-to-Video-with-Audio Models (Arena)
List image-to-video-with-audio models ranked by the Artificial Analysis arena, with Elo score and 95% confidence interval. Returns all models unpaginated in one call. No legacy sibling exists for this modality.
Inputs: limit
List Image-to-Video Models (Arena)
List image-to-video models ranked by the Artificial Analysis image-to-video arena, with Elo score and 95% confidence interval. Returns all models unpaginated in one call.
Inputs: limit
List Language Models (Intelligence & Pricing)
List LLMs from the Artificial Analysis leaderboard with intelligence, coding and agentic indices, pricing, and performance (speed and latency). Pass the response's next_cursor back via `cursor` to fetch the next page; pass `start`/`end` to return just a slice of a page (e.g. start=0, end=25 for the top 25).
Inputs: end, start, cursor
List Instrumental Music Models (Arena)
List instrumental music-generation models ranked by the Artificial Analysis arena, with Elo score and 95% confidence interval. Returns all models in one call (unpaginated). Note: music items have no slug field.
Inputs: limit
List Music-with-Vocals Models (Arena)
List music-with-vocals generation models ranked by the Artificial Analysis arena, with Elo score and 95% confidence interval. Use to find the best-rated music-with-vocals model. Returns all models in one call (unpaginated).
Inputs: limit
List Speech-to-Speech Models
List speech-to-speech models with Big Bench Audio (bba_score), Full Duplex Bench (fdb_score), and Tau-Voice (tau_voice_score) quality scores. Tau-Voice uses each model's best result across providers. Use to compare voice-to-voice model quality. Returns all models in one call (unpaginated).
Inputs: limit
List Speech-to-Text Models (WER)
List speech-to-text (transcription) models with the Artificial Analysis overall word-error-rate index (aa_wer_index). Lower is better. Use to compare transcription accuracy across models. Returns all models in one call (unpaginated).
Inputs: limit
List Text-to-Image Models (Arena)
List text-to-image models ranked by the Artificial Analysis image arena, with Elo score and 95% confidence interval. Use to find the best-rated text-to-image model. Returns all models in one call (unpaginated).
Inputs: limit
List Text-to-Speech Models (Arena)
List text-to-speech (TTS) models ranked by the Artificial Analysis arena, with Elo score and 95% confidence interval. Use to find the best-rated text-to-speech model. Returns all models in one call (unpaginated).
Inputs: limit
List Text-to-Video-with-Audio Models (Arena)
List text-to-video-with-audio models ranked by the Artificial Analysis arena, with Elo score and 95% confidence interval. No legacy sibling exists for this modality.
Inputs: limit
List Text-to-Video Models (Arena)
List text-to-video models ranked by the Artificial Analysis arena, with Elo score and 95% confidence interval. Use to find the best-rated text-to-video model. Returns all models in one call (unpaginated).
Inputs: limit