Artificial Analysis logo

Artificial Analysis

Artificial Analysis provides independent benchmarks and analysis of AI models and API providers — intelligence, coding and math indices, pricing, latency and throughput performance, plus arena rankings for image, video, speech and music models.

12 actions Integration catalog
Request access
Connect Artificial Analysis once you're in Boring.
01 · WHAT THE AGENT CAN DO

Actions

Every capability is a discrete, logged action the agent calls by name — scoped to what you authorize and recorded in the run trace.

List Image-Editing Models (Arena)ARTIFICIAL_ANALYSIS_LIST_IMAGE_EDITING_MODELS
List image-editing models ranked by the Artificial Analysis arena, with Elo score and 95% confidence interval. Use to find the best-rated image-editing model. Returns all models in one call (unpaginated).
List Image-to-Video-with-Audio Models (Arena)ARTIFICIAL_ANALYSIS_LIST_IMAGE_TO_VIDEO_AUDIO_MODELS
List image-to-video-with-audio models ranked by the Artificial Analysis arena, with Elo score and 95% confidence interval. Returns all models unpaginated in one call. No legacy sibling exists for this modality.
List Image-to-Video Models (Arena)ARTIFICIAL_ANALYSIS_LIST_IMAGE_TO_VIDEO_MODELS
List image-to-video models ranked by the Artificial Analysis image-to-video arena, with Elo score and 95% confidence interval. Returns all models unpaginated in one call.
List Language Models (Intelligence & Pricing)ARTIFICIAL_ANALYSIS_LIST_LANGUAGE_MODELS
List LLMs from the Artificial Analysis leaderboard with intelligence, coding and agentic indices, pricing, and performance (speed and latency). Pass the response's next_cursor back via `cursor` to fetch the next page; pass `start`/`end` to return just a slice of a page (e.g. start=0, end=25 for the top 25).
List Instrumental Music Models (Arena)ARTIFICIAL_ANALYSIS_LIST_MUSIC_INSTRUMENTAL_MODELS
List instrumental music-generation models ranked by the Artificial Analysis arena, with Elo score and 95% confidence interval. Returns all models in one call (unpaginated). Note: music items have no slug field.
List Music-with-Vocals Models (Arena)ARTIFICIAL_ANALYSIS_LIST_MUSIC_WITH_VOCALS_MODELS
List music-with-vocals generation models ranked by the Artificial Analysis arena, with Elo score and 95% confidence interval. Use to find the best-rated music-with-vocals model. Returns all models in one call (unpaginated).
List Speech-to-Speech ModelsARTIFICIAL_ANALYSIS_LIST_SPEECH_TO_SPEECH_MODELS
List speech-to-speech models with Big Bench Audio (bba_score), Full Duplex Bench (fdb_score), and Tau-Voice (tau_voice_score) quality scores. Tau-Voice uses each model's best result across providers. Use to compare voice-to-voice model quality. Returns all models in one call (unpaginated).
List Speech-to-Text Models (WER)ARTIFICIAL_ANALYSIS_LIST_SPEECH_TO_TEXT_MODELS
List speech-to-text (transcription) models with the Artificial Analysis overall word-error-rate index (aa_wer_index). Lower is better. Use to compare transcription accuracy across models. Returns all models in one call (unpaginated).
List Text-to-Image Models (Arena)ARTIFICIAL_ANALYSIS_LIST_TEXT_TO_IMAGE_MODELS
List text-to-image models ranked by the Artificial Analysis image arena, with Elo score and 95% confidence interval. Use to find the best-rated text-to-image model. Returns all models in one call (unpaginated).
List Text-to-Speech Models (Arena)ARTIFICIAL_ANALYSIS_LIST_TEXT_TO_SPEECH_MODELS
List text-to-speech (TTS) models ranked by the Artificial Analysis arena, with Elo score and 95% confidence interval. Use to find the best-rated text-to-speech model. Returns all models in one call (unpaginated).
List Text-to-Video-with-Audio Models (Arena)ARTIFICIAL_ANALYSIS_LIST_TEXT_TO_VIDEO_AUDIO_MODELS
List text-to-video-with-audio models ranked by the Artificial Analysis arena, with Elo score and 95% confidence interval. No legacy sibling exists for this modality.
List Text-to-Video Models (Arena)ARTIFICIAL_ANALYSIS_LIST_TEXT_TO_VIDEO_MODELS
List text-to-video models ranked by the Artificial Analysis arena, with Elo score and 95% confidence interval. Use to find the best-rated text-to-video model. Returns all models in one call (unpaginated).