Website profile

Artificial Analysis

Comparison and analysis of AI models and API hosting providers. Independent benchmarks across key performance metrics including quality, price, output speed & latency.

  • 16articles · 30d
  • 3+ day agolatest article
  • Aug 20, 2026earliest in window
  • 63%with images
  • 262avg words
articles per day
Categories
  • Science & Technology 16
  • Computers & Electronics 7
  • Software Dev. 7
  • Education & Jobs 5
  • Business & Industrial 3
  • Science & Nature 3
  • Jobs & Education 2
  • STEM 2

Please confirm you are human

This browser or connection looks automated. Press and continuously hold the control for 3 seconds to enable Google-hosted web results and, when separately allowed, AI-assisted answers.

A successful check enables 100 search requests. Interactive access does not authorize scraping, systematic collection, or reuse of search output.

Hold with a pointer, or hold Space or Enter.

News


artificialanalysis.ai > articles > benchmarking-gpt-6-astra

Benchmarking GPT-6 Astra

4+ day, 14+ hour ago   (757+ words) GPT-6 Astra ties leadership with Claude Fable 5.1 in both of our flagship Indices, at lower cost. Astra equals Fable 5.1 in the Intelligence Index at ~40% of the cost, and in the Coding Agent Index at ~60% of the cost. Pricing is 2.5x GPT…...


artificialanalysis.ai > articles > openbmb-releases-minicpm5-2b

OpenBMB releases MiniCPM5-2B

1+ week, 3+ hour ago   (446+ words) OpenBMB is the open-source AI group behind the MiniCPM series of efficient small models. MiniCPM5-2B is a 2.6B parameter dense reasoning model with text input and output, released under Apache 2.0. Scoring 15 on the Intelligence Index, MiniCPM5-2B sits one point behind Ling 3.0 Tiny (16), which…...


artificialanalysis.ai > models > releases

Language Model Releases

1+ week, 4+ day ago   (741+ words) All reasoning and effort levels grouped together by model....


artificialanalysis.ai > microevals > convert-the-below-paragraph-into-an-ai-custom-instruction-fo-1786823953499

convert the below paragraph into an AI custom instruction fo... | MicroEval

1+ week, 1+ day ago   (93+ words) convert the below paragraph into an AI custom instruction fo... Artificial Analysis - 1.convert the below paragraph into an AI custom instruction for gemini ai which will follow this in each chat page Always reply in English only — never Tamil or…...


artificialanalysis.ai > articles > artificial-analysis-intelligence-index-v4-2

Announcing Artificial Analysis Intelligence Index v4.2

1+ week, 2+ day ago   (557+ words) We are accelerating elements of our upcoming v5 release with interim updates to keep pace with the frontier. Index v4.2 has more complex and realistic tasks, and more private test sets to prevent gaming + AA-Briefcase, our agentic knowledge work evaluation with a…...


artificialanalysis.ai > articles > muse-spark-1-3

Muse Spark 1.3: Meta reaches the frontier

1+ week, 4+ day ago   (236+ words) Muse Spark 1.3 (xhigh) enters the Artificial Analysis Intelligence Index at 61, up 4 points from Muse Spark 1.2 (57, August) and 8 points from Muse Spark 1.1 (53, July). It enters tied with GPT-5.6 Sol (max), Grok 4.6 (high), and Claude Opus 5 (high), and behind Claude Fable 5.1 (max,…...


artificialanalysis.ai > microevals > i-want-to-create-a-small-mes-system-for-my-production-compan-1786864132605

I want to create a small MES System for my production compan... | MicroEval

2+ week, 13+ hour ago   (30+ words) I want to create a small MES System for my production compan... Artificial Analysis I want to create a small MES System for my production company (eyewear). © 2026 Artificial Analysis...


artificialanalysis.ai > microevals > you-are-an-administrative-operations-lead-in-a-government-de-1787769054768

You are an administrative operations lead in a government de... | MicroEval

2+ week, 1+ day ago   (304+ words) You are an administrative operations lead in a government de... Artificial Analysis [+1] Addresses findings, risks, impacts or considerations identified in Key Findings in Implications for Government. [+1] Gives bullet points to each academic article’s Implications for Government. [+1] Gives bullet points of…...


artificialanalysis.ai > articles > mobile-phone-intelligence-inference

Intelligence at pocket scale: Benchmarking small models and mobile phones

3+ week, 1+ hour ago   (559+ words) Models under a few billion parameters can now follow instructions, call tools, and answer questions on mobile phones. But available benchmark results often describe neither the quantized model artifact nor the device and runtime combination a user will run. Some…...


artificialanalysis.ai > careers > forward-deployed-engineer-language-models

Forward Deployed Engineer - Language Models

3+ week, 3+ day ago   (186+ words) Our benchmarks and analysis are trusted by hundreds of thousands of users and are the go-to reference for leading AI labs including OpenAI, Google, Meta, NVIDIA and Anthropic, and major publications including the Wall Street Journal, Bloomberg, the Financial Times…...