Please confirm you are human
This browser or connection looks automated. Press and continuously hold the control for 3 seconds to enable Google-hosted web results and, when separately allowed, AI-assisted answers.
A successful check enables 100 search requests. Interactive access does not authorize scraping, systematic collection, or reuse of search output.
News
Launching v4.1.1 of the Artificial Analysis Intelligence Index
7+ hour, 54+ min ago (277+ words) We have updated the Artificial Analysis Intelligence Index to v4.1.1 - this patch release upgrades our grader models, and brings the latest 𝜏³-Banking version to Artificial Analysis To keep the Artificial Analysis Intelligence Index the most useful synthesis metric for developers, we…...
Muse Spark 1.2
1+ day, 3+ hour ago (363+ words) Meta’s Muse Spark 1.2 scores 54 on the Artificial Analysis Intelligence Index. Its Meta's third release in four months, significantly improving agentic knowledge work capabilities over prior releases and putting Meta next to SpaceXAI in a tie for third place amongst US…...
Product Manager (AI Hardware)
3+ day, 2+ hour ago (218+ words) Our benchmarks and analysis are trusted by hundreds of thousands of users and are the go-to reference for leading AI labs including OpenAI, Google, Meta, NVIDIA and Anthropic, and major publications including the Wall Street Journal, Bloomberg, the Financial Times…...
DeepSeek V4 Flash 0731 scores 50 on the Artificial Analysis Intelligence Index, 10 points above previous DeepSeek V4 Flash
6+ day, 15+ hour ago (315+ words) DeepSeek V4 Flash 0731 retains a 1M token context window, and its size remains unchanged from DeepSeek V4 Flash at 284B total parameters and 13B active at inference time ➤ DeepSeek V4 Flash 0731 improves over its predecessor on every evaluation in the Intelligence Index: Alongside the agentic gains,…...
Compare | MicroEval
6+ day, 19+ hour ago (57+ words) Compare Artificial Analysis - 1.Complete the following Python function: ```python from typing import List def has_close_elements(numbers: List[float], threshold: float) -> bool: """ Check if in given list of numbers, are any two numbers closer to each other than given threshold. >>> has_close_elements([1.0, 2.0, 3.0], 0.5) False >>> has_close_elements([1.0, 2.8, 3.0, 4.0, 5.0, 2.0], 0.3) True…...
Which one is the best premium AI for book recommendations, c... | MicroEval
6+ day, 19+ hour ago (52+ words) Which one is the best premium AI for book recommendations, c... Artificial Analysis Which one is the best premium AI for book recommendations, complex political and economic discussions, and personal projects/tasks relative to its price? Take the intelligence into account,…...
The Most AI Agent | MicroEval
6+ day, 19+ hour ago (14+ words) The Most AI Agent Artificial Analysis The Most AI Agent © 2026 Artificial Analysis...
bedrock comparison | MicroEval
6+ day, 19+ hour ago (22+ words) bedrock comparison Artificial Analysis Rank the models in terms of general coding capability for medium to complex tasks © 2026 Artificial Analysis...
Doomscroll - An Endless Feed of AI Model Charts
1+ week, 6+ hour ago (18+ words) Doomscroll Artificial Analysis An endless, randomly shuffled feed of charts from across Artificial Analysis. © 2026 Artificial Analysis...
Agnes AI releases Agnes 2.5 Pro Alpha
1+ week, 1+ day ago (513+ words) Agnes AI is a Singapore-based AI lab that trains its own full-modality foundation models in-house across text, image, and video, and offers them through a free omni-modal API that has passed 3 million users. Agnes 2.5 Pro Alpha is a text, image,…...