Ad
Skip to content
Read full article about: Ling 3.0 Flash is the smartest open model at its size

Ling 3.0 Flash is the smartest open model in its size class. On the Artificial Analysis Intelligence Index, Ling 3.0 Flash scores 38 points, a big jump over its predecessor. That puts it on par with Qwen3.6 27B while using far fewer active parameters. It still trails the leader, DeepSeek V4 Flash, which sits at 52 points. According to Artificial Analysis, it's the smartest open model under 124 billion total parameters. No smaller model matches its score.

Image: Artificial Analysis

On the AA Omniscience test, the hallucination rate dropped from 97 to 44 percent compared to the previous version. The model now refuses to answer questions it doesn't have reliable answers for far more often. Ling 3.0 Flash also shows strong gains in agentic tasks over its predecessor, including on the t3-Bench Banking benchmark.

On per-token pricing, Ling 3.0 Flash beats every comparably capable model. It does burn through more tokens on complex tasks than similarly strong alternatives. But even on a per-task basis, it stays cheaper than Qwen3.6 27B. Ant Group's inclusionAI is releasing the model under an MIT license on the inclusionAI API and through DeepInfra. Model weights are available directly on Hugging Face.

Silicon Valley’s rift over open source pushes back contemplated White House bans on Chinese AI

The Trump administration discussed sanctions and cloud bans targeting Chinese open-weight AI models, according to the New York Times. OpenAI and Anthropic pushed for restrictions, while Nvidia, Google, and Meta fought back. After pushback from Silicon Valley, Washington backed off for now, but a decision is expected before Xi Jinping’s visit in September.

Read full article about: Nvidia's Nemotron 3 Ultra becomes the smartest open US model, but China still leads

According to benchmark platform Artificial Analysis, Nvidia's new Nemotron 3 Ultra is the most capable open AI model from the US to date. It has roughly 550 billion total parameters, with about 55 billion active at any given time. On the Artificial Analysis intelligence ranking, Nemotron 3 Ultra scores 48 points, well ahead of other open US models like Gemma 4 31B (39), Nemotron 3 Super (36), and gpt-oss-120b (33). It doesn't reach the top open models from China, though. Kimi K2.6 scores 54 points there. The current strongest closed model, Opus 4.8, hits 61 points.

Nemotron 3 Ultra lands in the "most attractive quadrant" of the Artificial Analysis chart, combining high intelligence scores with fast output speed. | Image: AAII

On provider DeepInfra, Nemotron 3 Ultra also delivers more than 300 tokens per second, according to Artificial Analysis. Comparably sized models from DeepSeek or Moonshot currently manage only 50 to 100. Nvidia says the model will be released on June 4 on Hugging Face, OpenRouter, and other platforms.

Read full article about: Cohere open-sources its strongest model yet

Canadian AI company Cohere is releasing its most powerful language model, Command A+, as open source under the Apache 2.0 license. The mixture-of-experts model has 218 billion parameters with 25 billion active, and already runs on two Nvidia H100 GPUs or a single Blackwell GPU, according to Cohere.

Command A+ is built for enterprise workflows like agentic tasks, RAG, and multilingual document processing. It supports 48 languages, handles text and images, and has a 128,000-token context window. Compared to its predecessor Command A Reasoning, scores jumped from 37 to 85 percent on the agent benchmark τ²-Bench Telecom and from 3 to 25 percent on the coding test Terminal-Bench Hard.

On the Artificial Analysis Intelligence Index, the model hits just under 37 points - roughly on par with Claude 4.5 Haiku, Gemma 4 31B, and Mistral Medium 3.5. Weights are available on Hugging Face in several quantizations. Cohere recently acquired German AI company Aleph Alpha.

Xiaomi's open-weight MiMo-V2.5-Pro takes aim at Claude Opus with hours-long autonomous coding

Xiaomi’s new MiMo-V2.5-Pro nearly matches Anthropic’s Claude Opus 4.6 on coding benchmarks while burning 40 to 60 percent fewer tokens, according to the company. The release pushes Xiaomi deeper into the race among Chinese open-weight providers like Deepseek, where the fight is shifting from raw benchmark scores to how cheaply and how long a model can run autonomously on a single task.

Arcee AI spent half its venture capital to build an open reasoning model that rivals Claude Opus in agent tasks

US start-up Arcee AI spent roughly half its total venture capital to train Trinity-Large-Thinking, an open reasoning model with 400 billion parameters designed to take on Claude Opus in agent tasks.

Read full article about: Meta plans to open-source parts of its new AI models

Meta is planning to release versions of its new AI models as open source, according to Axios. These would be the first models developed under the leadership of Alexandr Wang, who joined Meta in 2025 as part of a nearly $15 billion deal with Scale AI.

Unlike its approach with the Llama models, though, Meta plans to keep some components proprietary and review safety risks before releasing anything. The largest models won't be made publicly available either.

According to the report, Wang sees Meta as a counterweight to Anthropic and OpenAI, which focus more heavily on government and enterprise customers. Meta's strategy instead centers on consumer reach through WhatsApp, Facebook, and Instagram. Axios's sources say Meta already knows the new models won't match the competition in every area.