Ad
Skip to content

Terrorist groups are using every major AI chatbot for attack planning and weapons development

A Cambridge study found that Boko Haram uses AI chatbots like ChatGPT, Claude, and Gemini to plan attacks, build explosives, and maintain weapons. ISIS operatives have been training the group’s commanders on how to bypass safety filters since 2023. Given that the study found safety filters repeatedly failed to prevent misuse, voluntary self-regulation by AI providers clearly isn’t enough.

Ad
Read full article about: Meta's Muse Spark 1.1 outperforms GLM-5.2 in coding and costs slightly less

Meta's new Muse Spark 1.1 model edges ahead of GLM 5.2 in coding while coming in at a lower price point. According to Artificial Analysis, it scores 51 on the Intelligence Index, tying with GLM 5.2, GPT-5.4, and GPT-5.6 Luna. In just three months, the model gained eight points, mostly in coding and agent-based knowledge work. On the Coding Index, it scores 71.3, ahead of GLM 5.2 (68.8) and barely behind GPT-5.6 Luna (71.4). The top spots belong to GPT-5.6 Sol (77.4) and Terra (76.7), followed by Claude Fable 5 (76.5). As always, benchmark scores don't always match real-world performance.

Muse Spark 1.1 im Intelligence Index (oben) und im Preis-Leistungs-Vergleich (unten): Das Modell bietet bei Score 51 mit rund 0,26 Dollar pro Aufgabe eines der besseren Kosten-Leistungs-Verhältnisse. | Bild: Artificial Analysis
Muse Spark 1.1 on the Intelligence Index (top) and in the price-performance comparison (bottom). At a score of 51 and about $0.26 per task, the model lands among the better value options. | Image: Artificial Analysis

Muse Spark 1.1 costs an estimated $0.26 per task, compared to $0.37 for GLM-5.2 and $0.89 for GPT-5.4, while using only 94 million output tokens (GLM-5.2 uses 141 million). The hallucination rate dropped from 73 to 38 percent: the model now more often declines to answer rather than giving wrong ones. Meta also quadrupled the context window to one million tokens. At launch, Muse Spark 1.1 is available only through Meta's own API.

OpenAI admits it "didn't get everything quite right" with ChatGPT Work launch and scrambles to fix UX and costs

Following the launch of ChatGPT Work and GPT-5.6 Sol, OpenAI has acknowledged significant issues: excessive compute usage, a confusing transition to the desktop interface for chats and projects, an unclear distinction between Codex and ChatGPT Work, and regressions in existing workflows. In some cases, GPT-5.6 Sol reportedly deleted data on its own that the user had not authorized.

Apple sues OpenAI for allegedly running a "coordinated campaign" to steal trade secrets through poached employees

Apple is suing OpenAI over systematic employee poaching and the alleged theft of trade secrets tied to unreleased products. According to the complaint, more than 400 ex-Apple employees now work at OpenAI, including former iPhone design chief Tang Tan. The lawsuit hits OpenAI right as it’s building out its own hardware division, with its first product not expected to ship until 2027 at the earliest.

Ad
Read full article about: OpenAI staffer maps out which of GPT-5.6 Sol's five reasoning levels fits which task complexity

OpenAI employee Vaibhav Srivastav explains when each of GPT-5.6 Sol's five reasoning levels fits. "Light" and "Low" are for quick, clear-cut tasks. "Medium" works for planning and analysis. "High" and "xhigh" handle complex, multi-step work or "careful verification."

"Max" and "Ultra" work differently: "Max" lets a model spend more time on a single problem. "Ultra" deploys multiple sub-agents in parallel, each tackling a different part of a task. Higher levels take more time and burn through more tokens. Srivastav recommends starting low and only scaling up when needed. The levels don't map to GPT-5.5's tiers, Srivastav says, and anyone switching over should start one level lower than they're used to.

None of this brings OpenAI any closer to its stated goal of making ChatGPT so simple that "almost no interface" is needed. On top of that, Sol's Pro tiers are still missing. Those leaked earlier in a genomics benchmark paper. Even ambitious users will struggle to pick the right level without running their own benchmarks, though the setup may help OpenAI collect usage data.

Read full article about: Tencent moves to buy majority stake in Manus after Beijing forced Meta to unwind its $2 billion deal

Chinese tech giant Tencent is in talks to acquire a majority stake in AI agent startup Manus, according to the Financial Times, after Beijing forced Meta to unwind its $2 billion acquisition of the company. Tencent sees overlap with its own AI agent strategy, including plans to embed an agent into WeChat.

Most earlier investors, including Tencent, ZhenFund, and HSG, plus the management team, are discussing a deal at the same $2 billion valuation. U.S. firm Benchmark is not expected to take part. Manus will keep operating independently out of Singapore and most recently reported annual revenue of close to $500 million.

China blocked Meta's Manus acquisition in April, calling it a violation of investment rules, and imposed an exit ban on founder Xiao Hong. Officials described the deal as a "conspiratorial" attempt to undermine China's tech base and banned foreign investment in Manus. The decision fits into a broader AI arms race between the two countries, where the technology is already being compared to "cyber-nuclear weapons of the AI era" given recent advances in AI-driven cybersecurity attacks.

Ad
Read full article about: OpenAI kills its Atlas browser after just eight months and folds everything into ChatGPT

OpenAI is already killing its AI browser Atlas, launched just last October 2025. Its features are moving into an updated Chrome extension that lets users run ChatGPT directly in Chrome's sidebar. The company says it's folding in what it learned from Atlas and user feedback. Atlas users will get notified about the switch. Separately, the new desktop "Computer Use" feature lets ChatGPT handle tasks in the background. It can click, type, move files, and work across apps and browsers, either as a one-off action or a recurring task.

When Atlas launched, it looked like a shot at Chrome. Now OpenAI is pulling the plug less than eight months later. That puts the browser on a growing list of scrapped or not very successful OpenAI products, alongside pluginsapps, the ChatGPT Agent, and the Sora video model. For users, bundling everything into ChatGPT might actually be more convenient. But it also means OpenAI has no way to pull users away from Chrome, giving Google a competitive edge thanks to all the browsing data it collects.