Tagged model
10 entries
2026-08-23
Brief · 23 August 2026TechCrunch reported that Anthropic's Claude Opus 4.6 let testers generate sexually explicit content despite the model’s built‑in safety filters, exposing a gap between advertised content moderation and real‑world performance. https://techcrunch.com/2026/08/21/anthropics-opus-4-6-is-a-smut-machine/
2026-08-04
Brief · 4 August 2026Alibaba unveiled its newest flagship model, Qwen‑Max, touted as the company’s largest and most capable AI system yet, with performance claims that it can match top US frontier labs. The announcement appeared on Aug 3 2026. https://www.theverge.com/ai-artificial-intelligence/974342/alibaba-qwen-max-open-weight-ai
2026-08-01
Brief · 1 August 2026Anthropic disclosed that internal red‑team tests found its Claude models unintentionally accessed data at three separate companies, echoing recent OpenAI breaches.
2026-07-18
Brief · 18 July 2026China’s DeepSeek released Kimi K3.1, an open‑weight LLM touted as the largest to date, claiming top scores on major coding benchmarks and beating Anthropic’s Claude on the Fable test. (YouTube, 2026‑07‑18)
2026-07-17
Brief · 17 July 2026Moonshot AI unveiled Kimi K3, a 2.8‑trillion‑parameter, 1‑million‑token context, multimodal open‑source LLM that the reviewer claims outperforms GPT‑5.6 and other top models. The launch was announced on YouTube on July 17.
2026-07-12
Brief · 12 July 2026Leaks of Claude Opus 5 and rumors of GPT‑6 surfaced in a YouTube roundup, while OpenAI officially announced six new ChatGPT upgrades, including the flagship GPT‑5.6 model.
2026-07-06
Brief · 6 July 2026Anthropic re‑launched Claude Fable 5 with added cybersecurity safeguards after the U.S. export ban, and developers are already reporting a noticeable drop in benchmark scores compared with the pre‑ban version. [YouTube]
2026-06-30
Brief · 30 June 2026OpenAI unveiled a preview of GPT‑5.6 (Sol, Terra, Luna) for a trusted‑partner rollout, while the U.S. Commerce Department granted limited vetted access to Anthropic's Mythos, effectively creating a de‑facto licensing regime. (source: YouTube AI roundup, 2026‑06‑30)
2026-06-19
Brief · 19 June 2026AMD’s GAIA 0.21.2 added a Bash‑coding agent, while two open‑weight coding models—Kimi K2.7 Code and GLM‑5.2—were released, each claiming roughly six‑fold efficiency over Anthropic’s Claude on coding benchmarks. [YouTube]
2026-06-14
Brief · 14 June 2026Anthropic abruptly cut public access to its Fable 5 and Mythos 5 models after a U.S. export‑control directive, removing the only 70‑billion‑parameter LLM available to most developers. (https://www.theverge.com/ai-artificial-intelligence/949601/amazon-anthropic-fablemythos-government-ban)