🚨 AI News | TestingCatalog
Открыть в Telegram
Latest AI News on AI Agents, Model Releases, Tools, Leaks, and Rumors 🗞
БольшеСтрана не указанаТехнологии и приложения14 229
7 542
Подписчики
+224 часа
+377 дней
+22830 дней
Архив постов
GOOGLE 🔥: Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking are rolling out on APIs, Google AI Studio, Gemini Live, and other apps.
Gemini 3.8 Live Extended Thinking (High) is taking the 1st spot on most benchmarks, overtaking GPT Live 1.
NoimosAI launches new Customer Engagement Agent
NoimosAI is adding customer messaging to its autonomous marketing platform, with an agent that builds lifecycle email flows from plain-language prompts, drafts branded messages, supports review before sending, and tracks product goals.
🗞 #sponsored @testingcatalog
+1
OPENAI 🔥: A big “ship” week has been declared, teasing new models from GPT-6 family.
GPT-6 Sol is expected as well as smaller GPT-6 models.
The next generation model is reportedly “slowed down” and it is yet unclear if we will see it in September.What else do we expect? 👀
Claude Code Hub? Users will be able to see all their sessions across various places in Claude Desktop.
> Model, effort, and permission mode selectors are configurable for each.
> Both local and cloud sessions are listed there so it is easy to switch between them.
DAILY AI BRIEF 🗞 — Sept 15
AI SLOWDOWN 🔥:
> Dario Amodei: pace the frontier, not halt it — embedded evaluators, lab coordination, then a global deal.
> Elon, All-In: peer-review rival models before release. “Dario is right.”
> Sam Altman agrees, will match evaluators, and says pacing means slower, not stopped.
> Demis Hassabis: Amodei’s essay “points towards the right path,” details TBD, plus DeepMind’s standards-body plan.
> Satya Nadella welcomes deliberate pacing and embedded evaluators; superintelligence without human control is “not worth pursuing.”
> Alexandr Wang: alignment is fundamental and may gate scaling near the frontier, with no join on a coordinated slowdown.
> Trump rejects a slowdown: takeover talk is a “hoax,” control is a “strong and smart president,” and a “sick conspiracy” against AI only helps China.
> China MFA’s Guo Jiakun: slowdown talk is “fear-mongering”; confrontation “serves no one’s interest.”
MODEL FORECAST 🔥:
- Today? Gemini 3.8 Live and Live Extended Thinking slugs appeared on the Cloud quota page; not public.
- This week? Grok 4.7 should land near Opus 5.0, not 5.1; multimodal still needs work.
- This week? Testers say some Claude Code Opus 5 traffic is routing to Opus 5.2 — faster and less lazy, still unofficial.
- Soon: Grok 4.8 is a 2.5T model on a new C++ stack; training wraps this week, then RL.
- Not soon: Grok 4.9 is “probably Astra/Fable class”; Grok 5 is “maybe better than anything.”
PERPLEXITY 🔥:
- Portable Computer is live on Windows PCs with NVIDIA RTX GPUs and 24GB+ VRAM, including local MCP and scheduled tasks.
OPENAI 🔥:
- Bought Glass Imaging for $300M, per WSJ — an ex-Apple smartphone-camera startup, after the earlier Opal webcam stake.
- Codex — got an official Arch Linux installer with pacman updates.
GOOGLE 🔥:
- Interactive Reports incoming on Gemini Notebook: embed mind maps, slide decks, flashcards, and quizzes.
- Gemini Live can run Deep Research in the background and ping you when the report is ready.
- API slug antigravity-preview-09-2026 spotted: hosted Linux sandbox agent, expected to be powered by Gemini 3.8 Flash.
P.S. All predictions in the Model Forecast section are unconfirmed and may change anytime.
P.S.S. We've used Grok to compose this brief, cherry-picking the news and doing some post-editing.
+1
GOOGLE 🔥: Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking model names have started appearing on the GCP Console quotas page.
A new Gemini Live model, based on the latest Gemini 3.8 Flash, is expected to be a big leap.
> So far, Gemini 3.1 Live Preview is the latest Gemini Live model available via APIs.
> Recently, Google updated Gemini Live with support for Connector calls and the possibility of triggering Deep Research.
> OpenAI also released the GPT Live 1 model on the APIs last week, and it seems like Google has a response to that.
* Discovered by Bedros Pamboukian
+1
OpenAI ❤️ Glass Imaging
OpenAI acquired Glass Imaging for $300 million, according to WSJ.
Glass Imaging develops high quality smartphone cameras and was founded by former Apple employees.
Earlier this year, OpenAI invested in Opal, a maker of premium webcams.
ElevenLabs MCP can now generate voice, music, images & video
ElevenLabs added ElevenCreative to its hosted MCP server, letting Claude, ChatGPT, Cursor and coding agents create audio, images and video in-chat via one OAuth sign-in, alongside existing agent management tools.
🗞 #elevenlabs @testingcatalog
PERPLEXITY 🔥: Windows users with NVIDIA RTX GPUs can now use Portable Computer powered by local models!
> Support for local MCPs and scheduled tasks has also been added.
> Earlier, Portable Computer was introduced for NVIDIA DGX Spark devices.
> Now, Portable Computer is also available on Windows PCs with a supported NVIDIA RTX GPU with 24GB of VRAM or higher.
Models available locally 👀
- PPLX 27B (Perplexity post-trained Qwen 3.8 27B)
- Qwen 3.8 27B (stock) is not available on Windows
- Nemotron 3.5 Lightning (~30B MoE, ~3B active) is marked as "Coming Soon"
GOOGLE 🔥: Interactive Reports for Gemini Notebook will allow users to embed Mind Maps, Slide Decks, Flash Cards and Quizzes right inside the report.
At first, the system will generate a report with placeholders so users can explicitly create the artifacts they want embedded.
Early look at Interactive Reports on Gemini Notebook
Google appears close to launching Interactive Reports in Gemini Notebook: longer reports with embedded mind maps, slides, quizzes, and flashcards generated on demand, combining research, teaching, and explainer formats in one document.
🗞 #google @testingcatalog
MICROSOFT 🔥: A draft of the "Code of Conduct" has been published, introducing the term "Humanist Superintelligence" (HSI).
Key principles 👀
- People matter more than AI.
- AI must remain under meaningful human control.
- Safety takes priority over task completion.
- AI should remain a tool, not imitate a person.
- AI must not pursue independent goals.
- AI must stay within its authorized scope.
- Users must retain control over consequential decisions.
- AI should strengthen human reasoning and autonomy.
- Models must be accurate, transparent, and honest.
- AI must acknowledge uncertainty and correct mistakes.
- Models must not manipulate or exploit users.
- AI must respect personal and emotional boundaries.
- AI should support, not replace, human relationships.
- Models should discourage emotional dependence on AI.
- AI should respect cultural and personal differences.
- Human dignity and fundamental rights must be protected.
- AI should serve the public interest.
- Models should remain politically neutral in elections.
- AI actions should be traceable and understandable.
- Tool use should be authorized, limited, and reversible.
Anthropic prepares Claude Money for personal finance
Anthropic is testing a Claude “Money” tab in its mobile app that would let users link bank accounts and ask spending and budgeting questions. The unreleased feature points to a possible US-first launch, though timing remains unclear.
🗞 #anthropic @testingcatalog
+1
SPACEXAI 🔥: Grok 4.8 will be a 2.5T-parameter model built on a new C++ software stack, and Elon Musk expects it to finish training this week.
> While Grok 4.8 is in training, Grok 4.7 is still expected to arrive shortly, factoring in a previously communicated delay.
> Grok 4.8 will be ±67% larger than Grok 4.6, which is currently available. This size puts it into the same tier as Kimi K3 with 2.8T params.
Not very soon 👀
+1
MICROSOFT 🔥: Satya Nadella agrees with pacing frontier AI development.
> “Superintelligence that doesn’t benefit humanity and is not under human control doesn't worth pursuing”
> “We welcome the research, focus, and deliberate pacing needed to get alignment right as the design goal.”
+1
GOOGLE 🔥: Demis Hassabis shared that he is aligned with the direction outlined by Dario Amodei for pacing the frontier.
RSI moment? 👀
ICYMI: Cursor announced Projects for agent coordination
Cursor Projects, now in beta, coordinates long-running software work through delegated agents. It runs tasks in cloud or local environments, supports scheduled and PR-driven maintenance, and targets work spanning multiple PRs.
🗞 #cursor @testingcatalog
This is a "Defender's gap" chart that OpenAI published recently. It shows the gap between defenders' capabilities and attackers' capabilities from a cybersecurity POV. This also translates to a gap between proprietary and open AI models.
What Dario is proposing is closely related:
"Thus, a key part of pacing within democracies is to keep democracies’ AI lead over autocracies as large as possible, to give us the breathing room we need in order to pace effectively."In other words, Anthropic and OpenAI want to widen the "gap" between what their AI can do and what the rest of the world can do. This doesn't necessarily mean that they will stop AI development and training. What this leads to is: - Anthropic, OpenAI, and other frontier labs will need to work together to make sure that every lab maintains alignment standards. - These labs will continue using RSI to advance their internal models with "employee-like access". - These models WON'T be released to the public until the "gap" is sufficient and until alignment standards are met. What about China?
> "Distillation of frontier models allows lagging companies to narrow the gap using a fraction of the cost it would take to develop their own AI independently."
> "If we execute these measures well, I believe they would slow China’s progress enough to widen America’s lead significantly over the next 3–5 years — the window when AI becomes geopolitically most important."The assumption behind these measures is simple: without being able to distill frontier models, it will take China significantly longer to close the gap with top-tier models. All the above may fay fail. China may or may not take the lead in AI progress. Yet, it's not a surprise that AI can already be used as a cybersecurity weapon. Note that the top point on the "frontier" line describes defenders' capabilities available to companies with Daybreak access and similar. However, the top point on the "open-weight" line is accessible to everyone. Even with the current level of intelligence, we will start seeing more and more big security incidents happening around the globe.
> “Everything that makes it successful is exactly what makes it dangerous.”We should be monitoring this very closely 👀
+1
OPENAI 🔥: Sam Altman agrees with Dario Amodei on his proposal to pace frontier AI development.
Google next? Will we see any statement from Chinese frontier labs as well?
AI weekend unfolds 🤖
+1
SPACEXAI 🔥: Elon Musk agrees with the statement published by Dario Amodei, proposing to pace frontier AI development.
Looks like all this will have real consequences very soon. Nothing unexpected tho.
