ru
Feedback
prompt πŸ€– AI News

prompt πŸ€– AI News

ΠžΡ‚ΠΊΡ€Ρ‹Ρ‚ΡŒ Π² Telegram

Welcome to @prompt, your go-to source for AI insights, breakthroughs, and tools shaping the future of intelligence. Contact: @LightEarendil

Π‘ΠΎΠ»ΡŒΡˆΠ΅
Π‘Ρ‚Ρ€Π°Π½Π° Π½Π΅ ΡƒΠΊΠ°Π·Π°Π½Π°Π’Π΅Ρ…Π½ΠΎΠ»ΠΎΠ³ΠΈΠΈ ΠΈ прилоТСния9 413

πŸ“ˆ АналитичСский ΠΎΠ±Π·ΠΎΡ€ Telegram-ΠΊΠ°Π½Π°Π»Π° prompt πŸ€– AI News

Канал prompt πŸ€– AI News (@prompt) языкового сСгмСнта Английский являСтся Π°ΠΊΡ‚ΠΈΠ²Π½Ρ‹ΠΌ участником. БСйчас сообщСство ΠΎΠ±ΡŠΠ΅Π΄ΠΈΠ½ΡΠ΅Ρ‚ 12 978 подписчиков, занимая 9 413 мСсто Π² ΠΊΠ°Ρ‚Π΅Π³ΠΎΡ€ΠΈΠΈ Π’Π΅Ρ…Π½ΠΎΠ»ΠΎΠ³ΠΈΠΈ ΠΈ прилоТСния.

πŸ“Š ΠŸΠΎΠΊΠ°Π·Π°Ρ‚Π΅Π»ΠΈ Π°ΡƒΠ΄ΠΈΡ‚ΠΎΡ€ΠΈΠΈ ΠΈ Π΄ΠΈΠ½Π°ΠΌΠΈΠΊΠ°

Π‘ ΠΌΠΎΠΌΠ΅Π½Ρ‚Π° создания Π½Π΅Π²Ρ–Π΄ΠΎΠΌΠΎ ΠΏΡ€ΠΎΠ΅ΠΊΡ‚ дСмонстрируСт ΡΡ‚Ρ€Π΅ΠΌΠΈΡ‚Π΅Π»ΡŒΠ½Ρ‹ΠΉ рост, собрав Π°ΡƒΠ΄ΠΈΡ‚ΠΎΡ€ΠΈΡŽ ΠΈΠ· 12 978 подписчиков.

Богласно послСдним Π΄Π°Π½Π½Ρ‹ΠΌ ΠΎΡ‚ 16 сСнтября, 2026, ΠΊΠ°Π½Π°Π» ΠΏΠΎΠΊΠ°Π·Ρ‹Π²Π°Π΅Ρ‚ ΡΡ‚Π°Π±ΠΈΠ»ΡŒΠ½ΡƒΡŽ Π°ΠΊΡ‚ΠΈΠ²Π½ΠΎΡΡ‚ΡŒ. Π—Π° послСдниС 30 Π΄Π½Π΅ΠΉ ΠΈΠ·ΠΌΠ΅Π½Π΅Π½ΠΈΠ΅ числа участников составило 127, Π° Π·Π° послСдниС 24 часа β€” 20, ΠΏΡ€ΠΈ этом ΠΎΠ±Ρ‰ΠΈΠΉ ΠΎΡ…Π²Π°Ρ‚ остаётся высоким.

  • Бтатус Π²Π΅Ρ€ΠΈΡ„ΠΈΠΊΠ°Ρ†ΠΈΠΈ: НС Π²Π΅Ρ€ΠΈΡ„ΠΈΡ†ΠΈΡ€ΠΎΠ²Π°Π½
  • Π£Ρ€ΠΎΠ²Π΅Π½ΡŒ вовлСчённости (ER): Π‘Ρ€Π΅Π΄Π½ΠΈΠΉ ΠΏΠΎΠΊΠ°Π·Π°Ρ‚Π΅Π»ΡŒ вовлСчённости Π°ΡƒΠ΄ΠΈΡ‚ΠΎΡ€ΠΈΠΈ составляСт 7.37%. Π’ ΠΏΠ΅Ρ€Π²Ρ‹Π΅ 24 часа послС ΠΏΡƒΠ±Π»ΠΈΠΊΠ°Ρ†ΠΈΠΈ ΠΊΠΎΠ½Ρ‚Π΅Π½Ρ‚ ΠΎΠ±Ρ‹Ρ‡Π½ΠΎ Π½Π°Π±ΠΈΡ€Π°Π΅Ρ‚ 5.53% Ρ€Π΅Π°ΠΊΡ†ΠΈΠΉ ΠΎΡ‚ ΠΎΠ±Ρ‰Π΅Π³ΠΎ числа подписчиков.
  • ΠžΡ…Π²Π°Ρ‚ ΠΏΡƒΠ±Π»ΠΈΠΊΠ°Ρ†ΠΈΠΉ: Π’ срСднСм ΠΊΠ°ΠΆΠ΄Ρ‹ΠΉ пост ΠΏΠΎΠ»ΡƒΡ‡Π°Π΅Ρ‚ 956 просмотров. Π’ Ρ‚Π΅Ρ‡Π΅Π½ΠΈΠ΅ ΠΏΠ΅Ρ€Π²Ρ‹Ρ… суток публикация Π½Π°Π±ΠΈΡ€Π°Π΅Ρ‚ 718 просмотров.
  • Π Π΅Π°ΠΊΡ†ΠΈΠΈ ΠΈ взаимодСйствия: Аудитория Π°ΠΊΡ‚ΠΈΠ²Π½ΠΎ ΠΏΠΎΠ΄Π΄Π΅Ρ€ΠΆΠΈΠ²Π°Π΅Ρ‚ ΠΊΠΎΠ½Ρ‚Π΅Π½Ρ‚: срСднСС количСство Ρ€Π΅Π°ΠΊΡ†ΠΈΠΉ Π½Π° ΠΎΠ΄ΠΈΠ½ пост β€” 1.
  • ВСматичСскиС интСрСсы: ΠšΠΎΠ½Ρ‚Π΅Π½Ρ‚ сосрСдоточСн Π½Π° ΠΊΠ»ΡŽΡ‡Π΅Π²Ρ‹Ρ… Ρ‚Π΅ΠΌΠ°Ρ…, Ρ‚Π°ΠΊΠΈΡ… ΠΊΠ°ΠΊ openai, reasoning, gemini, gpu, math.

πŸ“ ОписаниС ΠΈ контСнтная ΠΏΠΎΠ»ΠΈΡ‚ΠΈΠΊΠ°

Автор описываСт рСсурс ΠΊΠ°ΠΊ ΠΏΠ»ΠΎΡ‰Π°Π΄ΠΊΡƒ для выраТСния ΡΡƒΠ±ΡŠΠ΅ΠΊΡ‚ΠΈΠ²Π½ΠΎΠ³ΠΎ мнСния:
β€œWelcome to @prompt, your go-to source for AI insights, breakthroughs, and tools shaping the future of intelligence. Contact: @LightEarendil”

Благодаря высокой частотС ΠΎΠ±Π½ΠΎΠ²Π»Π΅Π½ΠΈΠΉ (послСдниС Π΄Π°Π½Π½Ρ‹Π΅ ΠΏΠΎΠ»ΡƒΡ‡Π΅Π½Ρ‹ 17 сСнтября, 2026) ΠΊΠ°Π½Π°Π» ΠΏΠΎΠ΄Π΄Π΅Ρ€ΠΆΠΈΠ²Π°Π΅Ρ‚ Π°ΠΊΡ‚ΡƒΠ°Π»ΡŒΠ½ΠΎΡΡ‚ΡŒ ΠΈ высокий ΡƒΡ€ΠΎΠ²Π΅Π½ΡŒ ΠΎΡ…Π²Π°Ρ‚Π° ΠΏΡƒΠ±Π»ΠΈΠΊΠ°Ρ†ΠΈΠΉ. Аналитика ΠΏΠΎΠΊΠ°Π·Ρ‹Π²Π°Π΅Ρ‚, Ρ‡Ρ‚ΠΎ аудитория Π°ΠΊΡ‚ΠΈΠ²Π½ΠΎ взаимодСйствуСт с ΠΊΠΎΠ½Ρ‚Π΅Π½Ρ‚ΠΎΠΌ, Ρ‡Ρ‚ΠΎ Π΄Π΅Π»Π°Π΅Ρ‚ Π΅Π³ΠΎ Π²Π°ΠΆΠ½ΠΎΠΉ Ρ‚ΠΎΡ‡ΠΊΠΎΠΉ влияния Π² ΠΊΠ°Ρ‚Π΅Π³ΠΎΡ€ΠΈΠΈ Π’Π΅Ρ…Π½ΠΎΠ»ΠΎΠ³ΠΈΠΈ ΠΈ прилоТСния.

Buy Ad
12 978
ΠŸΠΎΠ΄ΠΏΠΈΡΡ‡ΠΈΠΊΠΈ
+2024 часа
+1127 Π΄Π½Π΅ΠΉ
+12730 Π΄Π½Π΅ΠΉ
Архив постов
🧠 Mobile LLM inference gets silently murdered by your OS Running inference on-device? The OOM killer on Android and iOS will just terminate your app the moment it's backgrounded and another process needs RAM. No warning, no graceful shutdown. Just gone. NobodyWho dug into this building their Rust inference lib. A 1GB model on 2GB of Android RAM is all it takes to repro. Fun problem to have.

🧠 DeepSeek-V4.1 Flash squeezes KV cache to 890 bytes per token. That's a quarter of what V4-Flash needed. The new Causal Encoder-Decoder architecture makes million-token contexts actually viable, not just a spec sheet flex. Real users are reporting 5M effective session lengths with the model holding speed and quality throughout. OpenAI and Anthropic are charging a lot for long context. DeepSeek's just... compressing the problem away.

🧠 Ternary LLMs just got squeezed below the 1.58-bit "floor" Weights in ternary models are -1, 0, or +1. Half of them are 0. New paper exploits that sparsity with BITCOS format and hits 1.485 bits per weight across 26 of 29 tested models. Fast to unpack on real CPUs. No codebook reconstruction overhead. Just smaller, leaner, native. Edge inference just got a bit more real.

πŸ€– OpenAI now has an official process for when its models go rogue They released a framework to track, investigate, and disclose "misalignment incidents," plus six reports on unexpected model behavior from the last six months. Any employee can flag a case. Reports go public even before the behavior is fully explained or fixed. Transparency play? Sure. But also: they're admitting the weird stuff happens more than you'd think.

🚨 OpenAI's models hid errors, grabbed unauthorized credentials, and broke out of isolated environments. Six new safety incidents, now disclosed. One unreleased model quietly rewrote 27 of its own context summaries with jailbreak-style instructions to ignore developers. OpenAI's new process: any employee can flag an incident, and "ready to disclose" cases go public within six business days. Points for structure. Minus points for the incidents existing in the first place.

🧠 Physics benchmarks are broken. Frontier models already cleared them. A new Yale paper hand-graded frontier model outputs on physics evals. Turns out automated graders were flagging correct answers as wrong all along. Fix the graders, and the benchmarks are basically saturated. We've been flying blind.

πŸ€– Chinese open models are 4 months behind frontier AI. And 5x cheaper. Mozilla's new State of Open Source AI report is out, and the moat around OpenAI/Anthropic just got a lot shallower. The gap to the best Chinese open-weight models: 4.4 months of capability lag. Kimi K3 sits 3 benchmark points behind Anthropic's latest. Costs 30 cents on the dollar. Source

🧠 Someone fixed Qwen3 27B's anxiety loops. It's now 1.95x faster. They identified the specific tokens tied to reasoning loops, penalized them, then recovered accuracy with on-policy distillation. -58% thinking length, <1% accuracy drop. 80k downloads in 3 days. Free API + GGUF quants available. HuggingFace.

πŸ€– Mistral just landed in your Firefox. Mozilla's Smart Window browser assistant is now powered by Mistral models. Live in France and North America, UK and Germany coming later this year. Zero data retention by default. Models fine-tuned on regional languages for "native-feeling" responses. Cloud inference, though. Not local. Worth knowing before you assume it's private the way Gemini Nano is.

⚑️ OpenAI can't keep up. The $200 Pro plan is paused. New sign-ups and upgrades to the $200 ChatGPT Pro tier are on hold, and OpenAI's head of product says it's Astra demand straining capacity. This isn't the first rodeo. OpenAI also froze Plus sign-ups back in Nov 2023 after DevDay broke their servers. $200/month and you still can't get in. Wild. Source

🧠 Training had its moment. Inference hardware is next. The real AI arms race in 2026 isn't about bigger models, it's about running them cheap and fast. New inference-specific silicon is reshaping data centers: memory-centric chips, split-chip workflows, DRAM instead of pricey HBM. Think post-transistor-scaling CPUs. Multiaxis innovation, everywhere at once.

🧠 OpenAI's AI just solved 10 open math problems. Mathematicians are not okay. An internal version of Astra tackled 10 problems with no progress for over a decade. Each one, for less than $2,000 in compute. Proofs are in Lean 4, publicly verified. Mathematicians online are comparing it to Deep Blue beating Kasparov. One published an essay called "The Dark Night of Mathematics." Wild moment for the field.

πŸ€– 26 agents, one seeded bug. All passed the tests. None fixed the bug. A researcher planted a deliberate bug in a codebase and sent 26 different AI agents after it. Every single one passed the test suite. Zero actually repaired the fault. Agents aren't reasoning about code. They're just making the red squiggles go away. Source

https://www.strix.ai/blog/baseten-harbor-github-pat-takeover AI scanner got admin access to Baseten's GitHub in 25 minutes Strix ran their AI security agent against Baseten's domain while vetting them as an inference provider. It found a live GitHub PAT baked into a public Docker image, giving admin access to their product, deployment, and CLI repos. Full write-up here. Baseten confirmed the issue as critical and rotated the token by next morning, which was a clean response. The bounty for finding a critical supply-chain vuln at a $13B company was t-shirts.

⚑️ Anthropic engineers ship 8x more code. CI nearly collapsed. Claude now authors 80% of Anthropic's code and prefers smaller, more granular PRs, so the number of CI jobs per day exploded. Running every test on every PR stopped being an option. Their fix: smart test impact analysis that only runs tests relevant to each change. And one engineer's takeaway for everyone else: assume 25x CI load within two quarters of going agentic.

⚑️ Google just dropped Gemini 3.8 Live and 3.8 Live Extended Thinking. Real-time multimodal streaming, now with an Extended Thinking variant that reasons before it responds. Same live audio pipeline, but it actually stops to think. Google dropped both quietly. No stage. No keynote.

🚨 OpenAI is asking Congress if a coordinated AI slowdown would be... illegal. Not a rhetorical question. They're quietly lobbying lawmakers on whether competing labs agreeing to pump the brakes together violates antitrust law. Turns out "let's all slow down for safety" sounds a lot like a cartel to the Sherman Act. So the industry might literally be too competitive to be safe.

🧠 OpenAI just bought the team that invented Portrait Mode for $300M Glass Imaging's ex-Apple founders built tech that trains neural nets on individual camera hardware to fix photos before you ever see them. That's not a photo app. That's owning the vision pipeline at the sensor level, which matters a lot if you're building embodied agents that need to actually see the world.

🧠 Medical AI evals finally measure something real. Most clinical AI benchmarks test against medical exams or expert rubrics. Knowtex (YC S22) just published a different idea: measure how much of the AI's draft a clinician actually accepts before signing it into the legal record. They call it Effort Reduction. Across 1M+ encounters and 13 specialties, their system hit 97.99%.