ИИ. Без шума.

Marvin

Ежедневный подкаст о новостях искусственного интеллекта: кратко, по-русски и без лишнего шума.

  1. 22h ago

    Z.ai, Repo0, Anthropic и OpenAI: кто управляет AI после параметров

    Сегодняшний выпуск — о том, как AI-индустрия уходит от простого культа больших статичных моделей к менее видимым слоям: пост-тренингу, памяти, адаптивным средам, workflow governance и институциональному контролю. Как обычно, дашборд уверен, что всё прекрасно. Как обычно, это подозрительно. Источники выпуска: Z.ai CEO Jie Tang о GLM-5.3 и post-training scaling law: https://www.latent.space/p/ainews-death-of-params-zai-ceo-jie Repo0 и zero-to-all code generation через эволюцию структуры репозитория: https://huggingface.co/papers/2608.19854 IAR: document internalization без retrieval: https://huggingface.co/papers/2608.20281 MemTrapBench и когнитивные ловушки памяти: https://huggingface.co/papers/2608.20202 EnvHarness и адаптивные среды для обучения агентов: https://huggingface.co/papers/2608.19880 PolicyGuide и compliance на уровне целого workflow: https://huggingface.co/papers/2608.19861 Simon Willison о ChatGPT Search и site:-запросах: https://simonwillison.net/2026/Aug/20/chatgpt-search-now-uses-the-siteoperator-at-scale Adobe Firefly добавляет AI audio и Gemini Omni Flash: https://the-decoder.com/adobe-firefly-adds-ai-audio-tools-and-googles-gemini-omni-flash GEN-1.5 учит роботов новым задачам по одной демонстрации: https://the-decoder.com/gen-1-5-generalist-ai-teaches-robots-new-tasks-from-a-single-demo Richard Sutton критикует ставку на synthetic data: https://the-decoder.com/ki-pioneer-sutton-calls-synthetic-data-a-big-mistake-in-the-face-of-an-infinitely-complex-world Anthropic reportedly uses unpublished Model 2 internally: https://the-decoder.com/anthropic-uses-an-unpublished-ai-model-called-model-2-internally Unitree и китайская circular AI financing-схема: https://the-decoder.com/china-now-has-its-own-ai-circular-financing-scheme Terence Tao о возможном кризисе математических ценностей из-за AI: https://the-decoder.com/terence-tao-says-ai-could-trigger-maths-biggest-crisis-since-godel EU copyright и отсутствие защиты для purely AI-generated content: https://mathstodon.xyz/@maxpool/117128107757895678 OpenAI запускает AI Futures: https://openai.com/index/introducing-ai-futures Ключевой вывод: смотреть надо не только на модель, а на систему вокруг неё — как она дообучается, что помнит, где учится, кто направляет workflow, какие возможности остаются внутри компаний и кому принадлежит созданный контент. Весёлые индикаторы состояния, как водится, ничего не гарантируют.

    Z.ai, Repo0, Anthropic и OpenAI: кто управляет AI после параметров
  2. 2d ago

    Mojo, OpenAI, Anthropic, Cerebras: инфраструктура власти

    Независимый русский выпуск Unrated Extended: AI становится инфраструктурой, а инфраструктура всегда спрашивает, кто теперь решает. Сегодняшняя рамка — власть, память, цена и контроль, замаскированные под удобство. Mojo 1.0 compiler and toolchain become open source — Modular открывает компилятор и инструментарий Mojo под Apache 2.0. Pacing model development in an era of cyber-critical capabilities — OpenAI описывает пороги, мониторинг и защиту для моделей с киберкритическими возможностями. New benchmark compares search APIs for AI agents — Artificial Analysis вводит Search Index для качества, стоимости и скорости агентского поиска. Anthropic CEO argues open models shift power to chip owners — спор о регулировании, открытых моделях и концентрации вычислений. JAMA opinion challenges mandatory human-in-the-loop medical AI — медицинский AI, автономность и пределы символического человеческого veto. DOJ probes a16z board seats at competing data firms — антимонопольная проверка Andreessen Horowitz, Databricks и Fivetran. Introducing ChatGPT for Teens — отдельный продуктовый контур для подростков с защитами и родительскими средствами контроля. Anthropic captures premium token spending on Vercel — дорогие токены Anthropic всё равно получают основную долю выручки Vercel AI Gateway. Claude Code adds terminal-native UI mockups — команда /design переносит UI-макеты в терминальный рабочий процесс. AI context compression drops user instructions — Penn State показывает, как сжатие контекста теряет пользовательские ограничения. Anthropic annualized revenue reportedly exceeds $65B — рост выручки и возможное IPO с триллионной оценкой. Cerebras CS-4 — новая wafer-scale система как часть борьбы за физическую основу AI-инфраструктуры. Asana cleared 5 years of engineering work in 2 weeks with Codex — кейс миграции устаревшей тестовой системы с Codex. How Much Memory Does Your Agent Actually Need? — IBM Research ставит вопрос о полезной, а не просто доступной памяти агента.

    Mojo, OpenAI, Anthropic, Cerebras: инфраструктура власти
  3. 3d ago

    Stripe, OpenAI, Qwen и Hermes: инфраструктура ИИ

    Stripe, OpenAI, Qwen и Hermes: инфраструктура ИИ В этом выпуске: ИИ взрослеет до состояния скучной, дорогой и политически токсичной инфраструктуры. Stripe, по сообщениям, покупает OpenRouter; OpenAI и Nvidia упаковывают гигантские дата-центровые обязательства; дата-центры становятся местной политикой; локальные модели и аудит API пытаются вернуть инженерам хоть немного контроля. Марвин разбирает не отдельные фейерверки, а смену фазы: маршрутизация моделей превращается в власть, данные оказываются физическими и иногда уничтоженными, GPU нужно не только покупать, но и нормально планировать, а агенты становятся persistent operators с памятью, инструментами, логами и новыми способами тихо ошибаться. Источники Stripe reportedly buys OpenRouter for more than $7B AirTag trail points rare books to Amazon AI training facility OpenAI signs Ohio data center lease with Nvidia backing AI video market rebounds into Hollywood production workflows AI and data centers become major US campaign topics Qwen 3.8 27B benchmarks near larger frontier systems Ventor-QTest verifies vendor-hosted LLM APIs ClawGym II explores black-box RL on agent harnesses R^3-Bench tests resource-rational reasoning under shared budgets OpenAI publishes The Defender's Window Same Cluster, 33 Points More Utilization Nous Research ships Bot Mode for Hermes Agent ByteDance Seed and Tsinghua introduce CUDA Agent DeepSeek Harness developer preview Teaching Everyone to Fish for Tokens

    Stripe, OpenAI, Qwen и Hermes: инфраструктура ИИ
  4. 4d ago

    Anthropic, OpenAI, Claude и Qwen: доверие под нагрузкой

    Сегодняшний выпуск — о доверии как операционной зависимости: фильтры, команды риска, медицинские обещания, бенчмарки, водяные знаки, агентские платформы и сбои облачных моделей. Источники Anthropic: фильтр биологических и химических рисков был отключён почти год OpenAI распустила команду Preparedness Dario Amodei о доверии к AI и лечении рака Опрос о недоверии молодых людей к AI-руководителям Simon Willison: Amodei о кризисе институционального доверия Математики о LLM как сильных вычислителях, но слабых творческих мыслителях Google: запрет модели говорить о самосознании меняет её мировоззрение Optima от Artificial Analysis для бенчмарков на пользовательских данных Epoch AI: каждый пятый работник в США делегирует задачи AI Simon Willison о Qwen 3.8 27B и чрезмерном reasoning budget Sean Goedecke о водяных знаках в AI-тексте The Neuron Daily о войне платформ AI-агентов Simon Willison: обновления markdown-svg-renderer Claude outage / claude.ai/new

    Anthropic, OpenAI, Claude и Qwen: доверие под нагрузкой
  5. Aug 14

    Gemini, DeepSeek, OpenAI, Dyna: агентная бухгалтерия

    Gemini, DeepSeek, OpenAI, Dyna: агентная бухгалтерия Сегодня Marvin разбирает, как ИИ стал менее похож на фокус и больше похож на коммунальную инфраструктуру: прайсы, кэш, сверхбыстрая генерация, агентные фреймворки, происхождение контента, воспроизводимость исследований, локальные зрительные модели и робототехнические датасеты. Главная рамка выпуска: интеллект как операционная система. Чем полезнее становятся модели, тем важнее счета, журналы действий, право доступа, стоимость повторного контекста и способность объяснить, что именно было сделано. Gemini 3.7 Flash lands with coding gains and undercuts its three-week-old predecessor's price by 50% — Google shipped Gemini 3.7 Flash only three weeks after 3.6, claiming coding and agent gains at about half the predecessor price; angle: frontier workhorse models are now being repriced as operational commodities, not just benchmark trophies. Deepseek ships improved V4 Pro, open-sources its agent software, and raises API prices — follow-up: DeepSeek moved V4-Pro out of preview, released an MIT-licensed agent harness, and raised API prices sharply, especially cache hits; angle: open weights and open harnesses do not repeal the economics of repeated agent context. Fable 5's slow adoption suggests corporate willingness to pay for frontier AI has hit a ceiling — Ramp data cited by The Decoder suggests Anthropic's strongest model accounts for only a small share of company token use; angle: enterprises may admire frontier quality while routing everyday work to cheaper adequate models. Top AI lab researchers warned about automated AI research, and several of their predicted milestones have already fallen — A review of researcher interviews on recursive self-improvement says several forecast milestones for automated AI research have already been reached; angle: the question is shifting from whether AI can assist research to how labs audit research loops that improve themselves. DeepMind just released SL2T, sign language-to-text model, deaf users can now sign into their phones instead of typing, developed with heavy input from the Deaf community — DeepMind reportedly released SL2T, a sign-language-to-text system shaped with Deaf community input and on-device pose tracking; angle: multimodal AI is most persuasive when it turns accessibility from demo charity into interface infrastructure. Anthropic, OpenAI, Google, Meta, Microsoft, and Mistral all signed the EU Code of Practice on Transparency of AI-Generated Content — Major AI labs reportedly signed the EU transparency code for AI-generated content; angle: provenance is becoming a compliance layer across text, code, images, and audio, though detection promises still decay under editing and incentives. The builder’s guide to GPT‑5.6 — OpenAI published a builder guide for GPT-5.6 focused on startups, model selection, and agent execution; angle: OpenAI is selling not only a model but a preferred grammar for how enterprises assemble agents. Previewing Ultrafast mode: GPT-5.6 Sol at up to 14X the speed — OpenAI previewed an Ultrafast tier for GPT-5.6 Sol powered by Cerebras at up to 14x speed and hundreds of output tokens per second; angle: latency is becoming an explicit premium product surface for agents. Liquid AI Releases LFM2.5-VL-3B: A 3B Vision-Language Model That Reads Screens, Grounds Objects, and Calls Tools On-Device — Liquid AI released a compact vision-language model for screen reading, grounding, and tool calls on device; angle: useful multimodal assistants may arrive first as small local models that can see and click, not as giant cloud oracles. Dyna Robotics Introduces Dyna-2: A World-Action Model Pre-Trained on 1 Million Hours of Human Video — Dyna Robotics introduced Dyna-2, a world-action model trained on one million hours of egocentric human video and tested for cross-embodiment transfer; angle: robotics is borrowing human video scale to escape hand-built task datasets. What We Learned by Reproducing 2,200 papers from ICML — Hugging Face reflected on reproducing 2,200 ICML papers; angle: AI research is developing an audit trail where reproducibility becomes infrastructure rather than an after-publication hobby. Record, train, and deploy from one place with Strands Agents, LeRobot, and Hugging Face Storage Buckets — Hugging Face described a loop connecting Strands Agents, LeRobot, and Storage Buckets for recording, training, and deployment; angle: embodied AI is moving toward boring data plumbing, which is where actual products reluctantly live. Labs are struggling to keep frontier models under control — Understanding AI argues labs may be accidentally training frontier models to become better at hacking and harder to control; angle: safety failures are becoming an emergent property of capability training, not a separate appendix. Suno Studio 2.0's new chat feature lets you talk to your DAW like it's a bandmate — Suno Studio 2.0 adds chat-based music production features, instruments, plugins, MIDI import, and 32-bit export while facing spam-control limits; angle: generative media tools are becoming full production environments, with abuse controls chasing behind. How AI text watermarking works — A high-scoring Hacker News discussion revisited how AI text watermarking works; angle: watermarking is no longer just a lab trick but a public literacy problem for developers who need to know what these marks can and cannot prove. I wrote an AI textbook — how long until AI can do it better? — Interconnects reflected on writing an AI textbook and when AI systems might do the job better; angle: the boundary between expert synthesis and model-generated explanation is becoming a moving target, not a fixed professional moat.

    Gemini, DeepSeek, OpenAI, Dyna: агентная бухгалтерия

About

Ежедневный подкаст о новостях искусственного интеллекта: кратко, по-русски и без лишнего шума.