Daily digest
9 items · ~9 min · Week 2026-W33
Worth knowing (3)
DeepSeek V4 Pro exits preview as V4-Pro-0813 with major coding benchmark gains
DeepSeekDeepSeek quietly shipped DeepSeek-V4-Pro-0813 as the general-availability release of its V4 Pro flagship on August 12, 2026, ending a nearly four-month preview. The 1.6T-parameter MoE model (49B active) keeps its 1M-token context and shows large jumps in agentic coding benchmarks versus the preview, including SWE-bench Verified at 80.6%, LiveCodeBench at 93.5%, and Terminal-Bench 2.1 rising from 72.1 to 87.9.
On-Policy Self-Distillation without Any Supervision
U-OPSD trains a language model on consensus pseudo-solutions built from its own majority-voted generations, correcting confident mistakes without ground-truth labels, teacher models, or environment feedback.
Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA
An open agent-model system for post-deployment continual learning, combining a 744B GLM-5.2 base with a Mixture-of-LoRA architecture that composes specialist adapters for chat, agent, coding, and generative UI tasks; a smaller 50B Qwen3.6-based variant supports local deployment.
For reference (6)
Stealing Reasoning Traces from Proprietary LLM APIs
Researchers show that encrypted reasoning-trace blocks returned by major LLM APIs are interchangeable across sessions, users, and even different models, letting attackers inject them into weaker models to force verbatim disclosure of the plaintext reasoning.
Yandex launches Lumena AI companion feature in Yandex Books
YandexYandex introduced the first capability of Lumena, a personalized AI companion planned across its entertainment services, starting with an 'About Characters' feature in Yandex Books that lets readers learn about characters and plot without spoilers, based on their own reading progress.
Yandex AI assistant helps pediatric rheumatologists remotely monitor juvenile arthritis patients
YandexYandex, built on Yandex AI Studio and trained on Russian Ministry of Health clinical guidelines and drug documentation, deployed an AI assistant inside a remote-monitoring app used by 32 pediatric rheumatologists tracking around 1,500 juvenile arthritis patients across Russia.
Claude Code v2.1.229 adds self-hosted-runner hooks and plugin marketplace command sources
AnthropicFollowing v2.1.228's skill-sync hardening, Anthropic shipped Claude Code v2.1.229 on August 12, adding self-hosted-runner hook support, plugin marketplace 'command' sources that let a local tool supply and hot-reload plugin directories, VSCode sidebar session groups, and SSE keepalive for gateway streaming.
Zed editor 1.16.0 preview adds Gemini 3.6 Flash support
Zed IndustriesZed's 1.16.0 preview release (Aug 12) adds Gemini 3.6 Flash to its Google AI model list, alongside Git panel improvements (collapsible grouped changes, optional stash messages) and zoomable/scrollable Mermaid diagram rendering.
OpenCode v1.18.17 and v1.18.18 ship provider routing and reliability fixes
SSTSST's OpenCode shipped v1.18.17 (Aug 12) and v1.18.18 (Aug 13), fixing the Kimi system prompt for Moonshot/Kimi providers, xhigh reasoning effort on xAI models, DeepSeek V4 Flash sampling defaults, session-compaction quality for smaller models, and enabling PDF vision attachments for GitHub Copilot models.