‹ 2026-08-13 06:34Z · 12 citations ›

AIINT BRIEF — 2026-08-13

BLUF

DeepSeek has released V4 Pro via API, offering a new high-end reasoning tier with observable control over reasoning depth. On the local inference side, Ollama v0.32.9 introduced NVIDIA Nemotron 3.5 Lightning, a 30B MoE model optimised for always-on agents, while v0.32.10-rc1 improved speculative decoding speeds. Research momentum is shifting towards the operational costs of agentic memory, with new benchmarks quantifying the serving overhead of long-context systems and identifying "catastrophic remembering" in coding agents.

Developments

DeepSeek V4 Pro Release

Ollama Adds Nemotron 3.5 Lightning and Performance Fixes

Benchmarking the Cost of Agentic Memory

New Benchmarks for Enterprise and Voice Agents

LangChain and Claude Code Updates

Trending

Assessment confidence

Corpus coverage is high for recent releases (Ollama, LangChain, Claude Code) and arXiv benchmarks/papers from 2026-08-11 to 2026-08-13. Not covered: specific security incident details beyond the mention of the Zoom hack, as security research is outside the beat.

Sources

  1. DeepSeek V4 Pro 0813 (on OpenRouter)Simon Willison · 2026-08-12 · corpus #1444https://simonwillison.net/2026/Aug/12/deepseek-v4-pro-0813/
  2. ollama/ollama v0.32.9Ollama releases · 2026-08-11 · corpus #1218https://github.com/ollama/ollama/releases/tag/v0.32.9
  3. ollama/ollama v0.32.10-rc1: v0.32.10Ollama releases · 2026-08-12 · corpus #1441https://github.com/ollama/ollama/releases/tag/v0.32.10-rc1
  4. Total Recall at What Cost? Benchmarking the Serving Cost of Agentic Memory SystemsarXiv cs.CL · 2026-08-12 · corpus #1514https://arxiv.org/abs/2608.11879v1
  5. Why Does CLAUDE.md Keep Growing? Catastrophic Remembering in Agentic CodingarXiv cs.AI · 2026-08-11 · corpus #1255https://arxiv.org/abs/2608.11095v1
  6. VAKRA: Evaluating Multi-Hop Reasoning Across APIs and Retrieval Under Tool-Use PoliciesarXiv cs.AI · 2026-08-12 · corpus #1454https://arxiv.org/abs/2608.12282v1
  7. VibeLifeBench: Can Your Life Agent Be Proactive and Persistent in a Living World?arXiv cs.AI · 2026-08-11 · corpus #1283https://arxiv.org/abs/2608.10875v1
  8. DuplexWorld: Can voice agents help you get through the day?arXiv cs.CL · 2026-08-11 · corpus #1319https://arxiv.org/abs/2608.10716v1
  9. ENTLORE: A Graph-Grounded Benchmark for Latent Organizational Reasoning in Enterprise Question AnsweringarXiv cs.CL · 2026-08-11 · corpus #1328https://arxiv.org/abs/2608.10679v1
  10. langchain-ai/langchain langchain-anthropic==1.5.6LangChain releases · 2026-08-13 · corpus #1446https://github.com/langchain-ai/langchain/releases/tag/langchain-anthropic%3D%3D1.5.6
  11. anthropics/claude-code v2.1.229Claude Code releases · 2026-08-12 · corpus #1442https://github.com/anthropics/claude-code/releases/tag/v2.1.229
  12. Language-Conditional Dequantization: Recovering What Quantization Steals from Non-English LanguagesarXiv cs.CL · 2026-08-12 · corpus #1525https://arxiv.org/abs/2608.11786v1

1 of 29 feeds silent · these sources have not been collected recently, so briefs may be missing their coverage: