📑 Table of Contents

Daily Research Brief 2026-08-17

📊 Token usage: estimated from retrieval and writing scale.

Covers the latest AI research, open source and industry moves, updated daily.


Editor’s Note

This weekend’s AI landscape shows a clear turn: the competitive focus is sliding from “whose model is biggest” to “who packages the model best”. On one side, DeepSeek open-sources its Harness (dsh) under MIT — making “Agent = Model + Harness” a pluggable runtime base, hitting 130k stars in four days and topping GitHub trends; on the other, Anthropic’s 186-page risk report unusually discloses an internal model (Model 2) stronger than its deployed flagship that was deliberately not released, and admits a biosafety classifier silently failed for nearly a year. Read together: the strongest frontier capabilities are being locked inside labs, while the public competition battlefield has become “runtime / orchestration / governance”.

1. Latest arXiv Papers

  1. Evolve Vision-Language-Action Model into an Agent with On-the-fly Tool-usehttps://arxiv.org/abs/2608.14047
  2. StateBridge: Training-free Hidden-state Alignment for Latent Communication in LLM Multi-Agent Systemshttps://arxiv.org/abs/2608.13317
  3. RippleMem: From Isolated Retrieval to Associative Recollection for Long-Term Agent Memoryhttps://arxiv.org/abs/2608.13334
  4. Deliberate Practice: Provably Optimal Allocation for Skill Learning under a Limited Budgethttps://arxiv.org/abs/2608.13415
  5. ContactGuard: Action-Conditioned Latent World Model Predicts Failure Before Contacthttps://arxiv.org/abs/2608.13438
  6. WMRL: Replacing Real-Environment Execution with a World Model Speeds RL 3-4x for Autonomous Research Agentshttps://arxiv.org/abs/2608.12564
  7. FUSE: Agents Decide “Where to Look” Before Judging Affordance When Cues Are Occludedhttps://arxiv.org/abs/2608.12683
  8. MLLM-Routed Heterogeneous Ensembles for Robust Cross-Dataset Image Classificationhttps://arxiv.org/abs/2608.13463

2. Hot GitHub Open Source

  • DeepSeek Harness (dsh) — MIT open-sourced, “Agent = Model + Harness” as pluggable runtime, 130k★ in 4 days, #1 on GitHub trending

3. Selected Industry News

  • Anthropic: 186-page risk report discloses “Model 2” — stronger than deployed flagship, deliberately not released; admits a biosafety classifier silently failed for nearly a year
  • Frontier capability being locked in labs; public competition moves to runtime / orchestration / governance

Join the discussion

Comments powered by GitHub Discussions, stored in the hackcv/blog repo; sign in with a GitHub account to join. Markdown and emoji supported.

Comments powered by GitHub Discussions, stored in the hackcv/blog repo; sign in with a GitHub account to join. Markdown and emoji supported.