📑 Table of Contents

Tech perspective · Today’s top picks: arXiv papers / GitHub open source / Industry news.


I · arXiv Latest Papers

WorldSculpt: Generating Compositional Worlds from Grounded Videos

Abstract: We study the problem of generating a compositional 3D representation of a cluttered scene containing hundreds of objects. The goal is to represent the scene as a collection of individual object meshes placed in a shared world frame, as required by downstream applications such as gaming, AR/VR, simulation, and robotics. This task is challenging in densely cluttered scenes, where objects heavily occlude each other.
Field: AI / LLM
Reason: Recently submitted, developer/research-oriented, worth a quick look.
Link: http://arxiv.org/abs/2609.05416v1

UniMate: One Unified Model to Animate Diverse Skeletons

Abstract: Recent advances in automatic rigging now deliver animation-ready 3D assets at scale, yet generating the motion to drive them remains a bottleneck. Existing learned animators are topology-constrained: they rely on category-specific templates or require per-skeleton fine-tuning and reference motions at inference. We present UniMate, a unified foundation model that synthesizes articulated motion for arbitrary skeletons.
Field: AI / LLM
Reason: Recently submitted, developer/research-oriented, worth a quick look.
Link: http://arxiv.org/abs/2609.05415v1

WearableQA: A Benchmark for Health Reasoning over Real-World Wearable Data

Abstract: Recent advances in wearable sensing enable continuous monitoring of physiological and behavioral signals, yet existing benchmarks rarely evaluate whether AI systems can reason over a real user’s longitudinal wearable record. We introduce WearableQA, a benchmark comprising 4,084 10-option multiple-choice questions constructed from the wearable time series, blood biomarkers, and demographics of 200 users.
Field: AI / LLM
Reason: Recently submitted, developer/research-oriented, worth a quick look.
Link: http://arxiv.org/abs/2609.05405v1

Diffusion TV: Experiencing Diffusion Models through Tangible, Embodied Interaction

Abstract: Diffusion TV is an interactive AI art installation that offers a tangible and embodied experience of diffusion models through a modified CRT TV. By physically manipulating the TV’s antenna, audiences control the clarity of AI-generated images and sounds, metaphorically enacting the denoising process that underlies diffusion-based generation. Using the tuning knob, participants switch between three modes.
Field: AI / LLM
Reason: Recently submitted, developer/research-oriented, worth a quick look.
Link: http://arxiv.org/abs/2609.05404v1

RegionFed: Federated Learning for Personalized Query Understanding in Heterogeneous Retail Environments

Abstract: Retail search systems serve diverse geographic regions with distinct query patterns, vocabularies, and product preferences, creating significant data heterogeneity that challenges both privacy-preserving training and model personalization. Federated learning offers a natural solution for privacy, but standard FL methods produce global models that sacrifice regional performance, while existing personalization methods have limitations.
Field: AI / LLM
Reason: Recently submitted, developer/research-oriented, worth a quick look.
Link: http://arxiv.org/abs/2609.05403v1

Same Trajectory, Contradictory Rewards: Paraphrase Fragility in Vision Language Reward Models

Abstract: Vision-language models are increasingly used as reward functions for robotic learning, but this role requires paraphrase invariance: the same trajectory should receive the same reward under semantically equivalent goal descriptions. We show that current VLM reward models often violate this property. Paraphrasing the instruction alone can substantially change predicted progress scores, and can even change reward rankings.
Field: AI / LLM
Reason: Recently submitted, developer/research-oriented, worth a quick look.
Link: http://arxiv.org/abs/2609.05401v1

Abstract: When there is not enough labeled data to properly train deep learning models, transfer learning can help. We still do not fully understand how effective it is in neuroimaging, especially for Alzheimer’s disease research. It is also not clear if these transferred models can work on new datasets without being retrained for each specific task. We evaluate whether a compact, supervised pretrained model can generalize.
Field: AI / LLM
Reason: Recently submitted, developer/research-oriented, worth a quick look.
Link: http://arxiv.org/abs/2609.05400v1

From Interpretability Methods to Interpretable Models

Abstract: More than a decade in, explainable AI (XAI) for computer vision has assembled a mature toolbox: attribution, feature visualization, concept-based, and circuit-based methods. Yet almost all of the field’s effort has gone into building and comparing these methods, and little into the question they were meant to answer—how interpretable are our models, and are we making progress as they evolve?
Field: AI / LLM
Reason: Recently submitted, developer/research-oriented, worth a quick look.
Link: http://arxiv.org/abs/2609.05399v1

openclaw/openclaw

Description: The AI that really does things. Any OS. Any Platform. The lobster way. 🦞
Stars: 389199⭐
Reason: Recently active with leading stars, worth following.
Link: https://github.com/openclaw/openclaw

obra/superpowers

Description: An agentic skills framework & software development methodology that works.
Stars: 283067⭐
Reason: Recently active with leading stars, worth following.
Link: https://github.com/obra/superpowers

NousResearch/hermes-agent

Description: The agent that grows with you.
Stars: 243245⭐
Reason: Recently active with leading stars, worth following.
Link: https://github.com/NousResearch/hermes-agent

n8n-io/n8n

Description: Fair-code workflow automation platform with native AI capabilities. Combine visual building with custom code, self-host or cloud, 400+ integrations.
Stars: 203712⭐
Reason: Recently active with leading stars, worth following.
Link: https://github.com/n8n-io/n8n

Significant-Gravitas/AutoGPT

Description: AutoGPT is the vision of accessible AI for everyone, to use and to build on. Our mission is to provide the tools, so that you can focus on what matters.
Stars: 187194⭐
Reason: Recently active with leading stars, worth following.
Link: https://github.com/Significant-Gravitas/AutoGPT

firecrawl/firecrawl

Description: The context API to search, scrape, and interact with the web at scale. 🔥
Stars: 177863⭐
Reason: Recently active with leading stars, worth following.
Link: https://github.com/firecrawl/firecrawl

f/prompts.chat

Description: f.k.a. Awesome ChatGPT Prompts. Share, discover, and collect prompts from the community. Free and open source — self-host for your organization with complete privacy.
Stars: 169642⭐
Reason: Recently active with leading stars, worth following.
Link: https://github.com/f/prompts.chat

Snailclimb/JavaGuide

Description: Java interview & backend general interview guide, covering computer basics, databases, distributed systems, high concurrency, system design and AI application development.
Stars: 158371⭐
Reason: Recently active with leading stars, worth following.
Link: https://github.com/Snailclimb/JavaGuide

III · Industry News

Native Full-Modal Tech Strategic Close Loop, HiDream Releases Embodied World Model HiDream-O1-Embodied

Content: HiDream releases new embodied world model HiDream-O1-Embodied, achieving full-modal tech strategic close loop.
Reason: From aggregated source, industry-focused.
Source: Quantum Bit

Baidu Integrates Xiaodu Hardware, Baidu Agent Enters Home Space

Content: On September 8th, at the Baidu AI Day Xiaodu New Product Launch in Beijing, Xiaodu announced that Super Xiaodu has completed agent upgrade, releasing multiple new home scenario agent applications. As the foundation of Xiaodu’s agent capabilities, Baidu Partner fully lands on Xiaodu smart screens, smart cameras, smart speakers and other new hardware products, promoting agents entering home spaces.
Reason: From aggregated source, industry-focused.
Source: Leifeng AI

Enflame Technology IPO Results Released! Raising 6.119 Billion Yuan, Domestic AI Chip Leader to Land on STAR Market

Content: On the evening of September 7th, Enflame Technology (688801.SH) officially released its IPO results. The issue price was 142.18 yuan per share, with 43.035 million shares issued, raising a total of 6.119 billion yuan. The online subscription reached 7.032 million accounts, with a final winning rate of 0.02455315%, showing strong market demand.
Reason: From aggregated source, industry-focused.
Source: Leifeng AI

Fields Medal Winner Joins LLM Race! 4B Mobile Qwen + Cloud GLM Breaks ARC-AGI 3

Content: “Finding a mathematical common foundation between two models is actually very difficult.”
Reason: From aggregated source, industry-focused.
Source: Quantum Bit

PHOGRAIN 400Gbps PIN PD Supports Global AI Computing Optical Interconnect to 3.2T Transceiver Modules

Content: Shenzhen, September 6, 2026 — The 27th China International Optoelectronics Expo (CIOE) will open next week (September 9-11) at Shenzhen International Convention and Exhibition Center. Global leading photodetector chip company PHOGRAIN will appear at Hall 11, Booth 11B33, officially releasing 400Gbps back-illuminated PIN photodetector (PIN PD) chip.
Reason: From aggregated source, industry-focused.
Source: Leifeng AI

Behind “ONE FOR ALL”: What Physical AI Close Loop is PaXini Building?

Content: Over the past month, PaXini AI’s strategic progress has accelerated significantly: releasing PX6AX GEN4 product matrix with GEN4 FUSE true 6D tactile sensing chip; Beijing headquarters landing, forming Beijing strategic R&D and Shenzhen manufacturing delivery dual-city collaboration; completing joint-stock reform and 1 billion yuan new round of financing.
Reason: From aggregated source, industry-focused.
Source: Leifeng AI

TokenRhythm Releases NeoHorse Model, Exploring Harness-Driven RSI Path

Content: In a project scheduling test, a 4B base model found files in the working directory but missed an email containing the latest dependency constraints. It generated plans based on outdated information and wrote files to the wrong location. This case comes from TokenRhythm’s recent technical report, jointly releasing the first Agent-Native model NeoHorse-1 with 4B and 9B versions.
Reason: From aggregated source, industry-focused.
Source: Leifeng AI

Silicon Valley AI Unicorn Switches to Alibaba Qwen: Perplexity Builds Local Agent with Qwen3.8

Content: On September 8th, according to US tech media Siliconangle, Silicon Valley AI unicorn Perplexity launched a new local Agent product Portable Computer, using Alibaba’s latest open source model Qwen3.8-27B. On NVIDIA DGX Spark hardware, Perplexity specially optimized the PPLX 27B series algorithm, enabling the Qwen model to better help users with file processing, data analysis and programming on local hardware.
Reason: From aggregated source, industry-focused.
Source: Leifeng AI

Token consumption statistics for this task: Script mode (opencode launch, no independent token metering)

Join the discussion

Comments powered by GitHub Discussions, stored in the hackcv/blog repo; sign in with a GitHub account to join. Markdown and emoji supported.

Comments powered by GitHub Discussions, stored in the hackcv/blog repo; sign in with a GitHub account to join. Markdown and emoji supported.