Weekly Trending

Brainhuggers Bureau
← PrevKW25 · 15.06–21.06.2026Next →
GitHub Trending
01
trycua/cua
The desktop becomes the agent test cage
trycua/cua is open-source infrastructure for computer-use agents, including desktop sandboxes, SDKs, and benchmarks across macOS, Linux, and Windows.
02
OpenMontage
The AI coding assistant becomes a production studio
OpenMontage is an open-source agentic video production system with pipelines, media tools, and hundreds of skills for automated content creation and editing.
03
BuilderIO agent-native
The app becomes a habitat built for agents first
BuilderIO's agent-native framing makes the interface shift explicit: applications are becoming structured habitats for autonomous operators.
HF Papers
01
Graph Memory for LLM Agents
Memory is reconstructed, not retrieved
Graph Memory for LLM Agents proposes reconstructing long-horizon agent memory through associative graph exploration rather than static retrieve-then-reason lookup.
02
HarnessX
The agent harness becomes a foundry
HarnessX proposes a composable, adaptive, and evolvable agent harness where prompts, tools, memory, and control flow are typed primitives assembled and improved from execution traces.
03
From Chatbot to Digital Colleague
The chatbot grows a desk
The paper frames the shift from chatbots to persistent autonomous AI systems with workspaces, skills, verification loops, governance, and state-action-observation trajectories.
04
RedAct
The agent trace becomes contraband
RedAct studies how agent execution traces can be redacted to protect procedural skills while preserving useful auditability and evaluation evidence.
05
Data Journalist Agent
The newsroom becomes an agent graph
Data2Story frames data journalism as a multi-agent production system with analysis, evidence inspection, narrative angle, and multimodal presentation as routed roles.
06
VisualClaw
The visual agent grows a skill memory
VisualClaw combines streaming perception, retrieved memories, hot/cold skill injection, and skill evolution for a personalized physical-world agent.
07
WebStep
The browser task becomes a forensic path
WebStep evaluates web agents through semantic trajectories, skill decomposition, and bifurcation analysis instead of only terminal success.
08
PhoneHarness
The smartphone becomes a mixed-action workbench
PhoneHarness routes phone-use agents across GUI, CLI, and structured tools while checking observable side effects.
09
Selective Control under Noisy Perception
The metric smiles while the bridge tissue tears
Selective Control under Noisy Perception shows that aggregate governance metrics can look healthy while harms concentrate on users bridging communities.
10
Dr-DCI
Retrieval becomes a workspace, not a result list
Dr-DCI turns retrieval into dynamic corpus interaction: agents pull documents into an expandable workspace, then search, filter, compare, and verify inside it.