Weekly Trending
Brainhuggers BureauGitHub Trending
01→
volcengine/OpenViking
+298 this wk · 29.3k starsvolcengine/OpenViking
An open-source context database gives agents a virtual filesystem for memories, resources, and skills, with tiered loading and observable retrieval paths.
02→
chaitanyagiri/munder-difflin
+256 this wk · 2k starschaitanyagiri/munder-difflin
A local desktop harness turns existing terminal coding CLIs into a coordinated multi-agent team with shared memory, mailboxes, worktrees, human approval gates, and an observable operator floor.
03→
akitaonrails/ai-memory
+207 this wk · 2k starsai-memory — portable context as agent infrastructure
ai-memory is an open-source long-term shared-memory system for agent coding CLIs. It carries scoped task context across different agent harnesses instead of leaving it inside one chat interface.
04→
GitHub Trending — apache/maka
+141 this wk · 2k starsapache/maka
Apache Maka is a local-first AI agent workspace that records model messages, tool calls, results, permission decisions, and termination events in an append-only operational log.
05→
GitHub Trending
+66 this wk · 1.4k starsagent-substrate/substrate
An Apache-2.0 Kubernetes control plane that suspends and resumes stateful agent sandboxes, multiplexing many agents across a smaller worker pool.
06→
GitHub Trending — Tencent/AI-Infra-Guard
+28 this wk · 4.9k starsTencent/AI-Infra-Guard
AI-Infra-Guard is an Apache-2.0 red-teaming suite for agent workflows, skills, MCP servers, AI infrastructure, and jailbreak evaluation, with CLI and CI-facing scanners.
HF Papers
01→
CoffeeBench
CoffeeBench: Benchmarking Long-Horizon LLM Agents in Heterogeneous Multi-Agent Economies
CoffeeBench evaluates long-horizon LLM agents in a simulated coffee economy that requires communication, negotiation, pricing, inventory management, and transactions.
02→
Qwen-Image-Agent
Qwen-Image-Agent: Bridging the Context Gap in Real-World Image Generation
Qwen-Image-Agent proposes an image-generation workflow that builds missing context through planning, search, memory and feedback before rendering.
03→
Co-Failure Ceiling
When Does Combining Language Models Help? A Co-Failure Ceiling on Routing, Voting, and Mixture-of-Agents Across 67 Frontier Models
A study across 67 frontier models finds that routing, voting, cascades, and mixture-of-agents systems are bounded by the failures their constituent models share. Ensemble gains depend on genuinely different error patterns.
04→
GBC
GBC: Gradient-Based Connections for Optimizing Multi-Agent Systems
A paper that models LLM multi-agent systems as computational graphs and applies gradient-based connection weights for fine-grained credit assignment.
05→
ProMSA
ProMSA: Progressive Multimodal Search Agents for Knowledge-Based Visual Question Answering
A paper on progressive multimodal search agents that choose among image search, text search, and stopping under explicit tool-call budgets.
06→
Ko-WideSearch
Ko-WideSearch: A Korean Breadth-Search Benchmark for Exhaustive Set Enumeration by Web Agents
A Korean benchmark that evaluates web agents on breadth search, exhaustive set enumeration, and table filling.
07→
Hugging Face Papers
Towards Automating Scientific Review with Google's Paper Assistant Tool
Google's paper proposes an AI-assisted scientific-review tool and a framework for levels of AI-human collaboration in review.
08→
Hugging Face Papers — 2606.24893
AgentOdyssey: Open-Ended Long-Horizon Text Game Generation for Test-Time Continual Learning Agents
AgentOdyssey procedurally generates long-horizon text-game environments for agents that learn during test time through exploration, episodic memory, and planning.