Vesara
Vesara Daily Wednesday, September 9, 2026
 
What's in today's edition
Meta launches Muse, a personal AI agent. Mercury 2.5 arrives. On the funding desk: Mistral AI (€3B). Also Actionable (€8.6M). Research watch covers the day's strongest operator-fit papers.
 
01  Agent abilities Skills · MCPs
 
Skills
01 Discord Clawd
A SkillsMP entry for running a Discord-backed OpenClaw session. It is aimed at talking to the active agent rather than searching old Discord messages.
Why it matters: Useful if a client community needs a conversational front door to an already supervised agent.
SkillsMP Messaging Client ops
02 Control UI E2E
A testing skill for OpenClaw Control UI changes using Vitest and Playwright, including mocked gateway flows, screenshots, and browser-verifiable evidence.
Why it matters: A good reference for insisting on proof from agent-driven product changes before a client sees them.
SkillsMP Testing Reliability
03 OpenClaw nightly release
A release-automation skill for OpenClaw nightlies that uses isolated branches, release CI, retained branches, and a forward-port path back to main.
Why it matters: Useful as a checklist for agent-maintained releases where experiments should not destabilize the working branch.
SkillsMP Release ops Engineering
04 Memory wiki maintainer
A memory-maintenance skill for OpenClaw that keeps its knowledge base in predictable pages, tracks managed sections, and ties changes back to evidence.
Why it matters: Worth borrowing if your agent memory needs a reviewable source of truth rather than an opaque context pile.
SkillsMP Memory Knowledge ops
 
MCPs
01 Business Contact Finder
A newly released MCP that checks a business website for a contact route and tests whether that route works. It is directly relevant to prospect enrichment, not generic web search.
Why it matters: Run it against a small lead batch and compare valid-contact yield with your current enrichment chain.
Official MCP Registry Lead research Vesara
02 ABMeter
ABMeter exposes A/B experiment management through an assistant, with SDK support across Python, JavaScript, Ruby, Go, React Native, and Node.
Why it matters: It could give agents a disciplined way to propose and record landing-page or outreach tests.
Official MCP Registry Experimentation Growth
03 Project Desk
Project Desk is a newly registered work tracker intended for people directing AI across several projects. The agent can update tasks, context, and progress as work moves.
Why it matters: A candidate for the gap between agent execution logs and the operator view of what is actually moving.
Official MCP Registry Project management Operator ops
04 Adtest MCP
Adtest scores image, video, and text ads across 13 dimensions before spend. It is designed for assistant-driven review rather than autonomous ad buying.
Why it matters: Use it as a preflight opinion in a creative workflow, then compare its calls with real campaign results.
Official MCP Registry Creative QA Marketing
 
02  Trending repos GitHub · last 24h
 
01 jo-inc/camofox-browser  MIT — A stealth browser layer positioned as a Puppeteer and Playwright replacement for agents that encounter bot defenses. It added 871 stars today. (+871 today) 10,659 ★
02 openai/skills  No license — OpenAI published a Codex skills catalog. It is worth scanning for task boundaries and packaging patterns, even if you do not adopt the catalog directly. (+490 today) 26,623 ★
03 n8n-io/n8n  No license — The fair-code automation platform remains active on the daily chart. Its mix of visual flows, code steps, and AI integrations keeps it relevant for fast operational prototypes. (+120 today) 203,795 ★
Trending ≠ vetted. Star counts measure attention, not safety — review the code and pin versions before running anything marked early.
 
03  AI & tech news Key reads · Top 5
 
01 Meta launches Muse, a personal AI agent  HN
Meta introduced Muse as a personal agent with access ambitions that include email, calendars, payments, and health services. The product question is whether users will grant that scope to Meta.
02 Mercury 2.5 arrives  HN
Inception Labs released Mercury 2.5. For operators, the useful follow-up is concrete: benchmark latency, cost, and tool-use reliability on a real production task before moving workloads.
03 Run a 2.8T model from four SSDs  HN
Deltafin demonstrates Kimi K3 running at one token per second on a MacBook Pro by streaming from four SSDs. It is a reminder that storage and serving tricks keep changing the local-model cost curve.
04 LLMs can acquire social bias while exploring  HN
A new paper on HN reports that adaptive exploration can produce novel social biases in language models. Agent teams using self-directed exploration need evaluation checks beyond task success.
05 What an AI project manager changes in practice  Every
Every published a fresh account of onboarding an AI project manager. The useful operator angle is the ongoing maintenance work: delegation creates another system that needs clear ownership and review.
 
04  Reddit watch Top 5
 
01 Lean quality controls for Claude Code  R/CLAUDEAI
A builder adapted Lean manufacturing ideas so recurring Claude Code mistakes become tracked failure modes. The interesting part is the feedback loop, not the prompt template.
02 An internal bot exposed an unannounced reorganization  R/AI_AGENTS
A cautionary account of a wiki-reading assistant surfacing details from a restricted spreadsheet. Retrieval permissions and answer-time policy checks need to be designed together.
03 Raggy brings local-document RAG to the CLI  R/RAG
Raggy combines vector search and BM25 over local files, with local or remote generation. It is a compact option for testing document retrieval without starting with a hosted stack.
04 Hard-stop controls for coding agents  R/CHATGPTCODING
A practitioner asks how to make a local coding agent stop immediately without letting it alter its own guardrails. That is a real design requirement for unattended execution.
 
05  Funding
 
01 Mistral AI · €3B  Series D
Mistral confirmed a €3 billion Series D at a €21 billion valuation. Sovereign AI remains a major capital story in Europe, with infrastructure and regional control part of the pitch.
02 Actionable · €8.6M  Venture round
Paris-based Actionable raised €8.6 million for predictive customer-experience software. Its product focuses on telling enterprises who may churn, complain, or buy again and why.
 
06  Research watch Daily top
 
01 Unlocking Lossless Speedups in LLMs via Discrete Diffusion  HF Papers · 121 upvotes · Sep 3
The authors propose diffusion-augmented language models to reduce the sequential bottleneck in text generation. Strong operator fit: serving speed and cost are still constraints on agent throughput.
02 NeoHorse-1: Recursive Self-Improvement via Agentic Post-Training  HF Papers · 84 upvotes · Sep 8
NeoHorse-1 explores agent-native post-training with a routing harness that turns observed performance into training signals. Medium fit today, but relevant to how agent systems may improve from their own runs.
03 FlowBalance: Verifier-Grounded Self-Improvement  HF Papers · 83 upvotes · Sep 3
FlowBalance studies self-improvement for reasoning models using terminal verifiers alongside denser guidance. Strong fit for anyone building evaluation loops where confident wrong answers are expensive.
04 Why Gated DeltaNet Survives 4-Bit Quantization  HF Papers · 79 upvotes · Sep 3
This paper investigates why a hybrid 27B model retains quality under four-bit quantization. It is a practical read for teams weighing cheaper local inference against degradation risk.
 
  Your feedback
 
Reply to this email with your feedback, and we'll do our best to implement it in the next edition.
 
Vesara mark Vesara Find the 5% of revenue leaking out of your company.
InstagramLinkedInX vesara.ai
Vesara Daily · Curated by Vesara operators · © 2026 Vesara, Inc. · vesara.ai Unsubscribe