|
|
|
|
What's in today's edition
Meta launches Muse, a personal AI agent. Mercury 2.5 arrives. On the funding desk: Mistral AI (€3B). Also Actionable (€8.6M). Research watch covers the day's strongest operator-fit papers.
|
| |
|
01 Agent abilities
|
Skills · MCPs |
Skills
| 01 |
Discord Clawd
A SkillsMP entry for running a Discord-backed OpenClaw session. It is aimed at talking to the active agent rather than searching old Discord messages.
Why it matters: Useful if a client community needs a conversational front door to an already supervised agent.
SkillsMP
Messaging
Client ops
|
| 02 |
Control UI E2E
A testing skill for OpenClaw Control UI changes using Vitest and Playwright, including mocked gateway flows, screenshots, and browser-verifiable evidence.
Why it matters: A good reference for insisting on proof from agent-driven product changes before a client sees them.
SkillsMP
Testing
Reliability
|
| 03 |
OpenClaw nightly release
A release-automation skill for OpenClaw nightlies that uses isolated branches, release CI, retained branches, and a forward-port path back to main.
Why it matters: Useful as a checklist for agent-maintained releases where experiments should not destabilize the working branch.
SkillsMP
Release ops
Engineering
|
| 04 |
Memory wiki maintainer
A memory-maintenance skill for OpenClaw that keeps its knowledge base in predictable pages, tracks managed sections, and ties changes back to evidence.
Why it matters: Worth borrowing if your agent memory needs a reviewable source of truth rather than an opaque context pile.
SkillsMP
Memory
Knowledge ops
|
MCPs
| 01 |
Business Contact Finder
A newly released MCP that checks a business website for a contact route and tests whether that route works. It is directly relevant to prospect enrichment, not generic web search.
Why it matters: Run it against a small lead batch and compare valid-contact yield with your current enrichment chain.
Official MCP Registry
Lead research
Vesara
|
| 02 |
ABMeter
ABMeter exposes A/B experiment management through an assistant, with SDK support across Python, JavaScript, Ruby, Go, React Native, and Node.
Why it matters: It could give agents a disciplined way to propose and record landing-page or outreach tests.
Official MCP Registry
Experimentation
Growth
|
| 03 |
Project Desk
Project Desk is a newly registered work tracker intended for people directing AI across several projects. The agent can update tasks, context, and progress as work moves.
Why it matters: A candidate for the gap between agent execution logs and the operator view of what is actually moving.
Official MCP Registry
Project management
Operator ops
|
| 04 |
Adtest MCP
Adtest scores image, video, and text ads across 13 dimensions before spend. It is designed for assistant-driven review rather than autonomous ad buying.
Why it matters: Use it as a preflight opinion in a creative workflow, then compare its calls with real campaign results.
Official MCP Registry
Creative QA
Marketing
|
|
| |
|
02 Trending repos
|
GitHub · last 24h |
| 01 |
jo-inc/camofox-browser
MIT
— A stealth browser layer positioned as a Puppeteer and Playwright replacement for agents that encounter bot defenses. It added 871 stars today. (+871 today)
|
10,659 ★ |
| 02 |
openai/skills
No license
— OpenAI published a Codex skills catalog. It is worth scanning for task boundaries and packaging patterns, even if you do not adopt the catalog directly. (+490 today)
|
26,623 ★ |
| 03 |
n8n-io/n8n
No license
— The fair-code automation platform remains active on the daily chart. Its mix of visual flows, code steps, and AI integrations keeps it relevant for fast operational prototypes. (+120 today)
|
203,795 ★ |
Trending ≠ vetted. Star counts measure attention, not safety — review the code and pin versions before running anything marked early.
|
| |
|
03 AI & tech news
|
Key reads · Top 5 |
| 01 |
Meta launches Muse, a personal AI agent
HN
Meta introduced Muse as a personal agent with access ambitions that include email, calendars, payments, and health services. The product question is whether users will grant that scope to Meta.
|
| 02 |
Mercury 2.5 arrives
HN
Inception Labs released Mercury 2.5. For operators, the useful follow-up is concrete: benchmark latency, cost, and tool-use reliability on a real production task before moving workloads.
|
| 03 |
Run a 2.8T model from four SSDs
HN
Deltafin demonstrates Kimi K3 running at one token per second on a MacBook Pro by streaming from four SSDs. It is a reminder that storage and serving tricks keep changing the local-model cost curve.
|
| 04 |
LLMs can acquire social bias while exploring
HN
A new paper on HN reports that adaptive exploration can produce novel social biases in language models. Agent teams using self-directed exploration need evaluation checks beyond task success.
|
| 05 |
What an AI project manager changes in practice
Every
Every published a fresh account of onboarding an AI project manager. The useful operator angle is the ongoing maintenance work: delegation creates another system that needs clear ownership and review.
|
|
| |
|
| 01 |
Lean quality controls for Claude Code
R/CLAUDEAI
A builder adapted Lean manufacturing ideas so recurring Claude Code mistakes become tracked failure modes. The interesting part is the feedback loop, not the prompt template.
|
| 03 |
Raggy brings local-document RAG to the CLI
R/RAG
Raggy combines vector search and BM25 over local files, with local or remote generation. It is a compact option for testing document retrieval without starting with a hosted stack.
|
| 04 |
Hard-stop controls for coding agents
R/CHATGPTCODING
A practitioner asks how to make a local coding agent stop immediately without letting it alter its own guardrails. That is a real design requirement for unattended execution.
|
|
| |
|
| 01 |
Mistral AI
· €3B
Series D
Mistral confirmed a €3 billion Series D at a €21 billion valuation. Sovereign AI remains a major capital story in Europe, with infrastructure and regional control part of the pitch.
|
| 02 |
Actionable
· €8.6M
Venture round
Paris-based Actionable raised €8.6 million for predictive customer-experience software. Its product focuses on telling enterprises who may churn, complain, or buy again and why.
|
|
| |
|
06 Research watch
|
Daily top |
| 01 |
Unlocking Lossless Speedups in LLMs via Discrete Diffusion
HF Papers · 121 upvotes · Sep 3
The authors propose diffusion-augmented language models to reduce the sequential bottleneck in text generation. Strong operator fit: serving speed and cost are still constraints on agent throughput.
|
| 03 |
FlowBalance: Verifier-Grounded Self-Improvement
HF Papers · 83 upvotes · Sep 3
FlowBalance studies self-improvement for reasoning models using terminal verifiers alongside denser guidance. Strong fit for anyone building evaluation loops where confident wrong answers are expensive.
|
| 04 |
Why Gated DeltaNet Survives 4-Bit Quantization
HF Papers · 79 upvotes · Sep 3
This paper investigates why a hybrid 27B model retains quality under four-bit quantization. It is a practical read for teams weighing cheaper local inference against degradation risk.
|
|
| |
|
Reply to this email with your feedback, and we'll do our best to implement it in the next edition.
|
| |
|
|
|