Vesara Daily

Friday, August 21, 2026

Skills

Lark meeting

A hot LarkSuite CLI skill for agent workflows that need to work with meetings rather than leave notes, follow-ups, and calendars in separate tools.

Meeting output becomes useful only when it can enter the same accountable workflow as the work it creates.

Skills.sh · collaboration · workflow

CodeStudio

A current Skills.sh capability from Acquia for CodeStudio work, aimed at putting product-specific development tasks behind reusable instructions instead of one-off prompting.

Specialized skills are most useful when they encode the awkward local rules that a general coding agent will otherwise miss.

Skills.sh · coding · developer

MLOps automation

A hot skill focused on automating MLOps tasks. It is a useful prompt to turn deployment, monitoring, and handoff work into explicit routines.

Models in production need a boring operating layer. This is the part that usually decides whether a prototype survives.

Skills.sh · mlops · reliability

DevOps engineer

A trending capability pack for infrastructure and DevOps tasks, positioned for agents that need to make changes with a clear operational frame.

Infrastructure actions need boundaries, evidence, and rollback paths, not a confident chat response.

Skills.sh · operations · safety

Guarding agent directives

A hot skill about protecting agent instructions from conflicting or malicious directives, directly relevant as prompt injection moves from theory into normal operations.

Treat instruction handling like an input-security problem, especially for agents with filesystem, browser, or deployment access.

Skills.sh · security · safety

Search memory

A current skill for searching an agent memory layer. It focuses on retrieving prior context rather than forcing every task to start from a blank conversation.

Persistent context is useful only when retrieval stays specific enough that old noise does not steer current work.

Skills.sh · memory · context

Repos

Tencent/AI-Infra-Guard

Tencent’s AI security platform scans agents, skills, MCP servers, infrastructure, and jailbreak exposure. It is worth reviewing as a map of the attack surface, not as a substitute for controls.

Apache-2.0 · 50 stars today · 5,071 total

agent-substrate/substrate EARLY

Agent Substrate is a young core system for agents. The project is early, but it belongs on a benchmark list for teams comparing runtime architecture and state handling.

Apache-2.0 · 22 stars today · 1,463 total

microsoft/agent-framework

Microsoft’s framework supports building, orchestrating, and deploying agents across Python and .NET. It is a serious option when an organization needs a conventional engineering surface around agent workflows.

MIT · 66 stars today · 13,015 total

apache/maka EARLY

Apache Maka records model messages, tool calls, results, permissions, and termination events in an append-only log. That is exactly the audit trail most autonomous workflows lack.

Apache-2.0 · 460 stars today · 1,949 total

magnitudedev/magnitude EARLY

Magnitude is an open-source agent with local models built in and an offline-first pitch. It offers a practical test bed for workflows where data locality matters.

MIT · 106 stars today · 1,462 total

pipecat-ai/pipecat

Pipecat is a framework for voice agents and real-time multimodal apps. It is a useful base layer if you are testing voice workflows beyond a scripted demo.

BSD-2-Clause · 41 stars today · 14,370 total

News

AI training and the disappearing physical record

A heavily discussed Hacker News post argues that physical books are being destroyed as AI demand reshapes access to written material. Whatever the framing, teams building data pipelines should document provenance and retention before sources disappear.

HN

GitHub documents its August 17 outage

GitHub’s incident write-up is a reminder that the developer platform is part of the agent stack. If an automated workflow assumes it is always available, its retry, fallback, and approval paths need testing.

HN

Stripe declares the routing layer part of its stack

TLDR flags Stripe’s growing AI infrastructure position after the OpenRouter deal. Routing choice, spend controls, and payments are moving closer together, which changes who owns the operating layer.

TLDR

ChatGPT gets an Apple Messages plug-in

TechCrunch reports a ChatGPT integration for sending texts through Apple Messages. When a model can act in a personal channel, approval and audit need to be part of the default flow.

TechCrunch

Reddit watch

Reddit search arrives as a local MCP server

A builder released a local MCP server for Reddit search and saved-post reading. Check the data path and rate limits before treating any unofficial social connector as production infrastructure.

R/MCP

Teams debate model A/B tests inside one agent

An operator asks how to compare a larger model against a cheaper one inside the same multi-step workflow. Compare task-level success and cost, not isolated benchmark scores.

R/AI_AGENTS

Deals

Callosum

London-based Callosum raised $100 million in an Atomico-led seed round to build infrastructure that unifies AI models and chips.

$100M · Seed

Domyn

Sifted reports that AI model maker Domyn raised more than $1 billion, another large bet on European model and infrastructure capacity.

$1B+ · Funding

Papers

Active Inference as Context Acquisition for AI Agents

This new paper frames context gathering as a decision between asking, retrieving, calling a tool, or proceeding with an assumption. That is a practical design problem for every production agent.

New upvotes · Aug 21

Bounded Sovereignty and the Control Tax

A new paper studies oversight when the deployer uses a frontier model through an API and cannot fully instrument the underlying system. Strong fit for regulated agent deployments.

New upvotes · Aug 21

EnvHarness

EnvHarness proposes adaptive environments for agent learning rather than fixed hand-built tasks. The operator angle is evaluation: static test suites stop teaching you much once agents learn their quirks.

116 upvotes upvotes · Aug 20

SemaPLC

SemaPLC evaluates generated industrial controller code inside an existing project with verification gates. It is a narrow domain, but its insistence on project-level validation is broadly useful.

112 upvotes upvotes · Aug 19

View the email version

No ads, no bullsh*t, one email a day. That’s it.