|
|
Vesara Daily
|
Thursday, July 30, 2026
|
|
|
Microsoft is making its AI stack a product, while agents keep getting more capable and more exposed.
|
|
Today at a glance
Microsoft used its earnings call to position homegrown models, agent harnesses, and a Mythos competitor as products in their own right, a sharper public break from its role as OpenAI’s main platform partner. The funding signal is Freehand’s $75M Series B for autonomous supply-chain spend management, while Centralize’s $15M Series A shows the same bet moving into enterprise sales. For operators, the useful question is not which model wins the chart this week. It is where a reliable agent can own a workflow, with enough controls to make the handoff safe.
|
| |
|
01 Agent abilities
|
Skills · MCPs |
Skills
| 01 |
ui-radar
A trending skill for scanning interfaces and surfacing UI patterns worth borrowing. It is useful when an agent needs a visual reference before it starts changing a product surface.
Why it matters: It gives design-aware agents a concrete reconnaissance step instead of asking them to invent a visual direction from a blank prompt.
Skills.sh
Design research
Product UX
|
| 02 |
sleek-design-mobile-apps
A trending mobile-app design skill from design-layers. It pushes an agent toward polished mobile layouts and implementation details rather than generic component dumps.
Why it matters: Useful for rapid customer-facing prototypes where the first mobile screen often decides whether a demo feels real.
Skills.sh
Mobile build
Prototype speed
|
| 03 |
coll_image-to-video
A trending workflow skill for turning still imagery into video output. It is a practical starting point for short product clips, ads, or social variants built from existing visual assets.
Why it matters: It can turn a static launch asset into a usable distribution asset without adding a separate production handoff.
Skills.sh
Content ops
Distribution
|
MCPs
|
| |
|
02 Trending repos
|
GitHub · last 24h |
| 01 |
sgl-project/sglang
Apache-2.0
— An LLM and multimodal serving framework that remains a practical choice for teams chasing lower latency and more control over inference workloads. (+73 today)
|
30,947 ★ |
| 02 |
lightseekorg/tokenspeed
No licenseearly
— A new inference engine focused on token throughput. Worth watching if model-serving cost or response latency is the constraint in an agent product. (+34 today)
|
1,760 ★ |
| 03 |
kangarooking/cangjie-skill
No license
— A project that distills books, long videos, and podcasts into executable agent skills. The interesting part is the conversion from reference material into repeatable operating behavior. (+182 today)
|
5,268 ★ |
| 04 |
calesthio/OpenMontage
No license
— An agentic video-production system with production pipelines, tools, and skill files. It is a substantial open-source bet on turning coding assistants into media operators. (+668 today)
|
43,988 ★ |
| 05 |
ag-ui-protocol/ag-ui
MIT
— An agent-user interaction protocol for bringing agents into frontend applications. It is relevant if your product needs rich state, approvals, and progress instead of a black-box chat box. (+45 today)
|
15,026 ★ |
| 06 |
ComposioHQ/composio
Apache-2.0
— A toolkit layer for agent integrations, authentication, tool discovery, and sandboxed execution. It is the plumbing behind agents that need to do real work across SaaS systems. (+23 today)
|
29,453 ★ |
Trending ≠ vetted. Star counts measure attention, not safety — review the code and pin versions before running anything marked early.
|
| |
|
03 AI & tech news
|
Key reads · max 5 |
| 02 |
Kimi Code lists K3-256k for long-context coding work
Kimi
Kimi’s documentation now lists K3-256k, putting a very large context window directly into its coding product. Long context helps only if retrieval and tool discipline keep the model from getting lost in the repo.
|
| 03 |
A self-replicating prompt-injection path through Word
Simon Willison
A researcher demonstrated hidden Word instructions that Copilot may treat as source material, then propagate into new documents. Treat document ingestion as untrusted input, especially in automated office workflows.
|
| 04 |
Every maps the infrastructure behind OpenAI’s scale
Every
Every published a fresh look at OpenAI infrastructure as compute demand becomes a business constraint, not merely an engineering detail. Infrastructure choices increasingly shape which agent experiences are affordable to ship.
|
|
| |
|
04 Reddit watch
|
Top 5 · practitioner signal |
| 01 |
The 2026-07-28 MCP specification is out
R/MCP
Developers are parsing a stateless MCP revision that removes the initialize handshake and session IDs. If you run remote servers, check your client and proxy assumptions before treating this as a routine version bump.
|
| 02 |
An MCP server that puts coding intent on the PR
R/MCP
A founder describes an MCP tool that captures intent while the coding agent still has it and attaches that context to the pull request. The idea targets a real review problem: the code survives longer than the reasoning.
|
| 04 |
A cozy 3D view for watching Claude Code work
R/CLAUDEAI
A builder replaced the plain terminal view of Claude Code with a 3D game-like simulation. It is playful, but it also points at a serious product gap: people want to see what their agents are doing.
|
| 05 |
Why Claude feels harsh to its subagents
R/CLAUDEAI
A thread on subagent behavior is a reminder that orchestration quality shapes the user experience as much as the model does. Delegation prompts need clear goals, constraints, and a way to escalate uncertainty.
|
|
| |
|
05 Funding
|
Pre-seed · Series · Growth |
| 01 |
Freehand
· $75M
Series B
Freehand raised a $75M Series B to scale autonomous agents for Fortune 500 supply-chain spend and back-office operations. It is a large wager on agents that touch operational systems, not just dashboards.
|
| 02 |
Centralize
· $15M
Series A
Centralize emerged from stealth with a $15M Series A for an enterprise-sales “Deal GPS.” Its pitch is a familiar agent pattern: assemble scattered account signals into guidance a rep can act on.
|
Only two fresh, qualifying operator-relevant rounds cleared the no-repeat bar today.
|
| |
|
06 Research watch
|
HF Papers · weekly top · max 5 |
| 01 |
CodeNib: A multi-view data system for repository context
HF Papers · 67 upvotes upvotes · Jul 28
CodeNib builds lexical, dense, and structural views for each repository commit so coding agents do less repeated discovery. Strong fit for teams trying to make long-running code work less forgetful.
|
| 02 |
DecoEvo: Solver and rubric-generator skills co-evolve in text
HF Papers · 34 upvotes upvotes · Jul 28
DecoEvo studies text-space optimization where an LLM edits inspectable artifacts rather than weights, while the grading rubric changes too. It is relevant to anyone maintaining prompts, playbooks, or agent policies.
|
| 03 |
JarvisHub: An open harness for multimodal creative agents
HF Papers · 117 upvotes upvotes · Jul 26
JarvisHub is an open harness for long-horizon creative agents working across images, video, audio, UI, and slides. The operator question is whether its task decomposition transfers beyond creative production.
|
|
| |
|
Vesara
|
Post-AI. Human-native.
|
|
Vesara Daily · Curated by Vesara operators · © 2026 Vesara, Inc.
|
|
|