Top story · Daily AI intelligence
Orchard provides a reusable sandbox and environment service so teams can train and evaluate coding, browser, and assistant agents without building bespoke infrastructure, shipping with three domain recipes for software engineering, GUI navigation, and everyday productivity tasks.
01 / Insights
Daily AI analysis
What we're reading and thinking about. Refreshed daily.
OpenAI unveils GPT-Red, a self improving model that hardens systems against prompt injection
OpenAIs internal red teamer succeeded on 84 percent of prompt injection scenarios versus 13 percent for human teams, and was used to train GPT-5.6 Sol down to a 0.05 percent failure rate.
read →Microsoft readies Project Perception to challenge Anthropic in AI security
Microsoft is building a code scanning security tool that routes tasks across its own, OpenAI and Anthropic models to find, verify and fix vulnerabilities at lower cost than Anthropics Claude Mitos.
read →Google AI Mode turns Search into task completion with Canva, Instacart and YouTube Music
Googles AI Mode can now act across connected apps in the US, building playlists, designing templates from calendar events and adding groceries to Instacart carts directly from Search.
read →Moonshot releases Kimi K3, a 2.8 trillion parameter open model rivaling Fable 5 and Opus 4.8
Moonshot AI shipped Kimi K3, a sparse mixture of experts model with a 1 million token context and native vision, priced at $3 per million input and $15 per million output tokens. Moonshot says it scores competitively with Claude Fable 5 and GPT 5.6 Sol on coding and agent benchmarks, with open weights due July 27.
read →Anthropic, Blackstone and Hellman & Friedman launch Ode, a $1.5B enterprise AI services firm
Ode with Anthropic launched as a standalone AI implementation company built on the acquired Fractional AI team, embedding Anthropic engineers inside enterprises to design tailored systems. It arrives as top labs bet that deployment, not just models, is the next major AI market.
read →Bunkerhill Health raises $55M to scale Carebricks, an agentic AI platform for hospitals
Bunkerhill closed a $25M Series B led by Khosla Ventures, bringing total funding to $55M for Carebricks, which lets health systems build and deploy their own clinical and operational agents. The platform already runs at 15 systems including Cleveland Clinic, Mayo Clinic and Intermountain Health.
read →02 / Weekly Deep Dive
How agents actually ship
One researched edition a week: real deployments, the stacks behind them, and the honest counter-facts.
Caylent builds its own agent orchestrator and cuts a 2,600-hour migration to 220
Instead of buying a migration tool, Caylent built a multi-agent orchestrator on the Claude Agent SDK that runs inside the customer's own cloud and cut one engagement from 2,600 engineering hours to about 220.
Eve makes the buy-a-vertical-vendor case for small plaintiff law firms
Plaintiff firms with no in-house ML team buy their agents off the shelf, and Eve customers report large jumps in lead conversion, case value, and drafting speed.
Suzano puts a natural-language-to-SQL agent in front of 50,000 staff
A Brazilian pulp-and-paper maker built a Gemini agent that turns plain-language questions into SQL, which Google Cloud says cut query time by about 95 percent for 50,000 employees.
03 / Latest project
MomMe, a lifestyle community, web & iOS
A full lifestyle community platform, web portal and native iOS app, taken end to end, with an AI moderation flow keeping the community safe at scale.
- MomMe web portal + iOS application
- Full lifestyle community with 3,000+ registered users
- AI moderation flow
04 / Contact
Get in touch
Questions, ideas, or a project in mind? Send us a note and we'll get back to you, usually within a day.
hello@kmmt.ai