Honest comparisons and stack breakdowns. No affiliate fluff.
DeepSeek open-sourced Harness v0.1 under MIT — a plugin-first agent runtime where models, tools, sandboxes and loops all swap in config. The harness is the product.
Cursor now boots Cloud Agents from pre-warmed environment snapshots instead of cloning and installing per run. The setup step was always billable — here's how to cut it.
Composio's new pricing lands August 15: the $29 tier drops from 200K tool calls to 50K, overage jumps to $4 per 1,000. Existing plans hold through December 31.
VS Code 1.133 opens the Agents window without GitHub sign-in and lets you swap model providers between turns. Your editor just got less locked in.
Microsoft is merging the consumer and Microsoft 365 Copilot apps and retiring Group Chats, Podcasts, Deep Research, and Copilot Labs. Export before the date.
Ryan Dahl's celld runs Cloudflare Workers and Durable Objects on your own S3 storage. The API is compatible, the license is Apache-2.0, the pricing claim is contested.
Automattic's Mesh CRM hit Android with an AI network query layer and a free 1,000-contact tier. What a personal CRM actually buys a small operator.
A researcher probed frontier models and found Claude Opus 5 answers like a January 2026 model despite a published May 2026 knowledge cutoff. Test your own.
A free 0-100 AI agent security score across six dimensions. Five of the six are ordinary supply-chain hygiene — which is the actual finding here.
GitHub Copilot for JetBrains now runs local Ollama models as a BYOK provider and remembers context across agent sessions. One is portability. One needs a policy.
Cross-session messaging lets Claude Code sessions hand off findings across terminals and machines. Here are the controls you should set before it matters.
GitHub shipped allowedMcpServers and deniedMcpServers for enterprise Copilot. Fail-closed by default. The pattern is worth copying even if you don't use Copilot.
Cloudflare shipped Kitesurf, a browser built for AI agents that uses 3-4x less CPU and up to 7x less memory than Chromium. Here's what it changes for your automation.
Nadella confirmed Microsoft is merging Copilot chat, GitHub Copilot, Cowork and Autopilots into one app. What consolidation does to your M365 bill and controls.
Atlassian's cloud revenue grew 31% while it guides Data Center down 17%. The self-hosted option ends March 28, 2029 — plan the migration on your calendar.
Reuters: the FCC is drafting an import ban on new Chinese optical transceiver models. The vendor with 27% share is Chinese. What that does to compute costs.
MacPaw is building on-device inference with Liquid AI and plans to hand it to Setapp developers. The credit meter is the part worth reading.
Microsoft Research's Orchard framework is MIT-licensed and public. Its own benchmarks show agents at 69.7% on SWE-bench and 59.6% on assistant tasks. Plan for that.
Genspark open-sourced GenOffice, an Apache 2.0 AI office suite for Mac and Windows. Read the routing: model calls go through Genspark's servers, not your key.
Cloudflare shipped a Billable Usage API for all self-serve accounts, FOCUS-aligned and live today. Why a queryable cloud bill beats a monthly PDF.
ShieldFont is a web font that feeds AI scrapers gibberish while humans see real text. Clever — and exactly wrong if you sell through AI search.
DeepSeek shipped V4-Flash-0731 on July 31 under an MIT license — 82.7 on Terminal Bench 2.1, $0.14 input and $0.28 output per million tokens. Route your cheap work here.
Cisco open-weighted two tiny models that find where a CVE lives in your codebase — Apache 2.0, 350M and 1B params, and your source never leaves the building.
PortSwigger's Burp AT puts AI agents inside Burp Suite Professional, with scope enforced outside the model. What agentic pentesting means for your web app.
Simile hit a $2B valuation selling AI stand-ins for consumers. Enterprises need simulation because they lost the customer. Small businesses haven't.
Pangram raised $9M from Menlo Ventures and shipped Pangram 4 plus an image detector. Every accuracy number is vendor-supplied. Here's how to test it on your own data.
Microsoft posted $90B in Q4 revenue and $115.9B in FY26 capex while pitching its own MAI models against OpenAI and Anthropic. Why the harness matters more than the model.
The FCC added humanoid robots, robot dogs, and power inverters to its Covered List on July 29, blocking new imports. What an import ban does to an automation roadmap.
A CVSS 8.7 n8n sandbox escape lets any workflow editor run OS commands on your host. Fixed in 2.31.5 and 2.32.1 — check your version today.
Ruff v0.16.0 raises the default rule set from 59 rules to 413 and formats Python in Markdown. Great defaults, and a very loud first CI run.
CodeMender hit public preview July 21: it finds a vulnerability, proves it's exploitable in your own sandbox, then writes the fix. Read the fine print.
GitHub made a 3-day Dependabot cooldown the default for version updates. Security updates still ship immediately. Why the delay is the point.
CVE-2026-29059 lets anyone read files off a Windmill server with no login. Patched in January, exploited in July. What self-hosting automation actually costs.
OpenAI's first hardware, a $230 keypad with status lights for AI coding agents, admits the real bottleneck: you can't tell what your agents are doing.
Anthropic's Claude Security plugin is in public beta for Claude Code — terminal vulnerability scanning with an adversarial verification pass. What it changes.
Runway launched a model router for generative media that picks by quality, speed, or cost. Useful — but routing logic is business logic you should own.
Claude Code Desktop can build, run, and tap through your iPhone app in a live simulator pane — no screen-recording permissions, but screenshots leave your Mac.
Abstract Security raised $25M for a streaming-first alternative to monolithic SIEM. The pattern — separate sources from destinations — applies far beyond security.
Katana shipped an MCP server and one-click AI replenishment for cloud inventory. Why a vendor-supported MCP endpoint matters more than the forecasting feature.
NVIDIA's SIGGRAPH 2026 news put MCP support into Adobe, Blender, Unreal and Houdini. Your content production pipeline just became scriptable by agents.
Infinity raised $15M to let an AI agent write inference kernels for any chip. The operator lesson: AI cost is a software layer, not just a hardware bill.
AWS launched CloudWatch coding agent insights for Claude Code, Codex, and Copilot. It's built on OpenTelemetry — which means you're not locked to AWS to get it.
VulnHunter is a free, Apache-2.0 agentic code security scanner from Capital One that hunts exploitable bugs instead of flagging patterns. Here's how to use it.
Epoch AI tested Pangram, GPTZero, and Originality.ai. Give a model a writing sample to imitate and detection collapses. Why you can't govern with AI detectors.
PyTorch 2.13 brings FlexAttention to Apple Silicon with a ~12x speedup on sparse patterns. Why that quietly lowers the cost floor for running your own models.
The Commerce Department moved the UAE to Country Group A:5, making advanced AI chip exports license-free overnight. What export policy means for where your AI runs.
Meta puts its custom Iris inference chip into production in September and is doubling compute to 14GW. What hyperscaler vertical integration means for your AI costs.
Claude Code desktop now has a sandboxed, isolated in-app browser the agent can read and click. The doc-lookup loop is gone — and browser-driving agents are now table stakes.
ZML released LLMD, a free inference server that runs open-weight LLMs across Nvidia, AMD, Google TPU, Apple, and Intel — breaking chip vendor lock-in.
SambaNova raised $1B at an $11B valuation for on-premises AI inference chips, with JPMorgan as a customer. Why owning the inference layer is now a bank-grade bet.
Squidbleed (CVE-2026-47729) is a Heartbleed-style memory leak in Squid proxy that survived 29 years. Here's who's exposed and what to do.
Baseten raised $1.5B at up to $13B for AI inference infrastructure. Why the boring serving layer — not the model — decides your cost and uptime.
Mistral OCR 4 adds bounding boxes, block types, and confidence scores at $2–4 per 1,000 pages — and runs in one container you control. What that unlocks for operators.
A single deterministic retrieval tool took biology-agent accuracy from 16.9% to 99.7%. The operator lesson: give agents real tools, not a bigger model.
Bhavin Turakhia is self-funding Neo, an AI-native, model-agnostic rival to Office and Workspace. The real lesson for operators: when to rebuild vs. bolt on.
Cursor shipped a native iOS app that fires off cloud coding agents to merge-ready PRs. What a small studio actually gets — and where the human checkpoint still goes.
X launched a hosted MCP server that lets Claude, Cursor, and Grok read the platform through your account — but not write. Why that read-only line is the right call for AI integrations.
We built Ceesvee, a CSV editor that handles 100MB+ files, on Tauri instead of Electron. Here's the honest comparison — including where Electron still wins.