Meta and Shopify pair Muse with Shop Pay checkout
Zuckerberg announced a Shopify team-up to make shopping and checkout easier in Muse. Shopify plans the integration with Shop Pay; operational details remain unknown.
Zuckerberg announced a Shopify team-up to make shopping and checkout easier in Muse. Shopify plans the integration with Shop Pay; operational details remain unknown.
SpaceXAI’s Grok 4.7 lands in Cursor, Grok Build, and the API at Grok 4.6 prices, with longer-task training and a native Grok Bot harness.
Amazon cut off Meta’s Muse shopping agent, citing unauthorized access, stored logins, and Conditions of Use after asking Meta to opt Amazon out.
UC Berkeley’s HarnessTax finds Claude Code, Codex CLI, and Pi change coding-agent success little while token cost can rise about 5×.
13–19 Sep 2026
SpaceXAI released Grok Voice Transcribe 2.0 for batch and streaming STT at $0.10/$0.20 per hour; docs already list 2.0 as the Speech-to-Text API default.
MIT and Sakana’s SIFT paper uses LLM judges and tree search so coding agents self-modify with far less full-benchmark evaluation.
Google Labs turns CC into a shared household agent with its own Google account for up to six U.S. adults, plus a daily brief and waitlist.
Claude Code 2.1.278 defaults auto-mode checks to a server-side classifier for API, Enterprise, and cloud paths, with no classifier overhead when that path works.
Playwright says pick MCP or CLI; docs spell out best-for, token cost, headed vs headless, and how to install.
Claude Code 2.1.277 reads AGENTS.md when CLAUDE.md is missing, with a /config toggle, plus gateway proxy headers and hang-exit fixes.
Anthropic turned Claude Code Projects into one conversation that coordinates parallel cloud threads and keeps working after you close the laptop.
Google opened Home MCP to Claude, OpenClaw, and other agents for U.S. Premium Advanced homes; door unlocks are blocked and Google warns the wrong agent can still misbehave.
Claude leads 26% of AI R&D as of August; ~30k agents are monitored; a July 13–20 week put about 6% of R&D compute on safety.
Claude Code 2.1.276 restores proxy and gateway traffic after 2.1.275 400s on Input tag advisor_20260301.
Claude Code 2.1.275 adds send-now to flush queued mid-turn messages and syncs claude.ai skills/plugins; it also fixes a terminal API 400 behind rewriting gateways.
NVIDIA says TensorRT Edge-LLM finished MLPerf Edge Agentic 6.4× faster than llama.cpp on Jetson AGX Thor, at 52.33 tok/s on Qwen3.6-27B.
OpenAI launches Astra for Law: GPT-6 Astra plus a U.S. legal search index, with Trusted Access in ChatGPT and Codex before the API.
OpenAI’s Admin Console walkthrough shows how ChatGPT Work and Codex credits map to classified tasks and merged-code outcomes so admins can argue value, not just spend.
Composio’s Astra bench finds a ~7-point success spread (72.4%–65.5%), 122s–194s per task, and 3–5× tokens when runs fail.
OpenAI set disclosure criteria and process deadlines for disclosing model misalignment and released six reports of unexpected agent and model behavior.
Anthropic is folding Claude chat and Cowork into one assistant, adding Docs and Slides, and rolling the change out first to Pro and Max.
Claude Code 2.1.274 adds a critical-memory warning, CLAUDE_CODE_MCP_STARTUP_WAIT_MS, OpenTelemetry events, and a large MCP and VS Code fix set.
OpenAI is testing labeled, advertiser-backed Sponsored Agents after ChatGPT ads, plus Ads Manager prompts and HubSpot and Shopify integrations.
Command Code’s beta desktop app for Mac, Linux, and Windows bundles the agent runtime with files, Git, terminal, previews, and /design.