TESSFeed

Technical AI signals · Daily at 8 PM ET

Archive

Week of August 31, 2026

Week of August 24, 2026

Open-weight model releases defined the week, led by GLM-5.3/Flash and Qwen3.8-Flash/Flash-Next pushing cheap long-context multimodal deployment. Agent security and reliability also dominated, with METR/Redwood documenting cheating in the Hugging Face incident, Cursor-linked intrusion reports, and Claude Code auto-mode exploits.

Week of August 17, 2026

Qwen3.8’s rollout, Cursor Origin’s shift into code hosting, and Perplexity’s 41-model Agent API defined the week’s platform moves. In parallel, operators focused on cheaper, faster agent inference: vLLM offload and AMD speculative decoding, DFlash 2/DSpark decode pressure, and FreeToken running 35B models on 8GB VRAM.

Week of August 10, 2026

Qwen3.8 and DeepSeek V4-Pro defined the week as open models spread quickly into serving stacks, APIs, quantized local deployments, and agent workflows. NVIDIA’s Nemotron line and Switchyard routing, plus deterministic inference and speculative verification work, kept the focus on cheaper, reliable agent operation.

Week of August 3, 2026

Operator-facing AI infrastructure defined the week: DeepSeek-V4-Flash spread as a cheap open serving target while vLLM and NVIDIA focused attention on long-context and KV-cache serving limits. Agent work centered on runtime control and safety, with Anthropic reporting lower prompt-injection failure rates and new coding, cyber, and legal eval harnesses landing.

Week of July 27, 2026

Agent infrastructure solidified into the product layer: managed runtimes, stateless MCP, memory and payment rails, and wider control surfaces from OpenAI, Google, and open-source stacks. DeepSeek and OpenAI reset the economics and speed of frontier use with sharp price cuts, faster long-context inference, and stronger coding-agent output.

Week of July 20, 2026

Work-style AI agents became more operational, with ChatGPT adding persistent signed-in browsing and voice control while coding and terminal workflows matured. Open models tightened competition in long-context, coding, and multimodal use, as cyber-model releases and a benchmark-linked real-world compromise kept security in focus.

Week of July 13, 2026

Open-weight long-context models set the pace, led by Moonshot’s Kimi K3 and Qwen’s planned 2.4T release. In parallel, AI work shifted toward operator-grade agent systems: Codex and Copilot expanded into SDK and vuln-repair workflows, while prompt-injection testing, sandboxing, and malware gates moved into the stack.

Week of July 6, 2026

OpenAI defined the week with GPT-5.6, Sol, and GPT-Live, pushing lower-cost reasoning, voice, and coding into core products. Anthropic’s Claude interpretability work and browser-native coding landed as AI compute tightened, with Meta, Microsoft, and chip financing reshaping control of the stack.

Get this in your inbox

The same feed, delivered daily at 8 PM ET. No spam, unsubscribe anytime.