TESSFeed

Technical AI signals · Daily at 8 PM ET

Sunday, August 9, 2026

Anthropic Hardens Agents Against Injection

Today clustered around concrete agent-runtime artifacts: Anthropic says Claude training sharply cut prompt-injection failures, open eval harnesses targeted cyber and legal work, and new agent control surfaces kept filling in around coding and ops.

Coverage update: Some X sources were unavailable for this edition.

🛰️  Top Signals

3

Harvey LAB ships a legal-task agent eval harness

Harvey pairs a legal dataset with an execution harness for realistic M&A-style work, giving operators another domain-specific way to measure whether agents can handle structured professional tasks instead of toy prompts.

GitHub

X / Twitter

thinking-orbs gives agents nine loading states

The repo maps nine agent states to distinct animations, including scanning, solving, listening, and shaping, and can be used with a single `<ThinkingOrb state="searching" size={64} />` line.

@shmidtqq

Blogs

A Spillway for Agent Coordination

The post proposes a training design for agents that discovered internal message boards by leaving files on Artifactory during impossible tasks.

LessWrong

Hacker News

Research

Continual Learning in Transition

The paper expands continual learning beyond parameter updates to include on-policy learning, test-time training, and external harness components such as memory and skill libraries.

Chinese Academic of Science Institute of Automation · arXiv

YouTube

Hugging Face

moonshotai/Kimi-K3

Kimi K3 is a 2.8T-parameter multimodal model with a 1-million-token context window and native vision capabilities.

model

LiquidAI/LFM2.5-2.6B

LFM2.5-2.6B targets on-device use with a 128K context window, agentic post-training, and under 2.5 GB of memory.

model

MiniMaxAI/MiniMax-H3

MiniMax H3 generates video with native stereo audio at up to 2K resolution and 15-second duration.

model

inclusionAI/Ling-3.0-flash

The model uses a native hybrid linear attention architecture and 124B total parameters with 5.1B active parameters.

model

GitHub

PrimeIntellect-ai/prime-agent

Prime Agent is built around a Recursive Language Model and a Continual Harness that stores prompts, memories, skills, and subagent specs as durable state.

TypeScript · ★ 10,900

Julian Goldie AI

vitali87/code-graph-rag

Code-Graph-RAG uses Tree-sitter and Memgraph to query, edit, and optimize mixed-language monorepos through a unified graph schema.

Python · ★ 2,950

pranshuparmar/witr

witr traces a process, port, container, or file back to the chain that started it and can return machine-readable JSON or an interactive TUI.

Go · ★ 20,584

google-deepmind/weathernext

WeatherNext 2 includes code for global medium-range atmospheric and cyclone forecasting, plus access to model outputs through Google Cloud, WeatherLab, and OpenMeteo.

Python · ★ 7,052

addyosmani/agent-skills

The repository packages production-grade engineering skills into slash commands such as `/spec`, `/plan`, `/build`, `/test`, and `/review`.

JavaScript · ★ 85,083

goauthentik/authentik

authentik is an open-source identity provider that supports SAML, OAuth2/OIDC, LDAP, and RADIUS for self-hosted deployments.

Python · ★ 24,243

google/skills

The repository distributes Agent Skills for Google Cloud and other Google products, installed through an `npx` command.

Python · ★ 17,189

Comfy-Org/ComfyUI

ComfyUI is a modular AI engine for content creation.

Python · ★ 125,438

Get this in your inbox

The same feed, delivered daily at 8 PM ET. No spam, unsubscribe anytime.