TESSFeed

Technical AI signals · Daily at 8 PM ET

Saturday, July 25, 2026

Open Models Push Into Operations

Today clustered around open-model scale and access, with Kimi and GLM extending the long-context frontier, open-weight politics getting louder, and agent tooling moving deeper into logged-in browsing, web search, and office workflows.

Coverage update: Some X sources were unavailable for this edition.

🛰️  Top Signals

1

Kimi K3 shows its scaling playbook

Moonshot put concrete architecture numbers behind Kimi K3—16 of 896 experts active per token plus delta attention for 1M context—turning this week’s benchmark pressure into a clearer systems story.

@0xMorlex · HN discussion

X / Twitter

sebkrier cites limits in LLM-assisted PoC generation

The quoted findings say progress in LLM-assisted PoC generation depends on stronger validation and failure analysis, and that LLMs should not fully replace deterministic test-generation or example-generation techniques.

@sebkrier

ChrisGPT says Codex voice mode is nearing Jarvis

The post says Codex voice mode still lacks deeper reasoning and that BrowseComp and OSWorld scores have improved over the past year, which is cited as evidence that computer-use systems are getting better fast.

@ChrisGPT

anti open-source

The post lists NVIDIA, Microsoft, Anthropic, and OpenAI examples to argue the major AI labs are all partly open in different ways.

@AtharvaIngle7

Blogs

Introducing Claude Opus 5

Simon Willison says Opus 5 is priced the same as Opus 4.8 and Anthropic describes it as a model that comes close to Claude Fable 5 at half the price.

Simon Willison

Ruff v0.16.0

This version turns on 413 rules by default, up from 59, which caused the author’s unpinned Ruff jobs to fail.

Simon Willison

Your software should build itself

The author proposes an Auto-Syntactic Model in which the agent lives inside the language’s type system and can edit the language itself.

LessWrong

Quoting Boris Cherny

The quote says Opus 5 is Anthropic’s least prompt-injectable model yet, based on PI evals and red teaming.

Simon Willison

Hacker News

Research

YouTube

Sam Altman - How to Start a Startup

The episode covers startup advice from OpenAI’s co-founder and CEO, including operating in chaotic environments and keeping suppliers on timeline.

Relentless

Hugging Face

upstage/Solar-Open2-250B

Solar Open 2 is a 250B-A15B open-weight model that activates only 15B parameters per token and is built for agentic use cases.

model

poolside/Laguna-S-2.1

Laguna S 2.1 is a 118B MoE model with 8B activated parameters per token and mixed sliding-window plus global attention.

model

openbmb/MiniCPM-RobotManip

MiniCPM-RobotManip is a 1.5B vision-language-action model that reduces per-step compute from 125 TFLOPs to 3.3 TFLOPs.

model

fdtn-ai/antares-1b

This model card lists security, vulnerability-detection, agentic, and terminal-agent tags.

model

thinkingmachines/Inkling

Inkling is a general-purpose multimodal model that takes text, image, and audio inputs and outputs text.

model

HuggingFaceCode/stack-v3-train

The Stack v3 is a large source-code dataset crawled from GitHub for pretraining code LLMs with full-repository context.

dataset

Glint-Research/Fable-5-traces

This dataset converts 4,665 Fable 5 coding-agent traces into Pi-compatible sessions for inspection and policy learning.

dataset

smolagents/hf-realtime-voice

This Space swaps the WebRTC SDP proxy for a direct WebSocket route in the Hugging Face speech-to-speech backend.

space

GitHub

alibaba/open-code-review

OpenCodeReview is an AI code-review CLI that reads Git diffs and generates structured comments with line-level precision.

Go · ★ 12,933

permissionlesstech/bitchat

bitchat uses Bluetooth mesh for offline messaging and Nostr for internet-based messaging, with no accounts or phone numbers.

Swift · ★ 28,679

CoreBunch/Instatic

Instatic runs its editor, content engine, forms, auth, plugins, and publisher in one Bun server.

TypeScript · ★ 5,049

shiyu-coder/Kronos

Kronos is an open-source foundation model for financial candlestick sequences trained on data from more than 45 exchanges.

Python · ★ 33,774

Automattic/harper

Harper is an English grammar checker built to avoid sending writing to external servers.

Rust · ★ 13,402

obra/superpowers

Superpowers is a methodology built from composable skills and initial instructions for coding agents.

Shell · ★ 261,081

mattpocock/skills

These agent skills are designed to be small, composable, and usable with any model.

Shell · ★ 188,211

palmier-io/palmier-pro

Palmier Pro is a Swift video editor for Mac where users and agents can generate and edit videos inside the timeline.

Swift · ★ 12,201

RyanCodrai/turbovec

turbovec stores 10 million documents in 4 GB of RAM and uses Rust SIMD search kernels with Python bindings.

Python · ★ 14,274

Get this in your inbox

The same feed, delivered daily at 8 PM ET. No spam, unsubscribe anytime.