TESSFeed

Technical AI signals · Daily at 8 PM ET

Wednesday, August 12, 2026

Qwen3.8 Meets Local Reality

Today centered on Qwen3.8’s open rollout into serving and quantization stacks, while Microsoft shipped its first in-house reasoning model and Vercel put hard production numbers behind agentic software engineering.

🛰️  Top Signals

X / Twitter

Upstage releases Solar Pro 4

Artificial Analysis says Solar Pro 4 scores 42 on its Intelligence Index, up from Solar Pro 3’s 14, and raises API pricing.

@ArtificialAnlys

Blogs

AI Product Engineering Notes

Hamel Husain says the notes distill 13 sessions on evals, context, and systems into about 20 minutes of reading.

Hamel Husain

Hacker News

My Agent Setup

Commenters say the setup uses personally hosted MCP servers for tasks like calendar and email management.

HN discussion

Delta

HN commenters say Zed has a built-in AI agent, but one commenter wants no multiplayer development in the editor and another says AI code summaries can be too verbose or miss edge cases.

HN discussion

Research

YouTube

Hugging Face

deepseek-ai/DeepSeek-V4-Flash-0731

DeepSeek says the official Flash release adds a speculative decoding module and improves agentic capabilities over the preview version.

model

moonshotai/Kimi-K3

Kimi K3 is a 2.8T-parameter open-weight multimodal agentic model with a 1-million-token context window.

model

LiquidAI/LFM2.5-2.6B

LiquidAI says the 2.6B model targets on-device use and runs at 220 tok/s on an Apple M5 Max in under 2.5 GB of memory.

model

deepgrove/maple-preview

DeepGrove says Maple-Preview is a 20B-A1B ternary-weight reasoning LLM with a 131,072-token context window.

model

MiniMaxAI/MiniMax-H3

MiniMax says H3 is an omni-modal system that handles text, images, video, and audio and can generate stereo-audio video up to 2K and 15 seconds.

model

HuggingFaceCode/stack-v3-train

Hugging Face says The Stack v3 is a source-code dataset crawled directly from GitHub for pre-training code LLMs with full-repository context.

dataset

HuggingFaceFW/fineweb

Hugging Face says FineWeb contains 15 trillion tokens of web data for model pre-training.

dataset

r0b0tlab/qwen3.8-max-distillation-50k

The dataset contains 49,772 teacher-generated traces from qwen3.8-max-preview for supervised fine-tuning and off-policy knowledge distillation.

dataset

GitHub

stablyai/orca

Orca runs Codex, ClaudeCode, OpenCode, or Pi in separate git worktrees and includes a mobile companion.

TypeScript · ★ 43,788

semantica-agi/semantica

Semantica says the system ingests enterprise data, builds a context graph and knowledge graph, and adds decision provenance.

Python · ★ 5,657

cathrynlavery/diagram-design

Diagram Design says version 2.0 adds the Loop, a flywheel pattern with a shared-memory hub.

HTML · ★ 10,058

infiniflow/ragflow

RAGFlow says the engine combines retrieval-augmented generation with agent capabilities and a converged context engine.

Go · ★ 87,515

paperclipai/paperclip

Paperclip says the Node.js server and React UI orchestrate teams of AI agents with goals, budgets, and governance.

TypeScript · ★ 77,681

NVIDIA-NeMo/Switchyard

Switchyard routes LLM traffic across providers, translates OpenAI and Anthropic API formats, and records operational metrics.

Rust · ★ 783

cactus-compute/needle

Needle 2 is a 45M-parameter tool-calling model packaged as a single 14MB binary that runs a session in about 28MB of RAM.

Python · ★ 4,172

embabel/embabel-agent

Embabel says the Kotlin framework mixes LLM-prompted interactions with code and domain models for agentic flows.

Kotlin · ★ 4,214

hugohe3/ppt-master

PPT Master generates natively editable PowerPoint files from PDFs, DOCX files, and web pages using Kimi K3’s 1-million-token context window.

Python · ★ 45,531

macro-inc/macro

Macro combines email, messages, docs, tasks, agents, and CRM in one fast interface with shared team-level memory and @linked search.

Rust · ★ 1,736

Get this in your inbox

The same feed, delivered daily at 8 PM ET. No spam, unsubscribe anytime.