TESSFeed

Technical AI signals ยท Daily at 8 PM ET

Friday, August 14, 2026

Qwen3.8 Lands in the Stack

Today was an operator-heavy day led by Qwen3.8โ€™s open-weight rollout into serving stacks, Perplexity turning Sonar into a tool-using agent API, and fresh routing and verification work for multi-model systems.

๐Ÿ›ฐ๏ธ  Top Signals

X / Twitter

Blogs

Do It Like Darwin

The post frames automated science as a general system for the full scientific discovery loop, not just domain-specific systems like AlphaFold.

LessWrong

Hacker News

AI by Hand

Commenters point to AI by Hand as reading on model interpretability and building LLMs from scratch.

HN discussion

Gemini 3.7 Flash

Top commenters describe image-to-HTML tests, say the model has introductory pricing that doubles on December 31, 2026, and compare it with earlier Flash releases.

HN discussion

Research

Thought-Level Beam Search for Reasoning

The paper frames test-time reasoning as constrained compute allocation over partial trajectories and argues that traditional parallel sampling and subtractive pruning waste hardware or starve it.

Princeton University · arXiv · code · project

YouTube

Hugging Face

Qwen/Qwen3.8-2.4T-A95B

The Hugging Face model card says Qwen3.8-2.4T-A95B is compatible with vLLM, SGLang, and TokenSpeed.

model

moonshotai/Kimi-K3

Kimi K3 is a 2.8T-parameter native multimodal model with a 1-million-token context window and 16-of-896 active experts.

model

MiniMaxAI/MiniMax-H3

MiniMax H3 supports text, image, video, and audio input and can generate video with native stereo audio up to 2K and 15 seconds.

model

Lightricks/LTX-2.5

The model card lists text-to-video, image-to-video, video-to-video, and audio-to-video support.

model

LiquidAI/LFM2.5-2.6B

LFM2.5-2.6B has a 128K context window and is described as running in under 2.5 GB of memory.

model

HuggingFaceCode/stack-v3-train

The Stack v3 is an open source code dataset crawled from GitHub for pre-training code LLMs with full-repository context.

dataset

agent-memory-leaderboard/leaderboard

The leaderboard compares long-term memory systems and memory-enabled agents under one evaluation contract with separate add, search, answer, and eval components.

space

GitHub

cactus-compute/needle

Needle 2 is a 45M-parameter tool-calling and device-use model that runs a full session in about 28MB of RAM.

Python · ★ 5,569

citrolabs/ego-lite

ego lite is a browser built so agents can run tasks in separate Spaces while the user's tabs stay open.

JavaScript · ★ 10,323

holaboss-ai/holaOS

holaOS is a local-first workspace where Claude Code, Codex, and the built-in agent share the same memory, tools, skills, and apps.

TypeScript · ★ 7,248

github/spec-kit

Spec Kit is an open source toolkit for spec-driven development with AI coding agents and a ready-to-use workflow.

Python · ★ 128,469

lightningpixel/modly

Modly turns a photo into a 3D model using open source AI models that run entirely on the user's GPU.

TypeScript · ★ 5,898

infiniflow/ragflow

RAGFlow is an open-source RAG engine that adds agent capabilities and a converged context engine for LLMs.

Go · ★ 88,383

semantica-agi/semantica

Semantica ingests enterprise data into a context graph and knowledge graph for graph analytics and causal reasoning with decision provenance.

Python · ★ 7,491

unslothai/unsloth

Unsloth Desktop can run, train, and deploy AI models locally across LLMs, diffusion, embedding, and audio models.

Python · ★ 71,469

ToolJet/ToolJet

ToolJet is an open-source base for ToolJet AI with a visual app builder and integrations for databases, APIs, SaaS apps, and object storage.

JavaScript · ★ 39,035

cursor/plugins

The repository contains standalone plugin directories with .cursor-plugin/plugin.json manifests for tools such as Continual Learning, Cursor Team Kit, and Thermos.

TypeScript · ★ 2,805

Get this in your inbox

The same feed, delivered daily at 8 PM ET. No spam, unsubscribe anytime.