Research explores recurrent depth, hidden reasoning chains, and looping transformer blocks in recent AI models.
What actually shipped in agent engineering, pulled from the labs, arXiv and Hacker News.
See who we followResearch explores recurrent depth, hidden reasoning chains, and looping transformer blocks in recent AI models.
Claude embeds watermarks in generated text through token sampling; video covers detection and removal methods.
Sebastian Raschka describes building an AI text detector including dataset construction, model training, local deployment, and RLVR.
LLMs can learn to operate in low-, medium-, and high-effort reasoning modes that can be controlled.
Open-weight models can replace Claude Code and Codex subscriptions in local coding agent harnesses.
Curated list of notable LLM research papers published from January to May 2026.
Recent open-weight LLMs use KV sharing, mHC, and compressed attention to reduce long-context costs.
Sebastian Raschka describes a workflow for learning new open-weight model architectures.
Coding agents combine tools, memory, and repository context to improve LLM performance.
Guide covers attention variants in LLMs including MHA, GQA, MLA, and sparse attention.
Sebastian Raschka compares ten open-weight LLM architectures released in January-February 2026.
Overview of inference-time scaling techniques for improving large language model reasoning capabilities.
Sebastian Raschka compiled curated research paper lists for LLM work from July to December 2025.
DeepSeek evolved from V3 to V3.2 with architecture, sparse attention, and reinforcement learning updates.
Sebastian Raschka discusses linear attention hybrids, text diffusion, code world models, and small recursive transformers.
Sebastian Raschka explained four main approaches to LLM evaluation with code examples.
Sebastian Raschka published a detailed guide implementing Qwen3, a leading open-source LLM.
Analysis compares architectural advances from GPT-2 to gpt-oss against Qwen3.
Sebastian Raschka compares modern LLM architectures from DeepSeek-V3 to Kimi K2.
Sebastian Raschka compiled 200+ LLM research papers from January to June 2025.
One pattern, one tradeoff, one production failure story. A short weekly briefing for people building agentic systems.
Weekly email, one-click unsubscribe. We only use your address to send the briefing.