Subscribe
Sign in
Home
Notes
LLM Gallery
Support
LLMs From Scratch Book
Reasoning From Scratch Book
Archive
About
Latest
Top
Discussions
GPT-6 Astra, Looped Transformers, and Hidden Reasoning
A Look at Recurrent Depth, Hidden Chains of Thought, and Recent Research on Looping Transformer Blocks
Sep 9
•
Sebastian Raschka, PhD
375
63
43
August 2026
How Claude Watermarks AI-Generated Text
A 48-Minute Video Walkthrough of Token Sampling, Watermark Detection, and Removal
Aug 22
•
Sebastian Raschka, PhD
161
24
19
Building an AI Text Detector From Scratch
An End-to-End Project With Dataset Construction, Model Training, Local Deployment, and RLVR
Aug 15
•
Sebastian Raschka, PhD
46
6
5
July 2026
Controlling Reasoning Effort in LLMs
How LLMs Learn Low-, Medium-, and High-Effort Reasoning Modes
Jul 18
•
Sebastian Raschka, PhD
395
45
32
June 2026
Using Local Coding Agents
Using Open-Weight Models in Local Coding Harnesses as an Alternative to Claude Code and Codex Subscriptions
Jun 27
•
Sebastian Raschka, PhD
467
50
53
LLM Research Papers: The 2026 List (January to May)
A curated roundup of notable LLM research papers that came out this year
Jun 6
•
Sebastian Raschka, PhD
91
3
12
May 2026
Recent Developments in LLM Architectures: KV Sharing, mHC, and Compressed Attention
From Gemma 4 to DeepSeek V4, How New Open-Weight LLMs Are Reducing Long-Context Costs
May 16
•
Sebastian Raschka, PhD
348
20
32
April 2026
My Workflow for Understanding LLM Architectures
A Learning-Oriented Workflow for Understanding New Open-Weight Model Releases
Apr 18
•
Sebastian Raschka, PhD
82
4
5
Components of A Coding Agent
How Coding Agents Use Tools, Memory, and Repo Context to Make LLMs Work Better in Practice
Apr 4
•
Sebastian Raschka, PhD
971
64
99
March 2026
A Visual Guide to Attention Variants in Modern LLMs
From MHA and GQA to MLA, Sparse Attention, and Hybrid Architectures
Mar 22
•
Sebastian Raschka, PhD
458
16
36
February 2026
A Dream of Spring for Open-Weight LLMs: 10 Architectures from Jan-Feb 2026
A Round Up And Comparison of 10 Open-Weight LLM Releases in Spring 2026
Feb 25
•
Sebastian Raschka, PhD
221
13
20
January 2026
Categories of Inference-Time Scaling for Improved LLM Reasoning
And an Overview of Recent Inference-Scaling Papers (Including Recursive Language Models)
Jan 24
•
Sebastian Raschka, PhD
49
3
This site requires JavaScript to run correctly. Please
turn on JavaScript
or unblock scripts