Latest Articles
Building an AI Text Detector From Scratch
Building an AI Text Detector From Scratch

An End-to-End Project With Dataset Construction, Model Training, Local Deployment, and RLVR

Controlling Reasoning Effort in LLMs
Controlling Reasoning Effort in LLMs

How LLMs Learn Low-, Medium-, and High-Effort Reasoning Modes

Using Local Coding Agents
Using Local Coding Agents

Using Open-Weight Models in Local Coding Harnesses as an Alternative to Claude Code and Codex Subscriptions

LLM Research Papers: The 2026 List (January to May)
LLM Research Papers: The 2026 List (January to May)

A curated roundup of notable LLM research papers that came out this year

Quick Notes
How Claude's Text Watermarking Works

Short illustration of how Claude's text watermarking is supposed to work based on Anthropic's released materials.

Build a Reasoning Model From Scratch Is Now on Amazon

Short note on the Amazon availability of Build a Reasoning Model From Scratch and a warning about counterfeit black-and-white copies sold...

Muse Glimmer 30B Architecture Notes

Short architecture note on Meta Muse Glimmer 30B, including gated local and global GQA, KV-cache efficiency, and release-time benchmark c...

LLMs From Scratch Reaches 100,000 GitHub Stars

Short note celebrating the LLMs-from-scratch repository passing 100,000 GitHub stars and summarizing its learning materials.

Kimi K3 Architecture Notes

Short architecture note on Kimi K3, including LatentMoE, Kimi Delta Attention, Attention Residuals, NoPE, multimodality, and inference-ef...