Archives
All the articles I've archived.
- dog
Peyto's 1st Year Photos Selection
A photo wall of Peyto's first year, from April 2025 to July 2026, earliest to latest.
- research
MARS: Several Tokens per Forward Pass, Without Giving Up Autoregression
One SFT checkpoint, two modes: full-quality AR at τ=1.0, or 1.8 tokens per forward pass for 1.56–2.16× speedup, with an exact, free fallback to plain AR.
- research
How Much Do Diffusion LLMs Memorize? (vs Autoregressive)
Same size, same memory ceiling, but the diffusion model needs ~10x the training to reach it, and eventually just can't.
- dog
Living With a Dog: Treat It Like an RL Agent
A dog doesn't get reasons, only consequences. Treat it like a reinforcement-learning agent and almost every training problem collapses into one: what exactly are you rewarding, and what are you punishing?
- publication
On the Role of Discreteness in Diffusion LLMs
Publication in Arxiv
- publication
Sailor2: Sailing in South-East Asia with Inclusive Multilingual LLMs
Publication in Arxiv
- life
Photoset: Japan
A full-screen photo essay from a week in Japan, September 2024. 54 frames, in order. Tap to open the cinematic view.
- publication
Self-Harmonized Chain of Thought
Publication in NAACL 2025 (Main Conference)
- publication
Tabular Chain of Thought
Publication in Findings of ACL 2023