66 Tokens Make a Diffusion Language Model Look Easy
A diffusion language model generates text by starting from masked or otherwise corrupted tokens and iteratively restoring them. In this…
Highlights advances in core systems, technical breakthroughs, experiments, and academic work driving progress.
A diffusion language model generates text by starting from masked or otherwise corrupted tokens and iteratively restoring them. In this…
The standard story is that LLMs work in words. They predict the next token, so surely their internal reasoning is…
Alexander Lerchner’s paper on conscious AI does something unusual: it does not start by asking whether today’s models seem conscious….
Kimi K2.6 is everywhere in preview chatter. Kimi K2.6 is also, based on the sources we can actually verify, not…
Most vision models get good by seeing absurd amounts of data. Zero-shot world models are interesting because they try a…
Qwen3.6-35B-A3B is being passed around as a major new open model release: 35 billion total parameters, 3 billion active, Apache…
A paper reports a new state-of-the-art result. The repo is public. The figures look clean. The conference is top-tier. In…
ARC Prize didn’t just tweak a benchmark. The ARC-AGI-3 human baseline now uses a median human run per level instead…
A spiking neural network allegedly reached 1.088 billion parameters and trained from random initialization to a reported loss of 4.4…