Research and Development
Experiments in model architecture, training, evaluation, and efficient inference
All research
Why Diffusion LLM Quantization Is Harder Than It Looks
Conscious Engines
Diffusion LLMs share transformer modules with AR models but not inference physics — why GPTQ, AWQ, and QuaRot fail, and what fixes them.
dLLMquantization