We’re putting $5M behind research and development of smaller, specialized AI models for real-world deployment — Read our manifesto →

    Research and Development

    Experiments in model architecture, training, evaluation, and efficient inference

    All research

    Why Diffusion LLM Quantization Is Harder Than It Looks

    Conscious Engines

    Diffusion LLMs share transformer modules with AR models but not inference physics — why GPTQ, AWQ, and QuaRot fail, and what fixes them.

    dLLMquantization