X (formerly Twitter)

Post

Post

Inception on X: "We are excited to introduce Mercury, the first commercial-grade diffusion large language model (dLLM)! dLLMs push the frontier of intelligence and speed with parallel, coarse-to-fine text generation."

  • user avatar

    We are excited to introduce Mercury, the first commercial-grade diffusion large language model (dLLM)! dLLMs push the frontier of intelligence and speed with parallel, coarse-to-fine text generation.

    00:00

  • user avatar

    Mercury Coder diffusion large language models match the performance of frontier speed-optimized models like GPT-4o Mini and Claude 3.5 Haiku while running up to 10x faster.

    user avatar

    We achieve over 1000 tokens/second on NVIDIA H100s. Blazing fast generations without specialized chips!

    user avatar

    On Copilot Arena, developers consistently prefer Mercuryโ€™s generations. It ranks #1 on speed and #2 on quality. Mercury is the fastest code LLM on the market.

    user avatar

    00:00

    user avatar

  • user avatar

    If you got curious by the is diffusion large language model but realized itโ€™s not open source: itโ€™s your lucky day ๐Ÿ€ LLaDA 8B - an open source apache 2 large diffusion language model is also just out! โœจ competitive with LLaMA 3 8B

    user avatar

    LLaDA (the first Large Language Diffusion Model) is *just* out ๐Ÿ’ฅ and I've built a demo, try out now ๐Ÿ‘จโ€๐Ÿ’ป It's mesmerizing to watch the diffusion process ๐ŸŒ€, and it being a diffusion model gives you superpowers like "the 4th word has to be pineapple" ๐Ÿฆธ Demo and weights ๐Ÿ‘‡

    00:00

Read the original on x.com โ†—