Infosec.Pub
  • Communities
  • Create Post
  • Create Community
  • heart
    Support Lemmy
  • search
    Search
  • Login
  • Sign Up
cm0002 to AI - Artificial intelligence@programming.devEnglish · 2 hours ago

Mercury: Ultra-Fast Language Models Based on Diffusion

arxiv.org

external-link
message-square
0
link
fedilink
  • cross-posted to:
  • hackernews@lemmy.bestiver.se
2
external-link

Mercury: Ultra-Fast Language Models Based on Diffusion

arxiv.org

cm0002 to AI - Artificial intelligence@programming.devEnglish · 2 hours ago
message-square
0
link
fedilink
  • cross-posted to:
  • hackernews@lemmy.bestiver.se
We present Mercury, a new generation of commercial-scale large language models (LLMs) based on diffusion. These models are parameterized via the Transformer architecture and trained to predict multiple tokens in parallel. In this report, we detail Mercury Coder, our first set of diffusion LLMs designed for coding applications. Currently, Mercury Coder comes in two sizes: Mini and Small. These models set a new state-of-the-art on the speed-quality frontier. Based on independent evaluations conducted by Artificial Analysis, Mercury Coder Mini and Mercury Coder Small achieve state-of-the-art throughputs of 1109 tokens/sec and 737 tokens/sec, respectively, on NVIDIA H100 GPUs and outperform speed-optimized frontier models by up to 10x on average while maintaining comparable quality. We discuss additional results on a variety of code benchmarks spanning multiple languages and use-cases as well as real-world validation by developers on Copilot Arena, where the model currently ranks second on quality and is the fastest model overall. We also release a public API at https://platform.inceptionlabs.ai/ and free playground at https://chat.inceptionlabs.ai
alert-triangle
You must log in or # to comment.

AI - Artificial intelligence@programming.dev

Aii@programming.dev

Subscribe from Remote Instance

Create a post
You are not logged in. However you can subscribe from another Fediverse account, for example Lemmy or Mastodon. To do this, paste the following into the search field of your instance: !Aii@programming.dev

AI related news and articles.

Rules:

  • No Videos.
  • No self promotion: Don’t post links to your articles.
Visibility: Public
globe

This community can be federated to other instances and be posted/commented in by their users.

  • 6 users / day
  • 82 users / week
  • 358 users / month
  • 1.13K users / 6 months
  • 2 local subscribers
  • 230 subscribers
  • 190 Posts
  • 161 Comments
  • Modlog
  • mods:
  • Vacant@programming.dev
  • cm0002@programming.dev
  • BE: 0.19.13
  • Modlog
  • Instances
  • Docs
  • Code
  • join-lemmy.org