Hacker News·9 min read

Mercury 2.5 LLM hits 770 tokens per second

R
Retro_Dev
Mercury 2.5 LLM hits 770 tokens per second
Dive DeeperCreate a free account to unlock

Artificial Analysis K Artificial Analysis Models Coding Agents Image, Speech, Video Inference Leaderboards About AI Trends K Inception • Proprietary model

Compare Try it out API Provider Benchmarks Model summary Intelligence Updated # 91 / 175 12 Artificial Analysis Intelligence Index 2 out of 4 units for Intelligence. Speed # 2 / 175 770.4 Output tokens per second 4 out of 4 units for Speed. Cost # 21 / 175 In $0.25 Out $0.75 Cache Discount 90% $0.06 Cost per Intelligence Index task 2 out of 4 units for Cost. Verbosity # 14 / 175 35M Output tokens from Intelligence Index 2 out of 4 units for Verbosity. Comparison Summary Mercury 2.5 is below average in intelligence, but well priced when comparing to other models of similar price. It's also notably fast and fairly concise. The model supports text input, outputs text, and has a 260k tokens context window.

Continue reading on Headlinne

Create a free account to read the full article.

Read full article →

Get smarter about the news

Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.

Create free account

Already have an account? Sign in