hn • r/hackernews
Comment on: Mercury 2: Fast reasoning LLM powered by diffusion
Have been following your models and semi-regularly ran them through evals since early summer. With the existing Coder and Mercury models, I always found that the trade-offs were not worth it, especially as providers with custom inference hardware could push model tp/s and latency increasingly higher.I can see some very specific use cases for an existing PKM project, specially using the edit model for tagging and potentially retrieval, both of which I am using Gemini 2.5 Flash-Lite still.The pricing makes this very enticing and I'll really try to get Mercury 2 going, if tool calling and structu