From the HackerLinks archive

Mistral Medium 3.5

128B model people weighed on local speed, cost, and quantization tradeoffs.

At a glance:
First seen:2026-04-29
Last seen:2026-04-29
Times seen:1
Website:mistral.ai

The short version

128B model people weighed on local speed, cost, and quantization tradeoffs.

Why it caught our attention

It triggered serious debate about what local LLMs can realistically do.

Where it surfaced on Hacker News

2026-04-29

Editorial paraphrase

HN compared its Pareto efficiency, local VRAM needs, and token throughput against Sonnet and other frontier models.

Original thread

Mistral Medium 3.5