Theme · 2 articles

LLM

Across all formats — watch, learn, research, benchmarks and demos.

AgentsLLMRAGEvaluationMLOpsClassical MLDataToolingSafetyPolicy
Watch · LLM

Speculative decoding is finally boring — and that's good news

Every serving engine now ships it, defaults are sane, and the 2× is real on the workloads that matter. Here is what to switch on.

A. Martin · 1 Sept · 1 min
Research · Evaluation

Can a judge model grade its own family? Preliminary results

LLM-as-a-judge is everywhere. We asked whether a judge is systematically kinder to models from its own vendor — and got an answer we did not like.

V. Levy dit Vehel · 25 Aug · 1 min