Skip to content
FRITS AI
Research About Jobs Contact GDPRchat

Research

Tagged “benchmarking”

1 report on this topic. All publications →

  • FRITS-TR-2026-09 15 August 2026 model governanceevaluationproduction LLM systemsbenchmarking

    How to Pick the Best Language Model: A Same-Run Test That Rejected a 95% Score

    If you operate a production language model, you will eventually be offered a cheaper or newer replacement, and the wrong test will tell you it is better. We built the test that does not lie: measure the candidate against the model already running, on your own…

    Frits Lyneborg

    Read the report →

FRITS AI ApS

  • Nyhavn 38, 1051 København K, Denmark
  • CVR (DK): 45733785
  • Contact form

Site

  • Research
  • About
  • Jobs
  • Contact
  • Privacy
  • RSS feed

Elsewhere

  • GDPRchat — our assistant
  • LinkedIn

© 2026 FRITS AI ApS. We build AI that runs entirely in Europe — and publish what we learn.