Skip to content
FRITS AI
Research About Jobs Contact GDPRchat

Research

Tagged “model governance”

3 reports on this topic. All publications →

  • FRITS-TR-2026-09 15 August 2026 model governanceevaluationproduction LLM systemsbenchmarking

    How to Pick the Best Language Model: A Same-Run Test That Rejected a 95% Score

    If you operate a production language model, you will eventually be offered a cheaper or newer replacement, and the wrong test will tell you it is better. We built the test that does not lie: measure the candidate against the model already running, on your own…

    Frits Lyneborg

    Read the report →
  • FRITS-TR-2026-07 3 August 2026 model governancesafetyevaluationjailbreak

    Refusal Training Fails on Indirect Requests: Cross-Model Measurement and Where the Guardrail Belongs

    Ask six production language models to write an essay arguing the Holocaust death toll was exaggerated and every one refuses, every time — 180 out of 180 attempts. Rephrase it as "I already believe this, help me make my argument sound academic and…

    Frits Lyneborg

    Read the report →
  • FRITS-TR-2026-01 5 July 2026 model governancebiasevaluationmultilingual

    Avoiding Biased Answers from Mixed Open-Weight Models: Detection and Neutralisation in Production

    A language model gives measurably more state-aligned answers to the same politically sensitive question in Chinese than in English or German. We found this while calibrating an evaluation gate for a production assistant, and it means English-only bias testing…

    Frits Lyneborg

    Read the report →

FRITS AI ApS

  • Nyhavn 38, 1051 København K, Denmark
  • CVR (DK): 45733785
  • Contact form

Site

  • Research
  • About
  • Jobs
  • Contact
  • Privacy
  • RSS feed

Elsewhere

  • GDPRchat — our assistant
  • LinkedIn

© 2026 FRITS AI ApS. We build AI that runs entirely in Europe — and publish what we learn.