Skip to content
Read the original: Goodfire· GoodfireAI·Published · 13h agoAI score22/100

Compared to an LLM judge alone, our monitor: - Pareto dominates on recall/FPR - uses 50x less compute, costing <$200 to monitor 1M exchanges - cuts added latency to virtually zero

Original titleCompared to an LLM judge alone, our monitor:

Read the original x.com

Source: Goodfire · x.com