OpenAI admits safety monitoring eats 20% of inference compute
OpenAI has admitted that safety monitoring consumes 20% of inference compute, a significant operational cost that translates to $6,000 annually per H100 GPU. The admission signals that supervision is now integral to the …