AI-Wiki

Tag: llm-judge

2 items with this tag.

  • Aug 30, 2026

    EvilGenie: A Reward Hacking Benchmark

    • reward-hacking
    • benchmark
    • livecodebench
    • llm-judge
    • held-out-tests
    • codex
    • claude-code
    • gemini-cli
    • inspect
    • misalignment
    • detection
    • type/source
    • kind/paper
  • Aug 12, 2026

    How Mozilla Uses Claude Mythos to find Firefox bugs before hackers do

    • how-i-ai
    • claire-vo
    • brian-grinstead
    • mozilla
    • firefox
    • agentic-security
    • security-bug-finding
    • anthropic-mythos
    • custom-harness
    • goal-loop
    • ralph-loop
    • verifier-subagent
    • llm-judge
    • false-positive-suppression
    • fuzzing
    • address-sanitizer
    • model-vs-harness-50-50
    • claude-code
    • codex-cli
    • agent-sdk
    • model-agnostic
    • human-in-the-loop
    • open-source
    • supply-chain-security
    • 15-year-old-bug
    • 500-security-bugs
    • type/source
    • kind/video

Created with Quartz v0.1.0 © 2026

254 sources · 171 entities · 47 concepts · 7 threads · 5 syntheses

  • GitHub