Find a story
Search Spins
Search titles, summaries, and missing voices across published articles — press releases, announcements, and media coverage.
4 results for “Classifiers”
Judging LLM-as-a-Judge: Concerning Rubric Artifacts in LLM-based Automated Text Generation Evaluation
A research paper demonstrates that LLM-as-a-Judge evaluation systems often rely on rubric text alone—not candidate responses—to generate scores, undermining their validity as objective evaluators of AI-generated text.
Sep 4, 2026
Asymmetries in Spontaneous and Instructed Deception
A new arXiv preprint reports empirical evidence that spontaneous (uninstructed) and instructed deception in Llama-3.1-70B-Instruct share latent geometric structure in model representations, revealing asymmetric transferability between detection and steering across deception modes.
Sep 2, 2026
Classifiers: Track What Your Agents Do and What It Costs - OpenRouter
OpenRouter introduced 'Classifiers', a new feature enabling developers to monitor and quantify the behavior and cost of AI agents using its API platform.
Jul 25, 2026
Claude Code now has a built-in browser that lets the AI read, click, and type on external websites
Anthropic's Claude Code developer tool now includes an integrated browser enabling AI-driven navigation, reading, and limited interaction with external websites within the IDE, with safety gates for write actions.
Jul 12, 2026