AI Alignment ForumResearch
A Conceptual Framework for Reasoning about Exploration Hacking
This is the second of two posts resulting from a recent Astra/MATS research project investigating exploration hacking in AI debate.
Read at AI Alignment Forum ↗Related

SiliconFrontier Reasoning Reaches the Edge: How to Deploy and Optimize Models on NVIDIA Jetson NVIDIA Technical Blog

VoicesLess about Models; More about Architecture Practical AI

VoicesMoving Beyond RAG with Precomputed Context Software Engineering Daily

Voices[AINews] Claude Fable/Mythos 5.1: new SOTA model, 75% cache price cut but 70% more output tokens Latent Space

LabsBenchMIRT: What are LLM benchmarks actually measuring? Hugging Face Blog
