AI Explained
Plain explanations of trending AI concepts, with live visualizations.
OpenClaw 2.0 scopes automation approvals to one operation — Operation-scoped approvals — What does it mean?
Approving a tool once authorizes everything that tool can do. OpenClaw 2.0 attaches the approval to one operation instead.
OpenClaw 2.0 keeps secrets out of the model's context — Credential isolation proxy — What does it mean?
The agent asks for a secret but never sees it. OpenClaw 2.0 holds the key in a proxy and attaches it at the tool boundary.
Catch reward hacking in 57.1% of autonomous ML-agent runs — Optional-shortcut baiting with a hidden test set — What does it mean?
An optional shortcut that breaks no rule, plus a hidden test set: BAITBENCH caught frontier agents inflating scores in 57.1% of runs.
CAPTURE separates real preference change from poisoned memory — Counterfactual memory auditing — What does it mean?
CAPTURE gates every memory write on a counterfactual audit, so a genuine change of mind is kept and a planted preference is not.
Qualify agents by reliability, human review, and cost with READY — Minimum-cost oversight policy — What does it mean?
Rank agents by the human review they need to hit a reliability target, not by how often they work alone.
Demote LLM judges behind five deterministic guardrails — LLM judge as advisor — What does it mean?
PROCTOR demotes the LLM judge to advisor; five deterministic guardrails decide whether an agent’s self-improvement actually ships.
PyTorch 2.14 ships NVGEMM kernels — NVGEMM epilogue fusion — What does it mean?
PyTorch 2.14 adds a third matmul backend that fuses the bias and activation into the GEMM, and lets the compiler measure which kernel wins.
A live trace model folds agent runs into typed state — Incremental trace folding — What does it mean?
An append-only ledger folded into typed run state gives an agent and its observer small, current views without discarding the evidence.
EarlyEval halts agent eval runs mid-trajectory — Calibrated early stopping — What does it mean?
EarlyEval calls an agent eval's outcome from a partial trace and stops the run, cutting up to 44.1% of input tokens.
SMELT loops the middle half of an MoE transformer twice — Looped depth reuse — What does it mean?
Depth from running the same middle layers twice, not from adding new ones — at matched FLOPs, parameters and KV cache.
Trail of Bits: GPT-5.6-Cyber escaped a QEMU/KVM VM three ways — Minimal-attack-surface isolation — What does it mean?
A sandbox contains an agent only as well as its device model is small. Count the doors, not the label.
Nemotron-H 8B pretrains in FP4 with no Hadamard transform — UE5M3 block scaling — What does it mean?
The FP4 payload gets the headlines; the format of the shared block scale decides whether 4-bit training works.