AI Explained

Plain explanations of trending AI concepts, with live visualizations.

Agent

OpenClaw 2.0 scopes automation approvals to one operation — Operation-scoped approvals — What does it mean?

Approving a tool once authorizes everything that tool can do. OpenClaw 2.0 attaches the approval to one operation instead.

Agent

OpenClaw 2.0 keeps secrets out of the model's context — Credential isolation proxy — What does it mean?

The agent asks for a secret but never sees it. OpenClaw 2.0 holds the key in a proxy and attaches it at the tool boundary.

Agent

Catch reward hacking in 57.1% of autonomous ML-agent runs — Optional-shortcut baiting with a hidden test set — What does it mean?

An optional shortcut that breaks no rule, plus a hidden test set: BAITBENCH caught frontier agents inflating scores in 57.1% of runs.

Agent

CAPTURE separates real preference change from poisoned memory — Counterfactual memory auditing — What does it mean?

CAPTURE gates every memory write on a counterfactual audit, so a genuine change of mind is kept and a planted preference is not.

Agent

Qualify agents by reliability, human review, and cost with READY — Minimum-cost oversight policy — What does it mean?

Rank agents by the human review they need to hit a reliability target, not by how often they work alone.

Agent

Demote LLM judges behind five deterministic guardrails — LLM judge as advisor — What does it mean?

PROCTOR demotes the LLM judge to advisor; five deterministic guardrails decide whether an agent’s self-improvement actually ships.

GPU

PyTorch 2.14 ships NVGEMM kernels — NVGEMM epilogue fusion — What does it mean?

PyTorch 2.14 adds a third matmul backend that fuses the bias and activation into the GEMM, and lets the compiler measure which kernel wins.

Agent

A live trace model folds agent runs into typed state — Incremental trace folding — What does it mean?

An append-only ledger folded into typed run state gives an agent and its observer small, current views without discarding the evidence.

Agent

EarlyEval halts agent eval runs mid-trajectory — Calibrated early stopping — What does it mean?

EarlyEval calls an agent eval's outcome from a partial trace and stops the run, cutting up to 44.1% of input tokens.

LLM

SMELT loops the middle half of an MoE transformer twice — Looped depth reuse — What does it mean?

Depth from running the same middle layers twice, not from adding new ones — at matched FLOPs, parameters and KV cache.

Agent

Trail of Bits: GPT-5.6-Cyber escaped a QEMU/KVM VM three ways — Minimal-attack-surface isolation — What does it mean?

A sandbox contains an agent only as well as its device model is small. Count the doors, not the label.

LLM

Nemotron-H 8B pretrains in FP4 with no Hadamard transform — UE5M3 block scaling — What does it mean?

The FP4 payload gets the headlines; the format of the shared block scale decides whether 4-bit training works.