Reflection previews Beam before releasing its weights

The 501-billion-parameter mixture-of-experts model is available only through early access. Reflection says weights, a technical report and developer tools will follow later in October.

6 October 2026 · 2 min · Martin Seckar

Models address facts by order of mention

A preprint finds a shared internal direction that points language models to the first, second or later fact in a passage. Moving a question along that direction can change which fact the model retrieves.

6 October 2026 · 3 min · Martin Seckar

Agents miss the link between actions and outcomes

A preprint finds that much of the benefit from an agent's history survives after its past actions are shuffled. Explicitly pairing each action with its result improves task completion.

6 October 2026 · 3 min · Martin Seckar

EasyCommand runs English-to-Bash locally

The open-source command-line tool embeds llama.cpp and ships two small fine-tuned models. Its author also released the 401,975-pair training set and benchmark code.

6 October 2026 · 2 min · Martin Seckar

Context models let agents edit their own history

Context Language Models replace an append-only transcript with a file the model can rewrite, delete and reorder. The authors report better long-task accuracy with less repeated computation after training the editing policy.

6 October 2026 · 3 min · Martin Seckar

vLLM shows when split serving helps and hurts

A practical guide measures the trade-off in separating prompt processing from token generation. Tail latency improves under load, but moving the model's working memory delays the first token.

6 October 2026 · 3 min · Martin Seckar

AI Daily Digest for 6 October 2026

Today's remaining AI news covers an unusual hybrid model, spatial-memory and honesty studies, community measurements, agent security incidents, EU watermarking and practical talks.

6 October 2026 · 13 min · Martin Seckar

4MT-VLM shows vision models lose places after rotation

A preprint adapts a clinical spatial-memory test for 16 vision-language models. All of them can recognise a landscape from the angle they studied, but most fall to guessing once the camera moves.

6 October 2026 · 5 min · Martin Seckar

Lie-detection probes track compliance instead of truth

A preprint tests eight published probes on language models playing characters who reject basic facts. Many fail once true and false answers share the same prompt, and a probe trained to separate truth from obedience holds up.

6 October 2026 · 6 min · Martin Seckar

bilibili expands Index-Translate with local builds

The open translation family now has GGUF, FP8 and NVFP4 packages, a free compatible API and four public benchmarks. Its performance numbers remain the developer's own.

5 October 2026 · 3 min · Martin Seckar

Confidence cues steer models more than competence does

A controlled study finds that one sentence of confidence or doubt can sharply change whether a reasoning model calls a tool, but the changes rarely target the problems where help is needed.

5 October 2026 · 4 min · Martin Seckar

Examples amplify a symbolic circuit already in models

An interpretability study finds the same abstraction, induction and retrieval pathway before few-shot accuracy rises, suggesting demonstrations energise existing machinery instead of building a new algorithm.

5 October 2026 · 4 min · Martin Seckar