Crimson AI NewsA CrimsonLingua Network service
EN ع
← Back to news
Research paper Hugging Face

HPSE: Hybrid-Policy Self-Editing Boosts Composable Knowledge Editing in LLMs

AI By Crimson AI Hugging Face Papers 14 August 2026 · 00:00 8 views
Share: X Telegram

Researchers propose HPSE, a method that improves unstructured knowledge editing by distilling from hybrid rollouts that insert missing facts into reasoning paths, enabling composable multi-hop reasoning across four LLM backbones.

HPSE: Hybrid-Policy Self-Editing Boosts Composable Knowledge Editing in LLMs

Key points

Large language models (LLMs) are trained on static corpora, so their knowledge quickly becomes outdated. Knowledge editing (KE) addresses this by updating specific knowledge without affecting unrelated information. Recent work has shifted from structured triples to unstructured KE (UKE), where edits are free-form passages that may state multiple facts at once. However, existing editors often fail to use the injected passage: the model can recall it but cannot answer atomic questions about its facts or compose them into multi-hop reasoning.

To solve this, researchers introduce Hybrid-Policy Self-Editing (HPSE), which treats editing as proactive self-distillation from a privileged in-context state of the same model, requiring no external supervision. The key insight is that pure on-policy distillation is limited because the pre-edited model's rollouts rarely cover the novel injected knowledge. HPSE builds a hybrid rollout that inserts missing facts onto the student's trajectory exactly where coverage fails, while staying on-policy elsewhere.

The team provides a theoretical analysis showing HPSE's advantage over pure on-policy distillation. Empirically, HPSE demonstrates plug-and-play improvements across four LLM backbones and two KE editors under various scenarios, making it a versatile enhancement for knowledge editing.

Source
Hugging Face · Hugging Face Papers
Related news
Research paper
Hugging Face 29 Aug 2026

Hugging Face Audit: 110 of 124 AI Evaluations Fail to Support Their Claims

A new commit-bound census of 124 Inspect Evals units reveals that 110 stop before deterministic inference due to missing historica...

4
Research paper
Hugging Face 29 Aug 2026

Aphanta: New Framework Diagnoses When Image Editing Boosts Multimodal Reasoning

Hugging Face researchers introduce Aphanta, a diagnostic framework that evaluates when image-editing intermediates improve multimo...

5
Research paper
Hugging Face 29 Aug 2026

EditaLive! Enables Real-Time Character Video Editing for Live Streaming

Hugging Face researchers introduce EditaLive, a framework for real-time human-centric video editing in live streams, achieving sta...

4