#self-distillation
7 briefs tagged #self-distillation.
Research
Researchers propose hybrid-policy self-editing for LLMs
The arXiv paper targets composability in unstructured knowledge editing for language models.
arXiv cs.AI+1 outlet·6d ago
Research
New papers test self-distillation for agentic RL
AgentOPSD and OCSD propose denser credit assignment, while another study reports failures on harder tasks.
arXiv cs.AI+1 outlet·Aug 7
Research
New papers test self-distillation for agentic RL
AgentOPSD and OCSD propose denser credit signals, while a critique finds self-distillation can fail on harder tasks.
arXiv cs.AI+1 outlet·Aug 7
Research
New papers probe self-distillation for LLM reinforcement learning
ArXiv reports propose ICE and OCSD while warning that privileged-information teachers can fail.
arXiv cs.CL+1 outlet·Aug 6
Research
Researchers target reasoning gaps in model post-training
New arXiv papers propose denser supervision for language and multimodal reasoning models.
arXiv cs.AI+1 outlet·Aug 5
Research
Researchers target sparse rewards in agentic RL
New arXiv papers propose self-distillation methods for denser credit assignment in LLM agents.
arXiv cs.AI+1 outlet·Aug 4
Research
New papers test self-distillation for LLM reinforcement learning
Studies propose ICE and OCSD while warning that privileged-information teachers can fail on harder tasks.
arXiv cs.AI+1 outlet·Aug 4