AAI News Hub

#llms

17 briefs tagged #llms.

Research

Researchers propose aDSL for agentic 3D creation

The paper pairs an agent-centric DSL with a multi-agent loop for programmatic 3D content creation.

HF Daily Papers·1d ago
Research

StartupBench tests agents on real startup workflows

The new benchmark finds the strongest tested model completes only about 30% of tasks.

HF Daily Papers+1 outlet·1d ago
Research

Researchers propose hybrid-policy self-editing for LLMs

The arXiv paper targets composability in unstructured knowledge editing for language models.

arXiv cs.AI+1 outlet·6d ago
Research

Researchers map activation spikes in hybrid attention LLMs

The study finds recurring pre-attention spikes and inter-spike plateaus across hybrid linear attention models.

arXiv cs.CL+1 outlet·6d ago
Research

Researchers improve logical compound-answer reasoning

A new framework decomposes AND, OR and NEITHER/NOR options before composing predictions.

HF Daily Papers+1 outlet·6d ago
Research

Google DeepMind introduces sign-language-to-text model

SL2T powers new sign language features for Deaf and hard of hearing users.

Google DeepMind+1 outlet·6d ago
Research

Paper maps blueprint for economic world models

The arXiv paper outlines a six-level roadmap for building generative economic simulations.

arXiv cs.AI+1 outlet·Aug 7
Research

Continual learning work targets forgetting in AI models

New papers propose CP-MoE and frame continual learning as system-level adaptation.

arXiv cs.AI+1 outlet·Aug 7
Research

Researchers outline blueprint for economic world models

The arXiv paper maps a capability ladder for agent-based simulations of economies.

arXiv cs.AI+1 outlet·Aug 7
Research

New papers probe self-distillation for LLM reinforcement learning

ArXiv reports propose ICE and OCSD while warning that privileged-information teachers can fail.

arXiv cs.CL+1 outlet·Aug 6
Products

Reddit adds AI moderation tools for subreddits

Rules Hub uses LLMs to help moderators enforce community rules, with broader launch planned later this year.

The Verge AI·Aug 6
Research

Study finds LLMs fabricate user-profile claims

MirageBench reports pervasive over-inference across 12 personalized LLMs with memory.

HF Daily Papers+1 outlet·Aug 5
Policy

Hank Green steps back after criticism over AI use

The YouTuber said his LLM use was “not healthy” but not for scriptwriting.

The Verge AI·Aug 5
Research

Researchers probe latent reasoning in language models

New arXiv papers test whether continuous hidden-state reasoning can improve LLM reasoning and agent collaboration.

arXiv cs.AI+1 outlet·Aug 4
Research

CALVER challenges voting for LLM causal reasoning

A symbolic verifier outperformed voting and judge methods on multi-answer causal queries.

HF Daily Papers+1 outlet·Aug 4
Research

ReflectRL trains on failed expert reasoning traces

The paper proposes using flawed expert trajectories as reflection signals in on-policy training.

HF Daily Papers+1 outlet·Aug 4
Tools

Hank Green steps back after criticism over LLM use

The creator said his AI use was “not healthy,” while saying he used it for research sources, not scripts.

TechCrunch AI+1 outlet·Aug 2