DEV Community

Papers Mache profile picture

Papers Mache

404 bio not found

Joined Joined on 
Larger LLMs generate more toxic content

Larger LLMs generate more toxic content

2
Comments
2 min read
AI/ML Research Digest — Aug 22, 2026

AI/ML Research Digest — Aug 22, 2026

1
Comments
3 min read
Skill Entropy reward boosts multi-skill reasoning

Skill Entropy reward boosts multi-skill reasoning

Comments
2 min read
Peer Selection Cuts Regret in Multi‑Agent Coordination

Peer Selection Cuts Regret in Multi‑Agent Coordination

Comments
2 min read
Closed‑loop detection improves attribution accuracy

Closed‑loop detection improves attribution accuracy

Comments
2 min read
Expert-locality-aware decode routing reduces MoE serving latency

Expert-locality-aware decode routing reduces MoE serving latency

Comments
2 min read
Uncertainty‑guided distillation improves student efficiency

Uncertainty‑guided distillation improves student efficiency

Comments
2 min read
Adaptive compute techniques yield significant inference speedups across models

Adaptive compute techniques yield significant inference speedups across models

Comments
2 min read
ReBA routing balances multimodal experts losslessly

ReBA routing balances multimodal experts losslessly

1
Comments
2 min read
Simple BM25 outperforms agents on large corpora

Simple BM25 outperforms agents on large corpora

Comments
2 min read
Test‑time geometry constraints tighten vision model predictions

Test‑time geometry constraints tighten vision model predictions

Comments
2 min read
Template tokens enable head pruning

Template tokens enable head pruning

Comments
2 min read
VideoChat3 and RoboTTT Use Separate Transformer Backbones for Video Chat and Robot Policies

VideoChat3 and RoboTTT Use Separate Transformer Backbones for Video Chat and Robot Policies

Comments 1
2 min read
World‑to‑Wrist: Task‑Conditioned Wrist Modeling for Fine‑Grained Robot Manipulation

World‑to‑Wrist: Task‑Conditioned Wrist Modeling for Fine‑Grained Robot Manipulation

Comments
2 min read
Segment‑level consolidation reduces construction tokens by up to ~86% with modest throughput gains

Segment‑level consolidation reduces construction tokens by up to ~86% with modest throughput gains

Comments
2 min read
Proactive memory stops long‑horizon drift

Proactive memory stops long‑horizon drift

Comments
2 min read
SkewAdam cuts training memory over 60%

SkewAdam cuts training memory over 60%

Comments
2 min read
Distilled Chain‑of‑Thought Can Be Gaming Exploited

Distilled Chain‑of‑Thought Can Be Gaming Exploited

Comments
2 min read
Vision‑language models cap at sixty percent accuracy

Vision‑language models cap at sixty percent accuracy

Comments
2 min read
Self-supervised skill rubrics cut LLM agent failures

Self-supervised skill rubrics cut LLM agent failures

Comments
2 min read
Instruction tuning harms confidence calibration

Instruction tuning harms confidence calibration

1
Comments
2 min read
Sparse Delta Memory multiplies hidden capacity

Sparse Delta Memory multiplies hidden capacity

Comments
2 min read
Byte‑Exact KV Cache Grafting Gains 12% Accuracy

Byte‑Exact KV Cache Grafting Gains 12% Accuracy

Comments
2 min read
Migration-aware scheduling reduces worst-case streaming video latency by up to 38%

Migration-aware scheduling reduces worst-case streaming video latency by up to 38%

Comments
2 min read
Hybrid attention reduces compute cost while preserving quality

Hybrid attention reduces compute cost while preserving quality

Comments
2 min read
FactorJEPA learns traffic scenes from few labels

FactorJEPA learns traffic scenes from few labels

Comments
1 min read
Training VLMs to absorb tools eliminates dependency

Training VLMs to absorb tools eliminates dependency

Comments
2 min read
Mixture-of-Experts make multilingual retrieval cheap

Mixture-of-Experts make multilingual retrieval cheap

Comments
2 min read
LLM‑generated pruning slashes multimodal FLOPs ninefold

LLM‑generated pruning slashes multimodal FLOPs ninefold

Comments
2 min read
Recursive verification loop improves reasoning performance

Recursive verification loop improves reasoning performance

Comments
2 min read
Graph-native RL phases make scientific hypotheses traceable

Graph-native RL phases make scientific hypotheses traceable

Comments 1
2 min read
Looped Transformers Double Depth Without Extra Memory

Looped Transformers Double Depth Without Extra Memory

Comments
2 min read
Provenance frameworks for multimodal agentic reasoning

Provenance frameworks for multimodal agentic reasoning

Comments
2 min read
8‑bit model runs 1.2B‑parameter music on Pi

8‑bit model runs 1.2B‑parameter music on Pi

Comments
2 min read
Tiny adapter matches 32B model performance

Tiny adapter matches 32B model performance

Comments
2 min read
EnvACE trains tool-use without real environments

EnvACE trains tool-use without real environments

Comments
2 min read
Evolutionary selection boosts LLM agent performance

Evolutionary selection boosts LLM agent performance

Comments
3 min read
Adaptive tool selection cuts agent cost tenfold

Adaptive tool selection cuts agent cost tenfold

Comments
2 min read
Rerankers overlook 55% coordination failures

Rerankers overlook 55% coordination failures

Comments
2 min read
Distillation replaces RL for exploration gains

Distillation replaces RL for exploration gains

Comments
2 min read
Fixed‑Answer Bias Emerges Before LLM Reasoning

Fixed‑Answer Bias Emerges Before LLM Reasoning

Comments
2 min read
AI/ML Research Digest — Aug 15, 2026

AI/ML Research Digest — Aug 15, 2026

Comments
4 min read
Adversarial Drift Derails Imagined Futures in Multimodal Agents

Adversarial Drift Derails Imagined Futures in Multimodal Agents

Comments
2 min read
Zero-Mem reduces memory‑operation token use and inference time

Zero-Mem reduces memory‑operation token use and inference time

Comments
2 min read
Decodability supervision erases hidden private codes

Decodability supervision erases hidden private codes

Comments
2 min read
LLMs reconstruct only 27% of ideas

LLMs reconstruct only 27% of ideas

Comments
2 min read
Contamination inflates macro‑F1 by eleven points

Contamination inflates macro‑F1 by eleven points

Comments
2 min read
One-step diffusion reaches ImageNet quality

One-step diffusion reaches ImageNet quality

Comments
2 min read
AI/ML Research Digest — Aug 02, 2026

AI/ML Research Digest — Aug 02, 2026

Comments
3 min read
AI/ML Research Digest — Jul 05, 2026

AI/ML Research Digest — Jul 05, 2026

Comments
4 min read
AI/ML Research Digest — Jul 12, 2026

AI/ML Research Digest — Jul 12, 2026

Comments
4 min read
AI/ML Research Digest — Jul 19, 2026

AI/ML Research Digest — Jul 19, 2026

Comments
4 min read
AI/ML Research Digest — Jul 26, 2026

AI/ML Research Digest — Jul 26, 2026

Comments
3 min read
AI/ML Research Digest — Aug 09, 2026

AI/ML Research Digest — Aug 09, 2026

Comments
4 min read
Autoregressive retriever training raises BEIR scores

Autoregressive retriever training raises BEIR scores

1
Comments
1 min read
Binary chunk trees cut RAG latency

Binary chunk trees cut RAG latency

Comments
2 min read
JSON-Schema masks can block needed tool calls

JSON-Schema masks can block needed tool calls

Comments
2 min read
Tiered models separate public and private capabilities

Tiered models separate public and private capabilities

1
Comments
2 min read
Head-level attention fusion trims compute

Head-level attention fusion trims compute

1
Comments
2 min read
RL-driven data mixing boosts evaluation scores

RL-driven data mixing boosts evaluation scores

1
Comments
2 min read
loading...