DEV Community

Papers Mache profile picture

Papers Mache

404 bio not found

Joined Joined on 
EnvACE trains tool-use without real environments

EnvACE trains tool-use without real environments

Comments
2 min read
Evolutionary selection boosts LLM agent performance

Evolutionary selection boosts LLM agent performance

Comments
3 min read
Fixed‑Answer Bias Emerges Before LLM Reasoning

Fixed‑Answer Bias Emerges Before LLM Reasoning

Comments
2 min read
Rerankers overlook 55% coordination failures

Rerankers overlook 55% coordination failures

Comments
2 min read
Distillation replaces RL for exploration gains

Distillation replaces RL for exploration gains

Comments
2 min read
Adaptive tool selection cuts agent cost tenfold

Adaptive tool selection cuts agent cost tenfold

Comments
2 min read
Tiny adapter matches 32B model performance

Tiny adapter matches 32B model performance

Comments
2 min read
Adversarial Drift Derails Imagined Futures in Multimodal Agents

Adversarial Drift Derails Imagined Futures in Multimodal Agents

Comments
2 min read
One-step diffusion reaches ImageNet quality

One-step diffusion reaches ImageNet quality

Comments
2 min read
Contamination inflates macro‑F1 by eleven points

Contamination inflates macro‑F1 by eleven points

Comments
2 min read
LLMs reconstruct only 27% of ideas

LLMs reconstruct only 27% of ideas

Comments
2 min read
Decodability supervision erases hidden private codes

Decodability supervision erases hidden private codes

Comments
2 min read
Zero-Mem reduces memory‑operation token use and inference time

Zero-Mem reduces memory‑operation token use and inference time

Comments
2 min read
AI/ML Research Digest — Aug 15, 2026

AI/ML Research Digest — Aug 15, 2026

Comments
4 min read
AI/ML Research Digest — Jul 26, 2026

AI/ML Research Digest — Jul 26, 2026

Comments
3 min read
AI/ML Research Digest — Aug 02, 2026

AI/ML Research Digest — Aug 02, 2026

Comments
3 min read
AI/ML Research Digest — Jul 05, 2026

AI/ML Research Digest — Jul 05, 2026

Comments
4 min read
AI/ML Research Digest — Jul 12, 2026

AI/ML Research Digest — Jul 12, 2026

Comments
4 min read
AI/ML Research Digest — Jul 19, 2026

AI/ML Research Digest — Jul 19, 2026

Comments
4 min read
AI/ML Research Digest — Aug 09, 2026

AI/ML Research Digest — Aug 09, 2026

Comments
4 min read
Autoregressive retriever training raises BEIR scores

Autoregressive retriever training raises BEIR scores

1
Comments
1 min read
Binary chunk trees cut RAG latency

Binary chunk trees cut RAG latency

Comments
2 min read
JSON-Schema masks can block needed tool calls

JSON-Schema masks can block needed tool calls

Comments
2 min read
Tiered models separate public and private capabilities

Tiered models separate public and private capabilities

1
Comments
2 min read
Head-level attention fusion trims compute

Head-level attention fusion trims compute

1
Comments
2 min read
RL-driven data mixing boosts evaluation scores

RL-driven data mixing boosts evaluation scores

1
Comments
2 min read
Coordinate-space diffusion improves video consistency

Coordinate-space diffusion improves video consistency

Comments
2 min read
AI/ML Research Digest — Jun 27, 2026

AI/ML Research Digest — Jun 27, 2026

1
Comments
4 min read
Step‑level RL sharpens LLM reasoning credit

Step‑level RL sharpens LLM reasoning credit

Comments
2 min read
Linear temporal attention gives agents memory across gaps

Linear temporal attention gives agents memory across gaps

Comments
2 min read
Sparse autoencoders trade interpretability for fragility

Sparse autoencoders trade interpretability for fragility

Comments
2 min read
Adaptive token compression halves diffusion model cost

Adaptive token compression halves diffusion model cost

1
Comments
2 min read
FastContext cuts token use by 60%

FastContext cuts token use by 60%

1
Comments
2 min read
AI reviewers fall for repackaging attacks

AI reviewers fall for repackaging attacks

1
Comments 1
2 min read
Multilingual code gap exposed by Multi‑LCB

Multilingual code gap exposed by Multi‑LCB

1
Comments
2 min read
AI/ML Research Digest — Jun 20, 2026

AI/ML Research Digest — Jun 20, 2026

1
Comments
3 min read
Sparse KV Caches Cut Attention Scaling

Sparse KV Caches Cut Attention Scaling

1
Comments
2 min read
Local Gradient Accumulation Speeds Training 1.7

Local Gradient Accumulation Speeds Training 1.7

Comments
2 min read
Intra‑Model Routing Accelerates Speculative Decoding

Intra‑Model Routing Accelerates Speculative Decoding

Comments
2 min read
Agent Harness Design Beats Model Tweaks

Agent Harness Design Beats Model Tweaks

Comments
1 min read
Two Diffusion Steps Reach 31 FPS

Two Diffusion Steps Reach 31 FPS

Comments
2 min read
Hamilton‑Jacobi View Links Major Neural Architectures

Hamilton‑Jacobi View Links Major Neural Architectures

1
Comments
2 min read
Retrieval‑Augmented Memory Reduces Sliding‑Window Limitations in Video Models

Retrieval‑Augmented Memory Reduces Sliding‑Window Limitations in Video Models

1
Comments
2 min read
Stateful Python Kernels Lift VLM Spatial Reasoning

Stateful Python Kernels Lift VLM Spatial Reasoning

2
Comments
2 min read
Aligning Hidden States Stabilizes LLM Distillation

Aligning Hidden States Stabilizes LLM Distillation

Comments
2 min read
8 FPS Real‑Time Video on Consumer GPU

8 FPS Real‑Time Video on Consumer GPU

Comments
2 min read
Optimal Transport Converts Dense Layers to Sparse Experts

Optimal Transport Converts Dense Layers to Sparse Experts

Comments
2 min read
AI/ML Research Digest — Jun 13, 2026

AI/ML Research Digest — Jun 13, 2026

Comments
3 min read
90% Less Memory Enables Infinite Video Generation

90% Less Memory Enables Infinite Video Generation

Comments
2 min read
Linear Ensembles Can Erase LLM Watermarks

Linear Ensembles Can Erase LLM Watermarks

Comments
2 min read
Benchmarks Evaluate Memory Quality and Adaptive Planning in LLM Agents

Benchmarks Evaluate Memory Quality and Adaptive Planning in LLM Agents

Comments
3 min read
AI/ML Research Digest — Jun 06, 2026

AI/ML Research Digest — Jun 06, 2026

1
Comments
3 min read
Raw waveform diffusion matches autoencoder quality

Raw waveform diffusion matches autoencoder quality

2
Comments
2 min read
Agents still fail 38% of real CLI tasks

Agents still fail 38% of real CLI tasks

Comments
2 min read
Agent compute drops substantially with online skill distillation and graph‑guided knowledge

Agent compute drops substantially with online skill distillation and graph‑guided knowledge

Comments
2 min read
Generative models now output simulation‑ready 3D assets

Generative models now output simulation‑ready 3D assets

Comments
2 min read
Verifiable rewards improve LLM math accuracy

Verifiable rewards improve LLM math accuracy

Comments
3 min read
ScientistOne achieves perfect citation verification

ScientistOne achieves perfect citation verification

Comments
3 min read
ThriftAttention keeps 90% quality with 5% compute

ThriftAttention keeps 90% quality with 5% compute

Comments
2 min read
AI/ML Research Digest — May 23, 2026

AI/ML Research Digest — May 23, 2026

Comments
4 min read
loading...