DEV Community

Machine Learning

A branch of artificial intelligence (AI) and computer science which focuses on the use of data and algorithms to imitate the way that humans learn, gradually improving its accuracy.

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Chain-of-Thought Faithfulness: Toggling 'Reasoning Mode' Made One Model 5x More Likely to Follow Its Own Mistakes

Kaggle Benchmarking Challenge Submission

Chain-of-Thought Faithfulness: Toggling 'Reasoning Mode' Made One Model 5x More Likely to Follow Its Own Mistakes

24
Comments 10
5 min read
8 LLMs, 480 Questions, 1 Kaggle Benchmark: Who Can Explain a Traffic Drop?

Kaggle Benchmarking Challenge Submission

8 LLMs, 480 Questions, 1 Kaggle Benchmark: Who Can Explain a Traffic Drop?

6
Comments 3
11 min read
AsyncGRPO: Eliminating GPU Idle Bubbles in Environment-Heavy RL Post-Training

AsyncGRPO: Eliminating GPU Idle Bubbles in Environment-Heavy RL Post-Training

Comments 1
9 min read
Promise Is Not Payment: Verification Errors That Amount Accuracy Misses

Kaggle Benchmarking Challenge Submission

Promise Is Not Payment: Verification Errors That Amount Accuracy Misses

Comments 1
6 min read
Laya: replace your LLM-as-a-judge with a 322M-parameter decision engine

Laya: replace your LLM-as-a-judge with a 322M-parameter decision engine

Comments 1
10 min read
Rule-based systems and machine learning aren't as different as people make them sound.

Rule-based systems and machine learning aren't as different as people make them sound.

Comments
2 min read
How I Took GLiClass FP8 from 59 ms to 16 ms on an RTX 4050

How I Took GLiClass FP8 from 59 ms to 16 ms on an RTX 4050

Comments
8 min read
Building an NVFP4 KV Cache for a Hybrid Qwen Model

Building an NVFP4 KV Cache for a Hybrid Qwen Model

Comments
9 min read
Julia 1: A 144M-Parameter Decision Model Trained for $104

Julia 1: A 144M-Parameter Decision Model Trained for $104

1
Comments 1
5 min read
How We Built a Web Agent That Remembers and Learns

How We Built a Web Agent That Remembers and Learns

Comments
5 min read
Owning vs. Renting Intelligence: Why Enterprises Are Building Sovereign AI

Owning vs. Renting Intelligence: Why Enterprises Are Building Sovereign AI

Comments 2
8 min read
I built a registry for System One models — here's what I learned comparing all of them

I built a registry for System One models — here's what I learned comparing all of them

Comments 2
1 min read
Running Qwen Flash-Next NVFP4 in vLLM: PLE Loading, B12x Fixes, and Stable Inference

Running Qwen Flash-Next NVFP4 in vLLM: PLE Loading, B12x Fixes, and Stable Inference

Comments 2
13 min read
What Happens When an AI System Is Built to Challenge Its Own Decisions?

Kaggle Benchmarking Challenge Submission

What Happens When an AI System Is Built to Challenge Its Own Decisions?

1
Comments
6 min read
Why our Amharic AI searches before it speaks

Why our Amharic AI searches before it speaks

Comments
3 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.