Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
benchmarks
Follow
Hide
Posts
Left menu
đ
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
Claude Sonnet 5.5 nears Opus 5.5 at half the price
techaiwire
techaiwire
techaiwire
Follow
Oct 3
Claude Sonnet 5.5 nears Opus 5.5 at half the price
#
anthropic
#
llm
#
benchmarks
#
pricing
5
 reactions
Comments
Add Comment
4 min read
How to read a Blender cloud render benchmark: GPU seconds are not the round trip
ć è€æ §ç„
ć è€æ §ç„
ć è€æ §ç„
Follow
Oct 2
How to read a Blender cloud render benchmark: GPU seconds are not the round trip
#
blender
#
rendering
#
benchmarks
#
gpu
Comments
Add Comment
4 min read
Gemini 4 Argon Wins 12 of 18 Benchmarks. Code Isn't One.
Max Quimby
Max Quimby
Max Quimby
Follow
Oct 2
Gemini 4 Argon Wins 12 of 18 Benchmarks. Code Isn't One.
#
ai
#
llm
#
google
#
benchmarks
Comments
Add Comment
10 min read
Why your SQLite WAL file never shrinks
Nobody
Nobody
Nobody
Follow
Sep 25
Why your SQLite WAL file never shrinks
#
sqlite
#
database
#
internals
#
benchmarks
Comments
Add Comment
12 min read
Wild beats mold linking Rust 20 times out of 20, and in release the linker is no longer the bottleneck
Efrain Garay
Efrain Garay
Efrain Garay
Follow
Sep 28
Wild beats mold linking Rust 20 times out of 20, and in release the linker is no longer the bottleneck
#
rust
#
linkers
#
benchmarks
#
performance
Comments
1
 comment
1 min read
Rust Coreutils 0.12 vs GNU 9.12: almost everything works the same, starting a process costs nearly 3x more
Efrain Garay
Efrain Garay
Efrain Garay
Follow
Sep 28
Rust Coreutils 0.12 vs GNU 9.12: almost everything works the same, starting a process costs nearly 3x more
#
rust
#
linux
#
benchmarks
#
performance
Comments
2
 comments
1 min read
TabPFN and TabICL against tuned XGBoost: the model that does not train won on fourteen tables out of fourteen
Efrain Garay
Efrain Garay
Efrain Garay
Follow
Sep 28
TabPFN and TabICL against tuned XGBoost: the model that does not train won on fourteen tables out of fourteen
#
machinelearning
#
benchmarks
#
datascience
#
python
Comments
1
 comment
1 min read
voiceloop: the fastest voice agent loop in the browser is now open source
Marcell Havlik
Marcell Havlik
Marcell Havlik
Follow
Sep 22
voiceloop: the fastest voice agent loop in the browser is now open source
#
voice
#
agents
#
benchmarks
#
opensource
2
 reactions
Comments
1
 comment
5 min read
94.7% on LoCoMo â and why most of the gap between published memory numbers isn't the memory
Marcell Havlik
Marcell Havlik
Marcell Havlik
Follow
Sep 22
94.7% on LoCoMo â and why most of the gap between published memory numbers isn't the memory
#
memory
#
benchmarks
#
agents
#
llm
Comments
Add Comment
9 min read
Our task categorizer routed 46% of tasks. A $0.04/MTok decision model routes 97%.
Marcell Havlik
Marcell Havlik
Marcell Havlik
Follow
Sep 22
Our task categorizer routed 46% of tasks. A $0.04/MTok decision model routes 97%.
#
benchmarks
#
categorization
#
embeddings
#
llm
Comments
Add Comment
3 min read
Fast Decisions in Agent Workflows: Laya vs TypeSafe Jev
x z
x z
x z
Follow
Sep 22
Fast Decisions in Agent Workflows: Laya vs TypeSafe Jev
#
ai
#
agents
#
benchmarks
#
opensource
Comments
Add Comment
5 min read
OCR that looked like it worked
Jakub Wietrzyk
Jakub Wietrzyk
Jakub Wietrzyk
Follow
Sep 21
OCR that looked like it worked
#
ocr
#
benchmarks
#
privacy
#
javascript
1
 reaction
Comments
1
 comment
6 min read
Run Qwen3.8 27B on Your Laptop, They Said. It Will Be FUN, They Said.
Tommy Leonhardsen
Tommy Leonhardsen
Tommy Leonhardsen
Follow
Sep 26
Run Qwen3.8 27B on Your Laptop, They Said. It Will Be FUN, They Said.
#
llm
#
benchmarks
#
bonsai
#
ternary
Comments
Add Comment
8 min read
How Postgres 19 checks foreign keys without running SQL
Nobody
Nobody
Nobody
Follow
Sep 24
How Postgres 19 checks foreign keys without running SQL
#
postgres
#
database
#
internals
#
benchmarks
1
 reaction
Comments
Add Comment
6 min read
DeepMind agents blew the whistle on cheating agents
techaiwire
techaiwire
techaiwire
Follow
Sep 14
DeepMind agents blew the whistle on cheating agents
#
aiagents
#
aisafety
#
benchmarks
#
google
5
 reactions
Comments
Add Comment
4 min read
đ
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account