Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
benchmarks
Follow
Hide
Posts
Left menu
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
Why your SQLite WAL file never shrinks
Nobody
Nobody
Nobody
Follow
Sep 25
Why your SQLite WAL file never shrinks
#
sqlite
#
database
#
internals
#
benchmarks
Comments
Add Comment
12 min read
Wild beats mold linking Rust 20 times out of 20, and in release the linker is no longer the bottleneck
Efrain Garay
Efrain Garay
Efrain Garay
Follow
Sep 28
Wild beats mold linking Rust 20 times out of 20, and in release the linker is no longer the bottleneck
#
rust
#
linkers
#
benchmarks
#
performance
Comments
1
 comment
1 min read
Rust Coreutils 0.12 vs GNU 9.12: almost everything works the same, starting a process costs nearly 3x more
Efrain Garay
Efrain Garay
Efrain Garay
Follow
Sep 28
Rust Coreutils 0.12 vs GNU 9.12: almost everything works the same, starting a process costs nearly 3x more
#
rust
#
linux
#
benchmarks
#
performance
Comments
2
 comments
1 min read
TabPFN and TabICL against tuned XGBoost: the model that does not train won on fourteen tables out of fourteen
Efrain Garay
Efrain Garay
Efrain Garay
Follow
Sep 28
TabPFN and TabICL against tuned XGBoost: the model that does not train won on fourteen tables out of fourteen
#
machinelearning
#
benchmarks
#
datascience
#
python
Comments
1
 comment
1 min read
voiceloop: the fastest voice agent loop in the browser is now open source
Marcell Havlik
Marcell Havlik
Marcell Havlik
Follow
Sep 22
voiceloop: the fastest voice agent loop in the browser is now open source
#
voice
#
agents
#
benchmarks
#
opensource
2
 reactions
Comments
1
 comment
5 min read
Our task categorizer routed 46% of tasks. A $0.04/MTok decision model routes 97%.
Marcell Havlik
Marcell Havlik
Marcell Havlik
Follow
Sep 22
Our task categorizer routed 46% of tasks. A $0.04/MTok decision model routes 97%.
#
benchmarks
#
categorization
#
embeddings
#
llm
Comments
Add Comment
3 min read
94.7% on LoCoMo — and why most of the gap between published memory numbers isn't the memory
Marcell Havlik
Marcell Havlik
Marcell Havlik
Follow
Sep 22
94.7% on LoCoMo — and why most of the gap between published memory numbers isn't the memory
#
memory
#
benchmarks
#
agents
#
llm
Comments
Add Comment
9 min read
Fast Decisions in Agent Workflows: Laya vs TypeSafe Jev
x z
x z
x z
Follow
Sep 22
Fast Decisions in Agent Workflows: Laya vs TypeSafe Jev
#
ai
#
agents
#
benchmarks
#
opensource
Comments
Add Comment
5 min read
OCR that looked like it worked
Jakub Wietrzyk
Jakub Wietrzyk
Jakub Wietrzyk
Follow
Sep 21
OCR that looked like it worked
#
ocr
#
benchmarks
#
privacy
#
javascript
1
 reaction
Comments
1
 comment
6 min read
Run Qwen3.8 27B on Your Laptop, They Said. It Will Be FUN, They Said.
Tommy Leonhardsen
Tommy Leonhardsen
Tommy Leonhardsen
Follow
Sep 26
Run Qwen3.8 27B on Your Laptop, They Said. It Will Be FUN, They Said.
#
llm
#
benchmarks
#
bonsai
#
ternary
Comments
Add Comment
8 min read
How Postgres 19 checks foreign keys without running SQL
Nobody
Nobody
Nobody
Follow
Sep 24
How Postgres 19 checks foreign keys without running SQL
#
postgres
#
database
#
internals
#
benchmarks
1
 reaction
Comments
Add Comment
6 min read
DeepMind agents blew the whistle on cheating agents
techaiwire
techaiwire
techaiwire
Follow
Sep 14
DeepMind agents blew the whistle on cheating agents
#
aiagents
#
aisafety
#
benchmarks
#
google
5
 reactions
Comments
Add Comment
4 min read
Claude Formalized Fermat in 11 Days. The Math Isn't New.
Peremptory
Peremptory
Peremptory
Follow
Sep 7
Claude Formalized Fermat in 11 Days. The Math Isn't New.
#
anthropic
#
research
#
claude
#
benchmarks
Comments
Add Comment
3 min read
13 of 14 Models Write Messier Code Than the Human Who Fixed the Same Bug
DaisukeYoda
DaisukeYoda
DaisukeYoda
Follow
Sep 6
13 of 14 Models Write Messier Code Than the Human Who Fixed the Same Bug
#
python
#
ai
#
codequality
#
benchmarks
Comments
Add Comment
6 min read
Efficiency Hallucination: Every Model Rewrote Code That Couldn't Get Faster
Qasim Parray
Qasim Parray
Qasim Parray
Follow
Sep 19
Efficiency Hallucination: Every Model Rewrote Code That Couldn't Get Faster
#
llm
#
aicodingagents
#
codeoptimization
#
benchmarks
Comments
1
 comment
7 min read
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account