Blog
Writing on GPUs, LLMs, MLOps, Kubernetes — and mindset · 3525 posts
#2026-03 765#english 592#culture 264#deep-dive 254#kubernetes 249#career 229#ai 219#llm 209#devops 196#2026-04 146#security 143#database 114#observability 113#communication 109#history 107#architecture 100#productivity 96#finance 88#economy 84#mindset 81#psychology 80#ai-papers 79#food 78#it 78#travel 78#deep-learning 77#japanese 77#networking 77#performance 72#business-travel 70#linux 70#gpu 69#ai-agent 66#cs-fundamentals 63#postgresql 61#rag 58#self-improvement 55#learning 53#mlops 53#python 52
Reasoning Effort Is Not a Model Choice but a Per-Request Deployment Parameter
The DeepSeek V4 Flash 0731 results page published by ARC Prize carries not one score but three, one per reasoning effort level. This post computes what can actually be read out of those three numbers: that the same step
2026-08-09 · 8 min read #llm#benchmark#arc-agi#inference#costBus Factor Is Not the Number of People Who Know the Code but the Number of People Who Can Decide
The Nixpkgs core team disbanded after ten months. In a repository with thousands of contributors, the people holding delegated decision-making authority numbered two, and when those two stepped down that jurisdiction was
2026-08-09 · 9 min read #devops#open-source#governance#nix#supply-chainYou Added Registry Instances and Availability Did Not Move — What zot Scale-Out Actually Sells
Taking zot as the example, this post takes apart the scaling design of a container registry. A structure that assigns repositories to instances by consistent hashing and has non-owner nodes proxy the request onward is sh
2026-08-09 · 8 min read #devops#oci#registry#zot#containerThe Real Reason Web Editors Have No Ruler — The Ruler Is Not Missing, the Paper Is
Of everything Word has and web editors do not, the most frequently requested feature is the ruler. It is not that nobody has built one. It is that the web has no page model, so you cannot even settle what a centimetre is
2026-08-09 · 8 min read #frontend#rich-text-editor#ui-design#contenteditable#developer-toolsWhy Two Projects at the Same Company Reached Opposite Conclusions on AI Contributions ♪ Listenable
OpenJDK banned contributions made with generative AI outright in April 2026, while GraalVM, under the same Oracle umbrella, explicitly permitted the use of AI coding assistants around the same time. Both projects use the
2026-08-09 · 8 min read #culture#open-source#ai#policy#code-reviewThe 300x Is Not a Number You Get by Tuning PostgreSQL — The Volcano Model and Vectorized Execution
A precise dissection of the 300x published alongside the pgrust 0.2 release. That figure did not come from changing a PostgreSQL setting; it is a ClickBench measurement of a database newly implemented in Rust, while the
2026-08-09 · 9 min read #postgresql#database#performance#query-engine#simdLookup Cost Decides What You Can Read — From Two-Thousand-Year-Old Texts to Code Navigation
A site has appeared that gathers 1,060 Greek and Latin works and, when you click any word, shows its lemma, its morphological parsing, and its dictionary entry. The impressive part is not the volume of text but that the
2026-08-09 · 8 min read #developer-tools#code-navigation#ide#reading#toolingCoding Agent Spend Is Controlled by Friction, Not by Caps
Two engineering posts Databricks published back to back in July and August 2026 show that the handling of coding agent spend is moving away from budget caps and toward a gateway plus progressive friction. This post takes
2026-08-09 · 8 min read #mlops#llm#cost#ai-gateway#developer-productivityThe Claim of 100x Cheaper Is True Only When the Task Was Narrowed — Verification and Break-Even
A case study published in August 2026 reports that a 4-billion-parameter-class open model, post-trained with reinforcement learning, matched frontier models on a retrieval task while cutting per-request cost by an order
2026-08-09 · 8 min read #llm#cost#fine-tuning#retrieval#open-modelsIt Was Not the CPU That Was Slow but the Syscall Path — Lessons From Turning a Phone Into a Server
Through the case of turning a CMF Phone 1 into a personal infrastructure server, this post lays out where the real cost of a Linux compatibility layer actually lands. After failing to flash postmarketOS, the author kept
2026-08-09 · 9 min read #linux#termux#chroot#proot#self-hostingA Screen Is Written With Ten Letters — The One Question to Ask Before Building a Custom Widget
Jakob Nielsen has laid out the claim that almost every user interface is assembled from about ten elements, the way every English word is written with twenty-six letters. The list is useful not because it tells you what
2026-08-09 · 7 min read #ui-design#ux#design-system#frontend#usabilityWhen Two Services That Passed the Type Checker Halt Waiting for Each Other — Choreography as a Different Approach
Using a language that guarantees memory safety does not stop two services from halting while each waits for a message from the other, because the field of view of a type checker ends at one process. Choreographic program
2026-08-09 · 8 min read #programming-languages#compiler#distributed-systems#type-systems#concurrencyWhen the Artifact Gets Cheap, Where Does Assessment Move — Why Denmark Chose the Oral Defense
The Danish Ministry of Education has issued an immediate package that makes an oral defense mandatory for exam assignments written at home. When the cost of writing approaches zero, the submission alone tells you nothing
2026-08-09 · 8 min read #assessment#hiring#code-review#education#engineering-cultureWhen Friction Disappears, Taste Does Not Remain — the Path to Growing Taste Disappears
Taste Is All That Is Left, the essay that drew attention in August 2026, says that as making things got cheap, the only ability left scarce is judging what is worth making. This post agrees with the diagnosis and then go
2026-08-09 · 8 min read #career#craft#ai#code-review#mentoringYour Deploy Is the Load Test — What Happens When Nobody Designs the Cost of Seeding the Cache
A walk through how Canva moved the in-memory session revocation cache in its gateway from MySQL to S3. The problem was not the steady-state lookup cost but the startup cost of hundreds of pods seeding their caches simult
2026-08-09 · 8 min read #architecture#caching#s3#scalability#deploymentWhy the Claim That Code Was Never the Hard Part Makes People So Angry
An essay that reached the top of Hacker News in August 2026 argues that saying code was never the hard part is an insult to every programmer. This post agrees with the rebuttal but locates the cause somewhere else. That
2026-08-09 · 8 min read #career#craft#ai#engineering-culture#skillsData Residency Is a Replication Topology, Not a Dropdown
Using the document Fastmail published when it opened an EU data region on August 3, 2026 as a textbook, this post lays out why the location of your data is not settled by picking a region once. The primary copy, the repl
2026-08-09 · 9 min read #architecture#data-residency#gdpr#replication#complianceIn Eval-Driven Development, the First Thing to Calibrate Is the Judge
The eval-driven development retrospective Airbnb Engineering published in July 2026 is less a plea to write the eval set first than a plea to earn the right to treat the grading model as an instrument. This post lays out
2026-08-09 · 9 min read #ai#llm#eval-driven-development#llm-as-judge#evaluationWhat Does a QR Code with a Photo Inside Pay For It — Error Correction Is a Budget
Take apart the technique for putting a photograph inside a QR code and it turns out to be a budget allocation decision rather than a design decision, because the slack spent on making it pretty was set aside for crumpled
2026-08-09 · 8 min read #algorithm#qr-code#error-correction#dithering#image-processingWhat LLMs Cannot Do Is Not the Proof, It Is Setting Up the Premise
The ICML 2026 position paper Position: LLMs can not jump argues that generative AI has mastered induction and is rapidly conquering deduction, yet remains structurally unable to reach abduction, the act of producing a ne
2026-08-09 · 8 min read #ai#llm#reasoning#abduction#research