📰 Tech Trends Daily — Monday, August 3, 2026
🔥 Today’s Focus
Monday’s front page belongs to Karpathy, who declared the classic “pelican on a bicycle” image benchmark graduated (375 points, today’s highest on HN) — a hint that the test Simon Willison popularized back in 2024 has finally been cracked by models. The comments split on cue: one camp points out that the pelicans AI still draws make elementary mistakes — both legs on the same side of the bicycle, sometimes two beaks — and calling it solved is self-deception; the other counters that the debate has devolved into “which pelican looks better,” which is itself the inflection-point evidence that models no longer fail in obvious ways. Both lines actually hold: the benchmark’s discriminating power has genuinely dried up, but “solved” is another matter. The second thread is the more interesting one: Lean’s kernel soundness bug #14576 was surfaced by an LLM “disproving” the Collatz conjecture — the counterexample the model claimed turned out to be an artifact of a hole in the prover’s own kernel. Formal verification was supposed to end AI hallucination; instead, the verifier’s own chain of trust broke first. The third thread is a nostalgia eruption: NetBSD 11.0 officially released (top of Lobsters at 99 points), RISC OS Open’s twentieth anniversary, an autoregressive language model running on a 6502, CP/M-386 revived — old platforms are becoming a refuge in the AI era.
🤖 AI & LLM
- Karpathy’s Pelican — Karpathy’s pelican. 375 points / 296 comments (HN). Today’s highest-scoring post: Karpathy hints that the “pelican on a bicycle” benchmark has been conquered. 💬 Morromist opens fire — the pelicans AI still draws have both legs on the same side of the bicycle and sometimes two beaks; “switching tests” isn’t “passing tests.” jonas21 counters that the argument has become a taste contest over “which one looks better,” which is itself the signal that models no longer fail in obvious ways.
- My personal AI benchmark: “Generate an SVG of a frog with a Habsburg jaw.” — A frog with a Habsburg jaw, in SVG. 76 points / 42 comments (HN). The ultimate personal benchmark — asking a model to render an SVG frog with the Habsburg dynasty’s signature jaw, combining SVG chops with obscure history. More personality than any benchmark suite.
- Autoregressive Language Model on the 6502 Processor — An autoregressive LLM on a 6502. 14 points / discussion ongoing (HN). BitNet-style quantization scaled down to an 8-bit CPU — an autoregressive model running on a 6502. Performance is beside the point; “it runs” is the hardcore romance.
- I’m (mostly) picking models on speed now, not intelligence — Speed over intelligence. △16 / 9 comments (Lobsters). Opinion piece: for everyday tasks, intelligence gaps between models barely matter — latency is the real experience divider. It echoes the Karpathy pelican thread: evaluation is shifting from “can it?” to “how fast?”
- Mathematics Without Mathematicians — Mathematics without mathematicians. △10 / 8 comments (Lobsters). Deconstructing the proposition that mathematics no longer needs mathematicians in the LLM era — AI can accelerate computation and verification, but asking the right questions is still human work.
- Prevent cognitive debt by manually retyping LLM-generated code — Retype it to understand it. △25 / 12 comments (Lobsters). The contrarian advice: retype LLM-generated code by hand before committing, trading muscle memory for understanding. The comments split roughly evenly between “performance art” and “it actually works.”
- Ten advances in mathematics and theoretical computer science — Ten advances in math and TCS. △13 / 1 comments (Lobsters). OpenAI’s rundown of its models’ mathematical progress — landing the same day as the Lean soundness bug postmortem; the timing is delicate.
- 23 languages, one I can check — 23 languages, one I can check. △1 / 2 comments (Lobsters). A project written in 23 languages, of which the author can review exactly one — the self-awareness problems of the vibe-coding era.
🛠️ Tools, Infrastructure & Developer Experience
- Meshdiff – visually compare two STL versions in the browser, client-side — Meshdiff: client-side STL diff. 170 points / 17 comments (HN). A visual diff for 3D-printing files, running entirely client-side. 💬 The comment consensus is unambiguous: synchronized rotation across the three viewports, GitHub PR integration for reviewing 3D files, and a CLI version for CI — all already on the author’s roadmap; the demand side has prioritized it for him.
- Show HN: MicroCodex Coding Agent — MicroCodex: a coding agent under 1MB. 11 points / 3 comments (HN). A coding agent that fits under 1MB — set against the hundreds-of-MB Electron kitchen sinks, it at least wins the “light” contest.
- Adding Go’s Defer to the TypeScript Compiler — defer for TypeScript. 13 points / 4 comments (HN). An experiment adding a defer keyword to the TypeScript compiler — entry-level compiler theory, clearly explained; hacking on compilers beats business code any day.
- Fasttracker II clone in C using SDL 2 — Fasttracker II, remade in C. 108 points / 36 comments (HN). An open-source remake of the classic tracker music software — the comments are uniformly “my youth is back.”
- FlickBoard — FlickBoard. △11 / 2 comments (Lobsters). A flick input method for Android — old-school gesture typing resurrected on the touchscreen.
- EPIPE on write might mean you’re doing it wrong — EPIPE means you’re doing it wrong. △18 / 4 comments (Lobsters). Classic rachelbythebay: EPIPE is not a random failure — it means you misjudged the pipe’s lifecycle. Unix-semantics education.
- The Fixi Project — The Fixi Project. △24 / discussion ongoing (Lobsters). A new project on Lobsters, tagged practices — the community keeps inventing fresh vehicles for engineering practice.
- Note-Taking and Personal Knowledge Management — PKM, again. 94 points / 26 comments (HN). Meta-discussion of PKM — why note systems always get built and then abandoned; a few of the 26 comments are sharper than the article itself.
- Developers are attached to tools because tools encode trust — Tools encode trust. 118 points / 58 comments (HN). Stack Overflow’s blog on tool loyalty — the 58-comment consensus: switching tools never costs you the learning curve, it costs you re-evaluating “can I trust this tool?”
🔒 Security & Privacy
- Rooting, firmware analysis and persistent credentials of TP-Link TL-841N — TP-Link TL-841N: root, firmware, persistent creds. 74 points / 14 comments (HN) + △6 / 1 comments (Lobsters). Old-router firmware reverse engineering: hardcoded credentials that survive even a factory reset — the security debt of cheap IoT devices, sold by the unit.
- Harvesting SSH Credentials: Insights from My Honeypot Network — SSH credential harvesting. 32 points / 24 comments (HN). Honeypot telemetry: SSH brute-forcing happens every minute across the internet, and the harvested credentials are better quality than you’d expect — the comments duly remind everyone that “fail2ban is just a placebo.”
- EU Age Verification Project Mandates Hardware-Bound Attestation — EU age verification goes hardware-bound. 85 points / 35 comments (HN). The EU’s age-verification scheme demands hardware-level attestation — browsers and device vendors get drafted as the “age police.” The comments argue that hardware attestation will end up as device-fingerprinting surveillance.
- Californians’ data deletion requests, DROP, become enforceable Aug. 1 — DROP becomes enforceable. 26 points / 1 comments (HN). California’s deletion-right requests enter the enforceable phase — regulation landing is only the beginning; the real question is whether companies can actually find and delete all of your data.
- ‘Crush this lady’: how eBay harassment campaign led to $56M payout — The eBay harassment payout. 137 points / 60 comments (HN). FT long read on the harassment campaign orchestrated by eBay executives — from live maggots mailed through social media to threat letters. How low can tech-executive judgment go? All 60 comments are “this actually happened.”
💻 Systems & Platforms
- NetBSD 11.0 released — NetBSD 11.0. △99 / 18 comments (Lobsters). Top of Lobsters today. 💬 Comment highlights: the team candidly owns the “vulnpocalypse” — AI tooling has exploded the volume of vulnerability reports, so rather than delay the release indefinitely they shipped transparently with open issues. New RISC-V support with 20,000+ pkgsrc packages for riscv64; one user switched from Linux to NetBSD as a daily driver over Linus’s comments about LLMs, and another bought a RISC-V dev board just for it.
- Twenty Years of RISC OS Open — RISC OS Open at 20. 143 points / 27 comments (HN) + △5 / 0 comments (Lobsters). Twenty years of the Acorn-era OS as open source. 💬 Old hand nickcw recalls writing his !Director app entirely in ARM assembly; someone notes that the scorewriting software Sibelius was born on RISC OS — it has outlived the “museum project” label most people give it.
- Show HN: Kakehashi – Experimental userspace to run macOS binaries on Linux ARM — Kakehashi: macOS binaries on Linux ARM. 152 points / 35 comments (HN). A “Wine for macOS”: Mach-O binaries running in userspace. 💬 Author vlad_kalinkin replies in person — 7-Zip works (5.2× slower than native, with an optimization roadmap already in place), curl passes 200+ commands in automated tests, Xcode’s basic Git commands work; the long-term goal is full Xcode Tools including iOS builds, plus macOS Homebrew. Commenters compare it to Darling; the author states plainly it is not a fork.
- Show HN: NixOS-DGX-Spark — NixOS on DGX Spark. 81 points / 22 comments (HN). NixOS on NVIDIA’s DGX Spark personal AI workstation — the declarative system taking over AI hardware. The comments chew on the friction between CUDA’s closed drivers and Nix.
- Resigning from Arch Linux — Foxboron resigns. △44 / 4 comments (Lobsters). Foxboron (core Arch security team, sbctl author) announces his departure — read alongside yesterday’s Arch decision to disable AUR package adoption, the two faces of governance pressure. 💬 novedeo: “the strain on the Arch security team and AUR maintainers has been visible these past months.”
- Sharing an X11 Server Across Hosts with FamilyWild — One X server, many hosts. 23 points / 5 comments (HN). Sharing a single X11 server across multiple hosts — X still pulls off new tricks in the Wayland era.
- MkLinux and the pimped-out Apple Workgroup Server 9150 — MkLinux archaeology. △12 / 3 comments (Lobsters). Nostalgia archaeology of Apple’s official ’90s Linux port — oldvcr.blogspot, with its usual impeccable images.
- CP/M-386: CP/M for 386 protected mode, derived from CP/M-68K — CP/M for the 386. △2 / discussion ongoing (Lobsters). An archaeology project porting CP/M to 386 protected mode — landing the same day as the 6502 LLM, the two poles of retrocomputing.
- But can your calculator run Linux? — Linux on a calculator. △27 / 2 comments (Lobsters). raymii hacks a graphing calculator to run Linux — “can it run Linux” has become the universal yardstick for hardware.
- Can you make a Wii U gamepad from a Raspberry Pi? — A Wii U gamepad from a Pi. △2 / discussion ongoing (Lobsters). Hardware-hacking video — once the Wii U GamePad’s screen-streaming protocol was reverse engineered, a Raspberry Pi became the cheap replacement.
💬 Programming Languages
- F*: A general-purpose proof-oriented programming language — F*, the proof-oriented language. 139 points / 61 comments (HN). Microsoft Research’s dependently typed language hits the front page. 💬 The top comment is a complaint: five pages in and still no code example — “for a new language I want two things: what the syntax looks like and why I should use it.” The marketing problem of formal-verification languages is more urgent than the technical one.
- Postmortem for Lean Kernel Soundness Bug #14576 — Lean kernel soundness postmortem. △48 / 3 comments (Lobsters). The postmortem, written by Leo de Moura himself. 💬 The backstory is wilder than the bug: it was surfaced by an LLM “disproving” the Collatz conjecture — the claimed counterexample turned out to be an artifact of the kernel flaw. The author knew the bug existed beforehand but declined to say whether the proof had been constructed against it; the bug hit both the main kernel implementation and the independent checker, breaking the independence assumption.
- sizeof is surprisingly difficult to parse in c — Parsing sizeof. △45 / 31 comments (Lobsters). sizeof’s type/expression ambiguity is the bane of every C parser. 💬 david_chisnall supplies the history: C never started with a formal grammar — syntax was “whatever the compiler accepted” — so C89 had to define parsing rules to accommodate existing code, and memory-constrained early compilers spawned a family of state-saving parser hacks. Today’s complexity is a legacy of 1970s hardware limits.
- Guarded methods in OCaml — Guards in OCaml. △11 / 1 comments (Lobsters). Language-design notes on expressing guard patterns in OCaml’s type system — everyday one-upmanship in the functional community.
- An old-new take on argument parsing in Rust — getopt-style Rust. △28 / 7 comments (Lobsters). jmmv argues for a return to old-school getopt-style design, challenging clap’s near-monopoly — cut from the same cloth as yesterday’s rand fork post: the Rust ecosystem is starting to rebel against “the official library is too big.”
- How fast is C++26’s std::hive? — std::hive, benchmarked. △9 / 3 comments (Lobsters). Daniel Lemire benchmarks the new container — how much hive wins in cache-unfriendly scenarios, with data.
- Faster floating point math with Rust’s new API — Faster floats in Rust. △6 / 4 comments (Lobsters). Benchmarks of Rust’s new floating-point API — a micro-optimization guide for performance-sensitive code.
- Atom is better than RSS, in ways that matter — Atom vs. RSS. △64 / 53 comments (Lobsters). A format war on Lobsters. 💬 emk — who worked with Dave Winer on RSS back in the day — writes a long retrospective: RSS 1.0’s RDF-ification complicated a simple format until nobody could implement it correctly, and Shirky’s predictions all came true. The proposal to “rename Atom to RSS 3” gets doused: the name is taken, and the Winer camp would fight over the trademark.
📚 Light & Fun
- Folding Paper Globes — Folding paper globes. 134 points / 28 comments (HN). Printable paper-globe templates — the Earth unfolded into mathematically foldable planes; the 28 comments are all “printed it, looks great.”
- When transit passes were designed by hand (2022) — Hand-designed transit passes. 87 points / 28 comments (HN). Letterform Archive’s Milwaukee transit-pass collection — a neglected corner of print-design history.
- The Myth of Snow Leopard — The Snow Leopard myth. 28 points / 23 comments (HN). Taking apart the “Mac OS X 10.6 was the perfect system” nostalgia — the comments dig into history while arguing over whether Snow Leopard was actually fast.
- Read the Novels and Forget Everything Else — Read the novels. 21 points / 5 comments (HN). Hedgehog Review’s literary exhortation — a case for reading long novels in an age of information overload.
- Getting Started with Google Wave (2010) — Google Wave, archived. △12 / 7 comments (Lobsters). Archaeology of Google Wave’s official tutorial video — the most famous “peaked at launch” case in internet product history.
- Where’s your website? — Where’s your website? △21 / 19 comments (Lobsters). Another manifesto for the personal-website revival — half of the 19 comments are showing off their own homepages.
- Show HN: Make your Framework 12 sound like a creaky door — A creaky Framework 12. 33 points / 2 comments (HN). An add-on that makes the repairable laptop creak — hardware-hacker humor; both comments are “why.”
- TinyNES Review – A Super Niche NES Console — TinyNES review. 16 points / 1 comments (HN). Lon.TV’s review of TinyNES — a 2023 video back on the front page; retro-hardware heat shows no sign of cooling.
- How the words we teach English language learners changed — The evolving English-teaching lexicon. Low-score newcomer (HN). Pudding’s data visualization: how the core vocabulary of English textbooks has shifted over decades — exemplary data journalism that barely made a ripple on HN.
📝 Summary
Today’s through-line: scrutiny of AI is shifting from results to the chain of trust. Karpathy graduated the pelican benchmark and more people are picking models by speed — benchmark anxiety is receding. But the Lean kernel soundness bug surfaced by an LLM’s Collatz disproof, and NetBSD’s own team coining “vulnpocalypse” for the AI-driven flood of vulnerability reports, both remind us: the trustworthiness of model output depends on whether the people and tools verifying it can themselves be trusted. Must-read Top 3: ① Karpathy’s Pelican post (375 points — watch the two camps argue about benchmark saturation); ② Lean Soundness Bug #14576 postmortem (Lobsters △48 — a first-hand incident report from the formal-verification trust chain); ③ NetBSD 11.0 release (Lobsters △99 — RISC-V support and the candid vulnpocalypse release strategy). Cross-cutting signal: retrocomputing (the 6502 LLM, CP/M-386, MkLinux, RISC OS at 20) shares front pages with the AI wave — vintage hardware is becoming the spiritual escape hatch for AI-fatigued readers.