LLMs reward expertise

https://www.seangoedecke.com/og-image.jpg
Using large language models like LLMs requires domain knowledge to get the most out of them. A skilled prompter like Terence Tao can steer the model with expertise in the domain, but this skill is not easily replicable by following tips.

Ten advances in mathematics and theoretical computer science

Devtools must be open source

https://blog.exe.dev/static/og-card.png
Software engineers used to rarely write personal software, but now agents can easily manage customizing software, making it easier to personalize and maintain. Agents can build and manage custom software with minimal programming, allowing for more efficient and flexible use of software.

200 Milliseconds

https://200ms.thenodebook.com/og.png
See what really happens in the ~200ms after you press Enter: a single HTTP request traced in real time through DNS, TCP, TLS, the Linux kernel, Node's event loop, and Postgres — then back to the pixel. Scroll distance equals time.

Smaller, faster, safer: running Kimi and GLM at scale

https://blog.cloudflare.com/_image?href=https%3A%2F%2Fblog.cloudflare.com%2F_emdash%2Fapi%2Fmedia%2Ffile%2F01KZ1PVA5SC9J87PS1R2B97N1F.png&w=1999&h=1125&f=webp&fit=cover&position=center
Workers AI optimizes large models on GPUs by quantizing the KV cache to FP8, compressing model weights to INT4, and protecting the cache with integrity checks. These techniques enable efficient serving of models like Kimi K-series and GLM with no change in accuracy, supporting more customers at lower costs.

Celebrating 45 Years of Kermit with the First New C-Kermit Release in 15 Years

https://changelog.complete.org/.within.website/x/cmd/anubis/static/img/pensive.webp?cacheBuster=1.22.0
Please wait a moment while we ensure the security of your connection.

MiniMax H3 Day-0 Support in ComfyUI: Open Weights, Native Audio, and 2K Video

https://substackcdn.com/image/fetch/$s_!59Bc!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F98d2552b-d84c-4385-9602-bf344dca2ebe_1000x727.png
MiniMax H3 is a next-generation open-weights video model that generates 2K video with real stereo sound up to 15 seconds. It combines text, images, video, and audio inputs with a prompt to create a single video clip, reducing five tasks into one model.

The Dunning-Kruger effect may just be a data artefact (2020)

https://www.mcgill.ca/oss/files/oss/figure_1_3.png
The Dunning-Kruger effect may not be a real bias in human thinking as it can be replicated with random data and is likely due to measurement error and unreliability. The effect was originally described as people overestimating their competence in areas they are not skilled at, but this may be an artefact of how the data was measured.

Andy Pavlo joins ClickHouse to establish ClickHouse Labs

https://clickhouse.com/_next/image?url=%2Fuploads%2Fandy_pavlo_clickhouse_b978ca187a.png&w=750&q=75
I'm joining ClickHouse to lead ClickHouse Labs, a research team focused on advancing database technology. Our goal is to conduct impactful research and transform ideas into technology that benefits users, building on ClickHouse's strong foundation and exploring emerging AI and agentic technologies.

Replacing the Kobo Libra H2O Battery

https://ei3lh.eu/wp-content/uploads/2025/11/6b6b3f7e-9b96-4505-ab02-f2d7569cb651-1024x558.jpg
User had a Kobo Libra H2O with a faulty battery, which was difficult to replace due to lack of suitable options. They found a replacement battery on Ali Express and successfully replaced it with a soldering iron and basic tools.

How Hollywood stopped making movies in Hollywood

https://substackcdn.com/image/fetch/$s_!PU5R!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F4ed3afc1-5576-4677-aa1a-bf941416e9d7_1548x1024.webp
The entertainment industry in Los Angeles has been experiencing a production migration due to rising costs and tax incentives offered by other cities, leading to a decline in physical production in Hollywood. This migration has resulted in a widening gap between where stories are set and where they are filmed, with many productions choosing locations based on financial considerations rather ...

AI's debt binge can't last, hidden borrowing reaches $1.65T

https://fortune.com/img-assets/wp-content/uploads/2026/07/ai-bonds-073126-e1785507587847.png?format=webp&w=1440&q=100
Tech giants' massive debt issuance, driven by AI spending, may soon overwhelm investors. Hidden debt, estimated at $1.65 trillion, is not reflected in official bond markets, adding to the financial burden.

Massively Parallel Postgres Backups

https://planetscale.com/assets/massively-parallel-postgres-backups-social-apuT3KE7.png
PlanetScale's backup system for sharded databases uses parallelism to minimize production impact and achieve petabyte-scale backups in hours, with techniques like spinning up backup-specific nodes and replaying Write-Ahead Logs. The system uses object storage, WAL archiving, and hybrid replay to ensure consistent backups, and is designed to be transparent and easy to use, with automated ...

Launch HN: Hoplite (YC S26) – Effortlessly deploy cloud coding agents

https://hoplite.sh/opengraph.png
Effortless cloud coding agents that feel good to use.

Kelly Criterion Simulator

ZX Spectrum System Tour: Sound

The ZX Spectrum has a 1-bit beeper and a General Instrument AY-3-8910 synthesizer chip for sound production, with the latter available in later models and as a peripheral for earlier ones. Three programs are created to play a C-major scale using the beeper and the AY-3-8910 chip, with the latter providing more precise and flexible sound control.

AirLLM 70B inference with single 4GB GPU

https://raw.githubusercontent.com/lyogavin/airllm/main/assets/airllm_logo_sm.png
AirLLM is a library that reduces inference memory usage for large language models, allowing them to run on a single GPU card without quantization or pruning. It supports various models, including Llama, Qwen, DeepSeek, and more, and can be used with a simple one-line initialization.

DDoS against Norwegian government IT infrastructure – status

Digitaliseringsdirektoratet's Status Page - ID-porten, Kontakt- og reservasjonsregisteret, Maskinporten, MinID, eFormidling, ELMA, eInnsyn, Ansattporten, Selvbetjeningsløsninger, Altinn - Driftsproblemer.

SearXNG in Rust

https://raw.githubusercontent.com/MikeLuu99/searxng-rust/main/assets/tui_screenshots.png
A Rust-based metadata search engine aggregates results from multiple search engines, deduplicates and ranks them using Reciprocal Rank Fusion. It supports concurrent querying, HTML scraping, and a ratatui-based TUI interface.

The Billable Usage API: programmatic cost visibility for Cloudflare

https://blog.cloudflare.com/_image?href=https%3A%2F%2Fblog.cloudflare.com%2F_emdash%2Fapi%2Fmedia%2Ffile%2F01KZ1WZB6KJGKY7A4HBARJW87R.png&w=1999&h=1132&f=webp&fit=cover&position=center
Cloudflare launched a Billable Usage API for self-serve accounts, providing a single endpoint to return account usage and cost, broken down by product and service period. This API is designed for automation, allowing developers to access usage data in real-time, enabling FinOps teams to attribute costs to internal projects and teams.

KisakCOD – open-source reimplementation of Call of Duty 4 Multiplayer

https://raw.githubusercontent.com/SwagSoftware/KisakCOD/master/GPLv3_Logo.png
Copy COD4 game files to bin/(BUILD_TYPE)/* for development. Use AI for assistance, but remain responsible for code committed.

Use Task Runners for Common Coding Tasks

https://hamvocke.com/img/default-preview.jpg
The user wants to create a convenience tool to run common tasks across multiple code repositories with different tech stacks, such as installing dependencies, building, linting, formatting, running tests, and deploying. They discuss various options, including bash scripts, make, just, and mise, to create a task runner that simplifies their workflow and allows muscle memory to be applied to ...

Bonsai: Janestreet's UI Library

https://raw.githubusercontent.com/janestreet/bonsai/master/docs/assets/bonsai-logo.png
Bonsai is a UI library for building performant, reactive web applications in OCaml, inspired by Elm. It allows for composable state machines and incremental rendering, making it easy to manage state and UI updates.

Don't be a meat proxy

You're frustrated with relying on AI output in conversations and code reviews, feeling it lacks value and understanding. The effort of reading, validating, and writing a response in your own words is what adds value.

Prevent cognitive debt by manually retyping LLM-generated code

The user uses coding assistants to fast-forward through boring parts of projects, but manually types generated code to understand and adapt it, valuing comprehension over productivity. This workflow helps them build a mental model of their codebase and detect potential issues, allowing them to work faster and more efficiently.

Show HN: Product analytics (and evals) for agent sessions on your MCP

https://armature.tech/assets/logos/claude.svg
Armature captures user sessions with AI agents, grouping them into use cases and ranking by volume and success rate. It detects failures, redacts PII and secrets, and scores sessions for user satisfaction, with free plans up to 1,000 sessions a month.

C++ float-to-int conversion can be undefined behavior

Converting a float to an int in C++ is undefined behavior when the value cannot fit into the destination type, even with -Wall and -Wextra. The correct fix is to bounds check before casting, as relying on undefined behavior can lead to different results on different hardware.

SQLite Critical CVEs or LLM Slop?

https://research.jfrog.com/img/RealTimePostImage/post/sqlite-critical-cves-or-llm-slops/image1.png
A GitHub account published 50+ SQLite vulnerability advisories, but JFrog security researchers found that 54 were fabricated and one contained a real bug with unverified metadata. This incident highlights a systemic issue with automated vulnerability ingestion, where plausible-sounding fake advisories can cause organizations to waste time investigating and patching non-existent vulnerabilities.

Rust project goals: Immobile types and guaranteed destructors

https://opengraph.githubassets.com/33d3cf50b3908ad51294f7ade72bf32057f24a694ef2052ca1c523d336cff23a/rust-lang/rust-project-goals
Rust proposes introducing new traits like Move and Forget to make explicit what operations are possible on a type, allowing types to opt out of being moved or forgotten. This change aims to simplify immovable types and enable safe scoped spawn for async, while eventually deprecating the Pin trait.

Explanation of INT8 ConvRot (FP8 is no longer needed)

https://assets.st-note.com/production/uploads/images/291809558/rectangle_large_type_2_a5df5f5d3cda1717c41284edb25c5b5c.png?width=1280
ComfyUI v0.27.0 supports INT8 ConvRot, a modeling and quantization method that provides performance exceeding FP8 and FP8 Scaled formats on GeForce RTX 40/50 series. INT8 ConvRot is expected to become the standard for 8-bit quantized models, offering benefits on GeForce RTX 20/30 series and surpassing NVFP4 in speed on RTX 50 series.