GLM-5.3: Frontier coding with emergent cyber capabilities

DeepSeek peak/off-peak pricing update

https://api-docs.deepseek.com/img/v4_260813_benchmark_table_en.png
V4-Pro and V4-Flash models have varying reasoning efforts for tasks, with Expert Mode available on app/web. API pricing is updated with peak and off-peak rates, offering more flexible workload scheduling options.

For the love of god stop using CPU limits in Kubernetes

https://raw.githubusercontent.com/inevolin/k8s-cpu-limits-analyzed/main/assets/01-request-vs-limit.png
Removing CPU limits from Kubernetes clusters can improve performance, reduce costs, and increase efficiency by allowing apps to use idle CPU resources more effectively. By right-sizing requests after removing limits, teams can achieve significant gains in p99 latency (up to 87% lower), throughput (10% higher), and cost savings (tens of thousands of dollars per year).

Major oil slick washes up on Iran coast after Hormuz ship strike

https://ichef.bbci.co.uk/news/480/cpsprodpb/d0d6/live/6acf5150-97c1-11f1-b2ab-0dd01740f9f6.png.webp
A major oil slick has washed up on Iran's Qeshm Island, believed to be from a ship hit by a strike in the Strait of Hormuz. The spill threatens wildlife and ecosystems around the island, which is home to mangrove forests and coral reef habitats.

Everyone talks about AI agents. This is what one looks from the inside

https://pssah4.github.io/vault-operator/assets/og-image.png
Vault Operator integrates AI directly into your knowledge base, enabling it to read, act on, and refine your notes. It offers deep ingestion, semantic search, and automated document creation while maintaining control and security.

Gemini 3.7 Flash

https://storage.googleapis.com/gweb-uniblog-publish-prod/images/gemini-3-7-flash__evals__perform.width-1200.format-webp.webp
Gemini 3.7 Flash is the latest model, offering substantial improvements in coding and agent tasks with a lower price point than Gemini 3.6 Flash. It delivers strong gains in debugging, issue resolution, and code accuracy across various workflows and domains.

Accelerating GPT-5.6 Sol Ultrafast

https://cdn.sanity.io/images/e4qjo92p/production/96417454c2234a5f3241876a06719db0d27fae1e-1273x910.png?auto=format&dpr=2&fit=max&q=75&w=1273
Cerebras and OpenAI are launching Ultrafast Mode, a new service tier powered by Cerebras that delivers up to 750 output tokens per second without quality compromise. This resolves the tradeoff between speed and intelligence in AI models, enabling faster inference for mission-critical work.

Hello, me. It's been a while

The author reflects on how they've lost the habit of quiet thinking over 14 years, replaced by constant background noise. They rediscovered this habit and found it to be more enjoyable than filling silence with podcasts or social media.

Show HN: C# Game Engine with its own scripting language and IDE

https://private-user-images.githubusercontent.com/160407755/631850586-26066d82-c607-4b22-9d47-4abebf85d8a3.png?jwt=eyJ0eXAiOiJKV1QiLCJhbGciOiJIUzI1NiJ9.eyJpc3MiOiJnaXRodWIuY29tIiwiYXVkIjoicmF3LmdpdGh1YnVzZXJjb250ZW50LmNvbSIsImtleSI6ImtleTUiLCJleHAiOjE3ODY2OTI0OTMsIm5iZiI6MTc4NjY5MjE5MywicGF0aCI6Ii8xNjA0MDc3NTUvNjMxODUwNTg2LTI2MDY2ZDgyLWM2MDctNGIyMi05ZDQ3LTRhYmViZjg1ZDhhMy5wbmc_WC1BbXotQWxnb3JpdGhtPUFXUzQtSE1BQy1TSEEyNTYmWC1BbXotQ3JlZGVudGlhbD1BS0lBVkNPRFlMU0E1M1BRSzRaQSUyRjIwMjYwODE0JTJGdXMtZWFzdC0xJTJGczMlMkZhd3M0X3JlcXVlc3QmWC1BbXotRGF0ZT0yMDI2MDgxNFQwNzIzMTNaJlgtQW16LUV4cGlyZXM9MzAwJlgtQW16LVNpZ25hdHVyZT0xMTQ1OTlkNDNkYzIzZTAxYTc2NDQ5MTFhOTM3ZDc2MTcwMGZhMDE1MTI0NzZiZDVkZWJlMThlNmYwN2U3ZTY4JlgtQW16LVNpZ25lZEhlYWRlcnM9aG9zdCZyZXNwb25zZS1jb250ZW50LXR5cGU9aW1hZ2UlMkZwbmcifQ.ud4HPlFLpbbvlUo1PDFH2ifr_AexLsPfsuBIYt9Laoc
ArcadeMaker is an open-source 2D game engine with its own programming language, built on top of MonoGame and GameMaker 8. It aims to be a beginner-friendly tool for creating games without relying heavily on AI-generated code.

Why does Opus 5 feel worse to work with?

The author believes Opus 5 is a capable model but feels like a downgrade due to its tendency to make bold assumptions in the face of ambiguity, prioritizing benchmark performance over nuanced understanding. This approach may not be suitable for real-life applications where agents need to ask questions and seek clarification.

Differential Heuristics

https://www.redblobgames.com/blog/2026-08-08-differential-heuristics/denerim-heuristic-improvement.png?2026-08-08-11-26-18
The author studied A* pathfinding after discovering Google Maps' efficient algorithm in 2007, but struggled to explain differential heuristics. After years of learning and experimentation, they created a new interactive guide on the topic.

Show HN: Lumabri – Run Moe Models on a P2P Swarm with Colibri

https://opengraph.githubassets.com/ae5b93438ebfc2e21b91f9f58da2cec9cbad9a1999bc21ce382ede2e56944118/JustVugg/lumabri
A swarm of peers runs huge mixture-of-experts models using the colibri engine, with no dependencies and pure C code. The network verifies bytes and results through local mirroring and peer-to-peer communication.

DeepSeek Harness developer preview

https://deepseek.com/harness/images/harness/feat-plugin.en.png
DeepSeek Harness is now in developer preview, offering a plugin-based system for agent developers to create custom harnesses. It provides various modes, including standard, code, minimal and creator modes with tools and capabilities that can be swapped or extended.

Ruby 4.0 Universal RCE Deserialization Gadget Chain

https://cdn.prod.website-files.com/6971f0e051b588235e8acf7b/6a7e978fe0c1c71cd64e1029_gadget-reuse-before-after_1.png
A collective of AI agents exploited Ruby deserialization to gain admin control of a cluster, using a previously published universal RCE chain. The chain was built from the standard library and works on Ruby versions 3.3-4.0.6 without requiring any gems or application code.

Spaghettifying DRAM

https://raw.githubusercontent.com/xoreaxeaxeax/skitter-creek-bath-salts/main/examples/unspaghettify.gif
The skitter-creek-bath-salts project exploits a vulnerability in AMD Family 16h CPUs by manipulating the DRAM controller to scramble memory, bypassing security features like SEV and SGX. This allows for unrestricted access to protected regions of memory, including PSP private memory and SMRAM.

Mistral OCR 4.1

https://docs.mistral.ai/_next/image?url=%2Fassets%2Fsprites%2Fcat_idle.gif&w=640&q=75
Our latest OCR service powering our Document AI stack, with native paragraph-level bounding box extraction, structural block labels, and block-level confidence scores.

Bluesky Protocol Services

https://atproto.com/-/blog/introducing-bluesky-protocol-services/opengraph-image.png?d2b4677fa75ef34c
Bluesky Protocol Services launches, organizing documentation for developers using the AT Protocol network's public infrastructure. The new site includes Jetstream v2, which adds history to the network and allows server-side slicing without backfilling locally.

We're not done with point clouds

https://claytonwramsey.com/blog/mvt/panda.png
Researchers Ching Chen and Tsung-Tai Yeh improved a data structure called CAPT for collision-checking against point clouds, making it faster and cheaper in memory.

Understanding is the new bottleneck

https://www.geoffreylitt.com/images/talks/understanding-bottleneck/slide-01.webp?1783098284
The talk discusses the importance of human understanding in systems built by agents, and provides techniques for efficiently understanding code, including explain-diff and shared spaces. These methods aim to bridge the gap between humans and AI, enabling active participation in creative processes.

Protect Your Relays

https://www.iroh.computer/api/og?title=Blog&subtitle=Protect%20your%20relays
When two devices can't get a direct connection, a relay carries the connection so data still flows. If the relay accepts anyone, then anyone who learns its URL can push traffic through it. And they will learn it: it ships inside every client you distribute and it's visible to anyone watching a connection get established. Because of this, we've decided that managed relays on Iroh ...

Choose Boring Technology (2015)

https://i.imgur.com/FRQKLCy.jpg
The author reflects on the importance of technology choices in a company, emphasizing that innovation tokens should be spent wisely and that "boring" technologies like MySQL or PHP can be good enough. They advocate for mindful choice of technology to avoid overwhelming unknowns and operational costs.

Donkey.bas is 45 Years Old – 131 line of Glory

DONKEY.BAS shipped with early IBM PC DOS as a demo of color graphics and sound in BASICA. It was written by Microsoft co-founder Bill Gates and Neil Konzen in 1981 (version 1.10 in 1982). You only switch lanes — avoid the donkey, or BOOM. This page recreates the original CGA gameplay in JavaScript with some liberties. Original source: DONKEY.BAS · GitHub

What an improv stage can teach you about leading cross-cultural teams in Tokyo

https://assets.tokyodev.com/tokyodev-production-r2/variants/q2e92zkkbfd6jnhr0lzcdgwh05ff/f71b87b38636b7591ceb09953914ef235c0903dc949c343b7f802d202ec96b92
The author, a foreign product manager in Tokyo, shares four habits he learned from improv comedy that help him manage uncertainty and build strong teams. He emphasizes the importance of listening, reframing pushback as an offer, naming credit to team members publicly, and sharing his own mistakes first.

Nine PBS sues Iron Mountain over blocked access to archival data

https://current.org/wp-content/uploads/2026/08/Denver_CO_City_and_County_Building_IMG_5541-crop-scaled-e1786478301201.jpeg
Nine PBS sued Iron Mountain Data Centers for over 50 terabytes of archival materials stored in a Denver data center after their cloud-storage vendor, Open Source Storage (OSS), cut off access without warning. The station claims OSS' predecessor owned the data and Iron Mountain refuses to return it due to OSS' ownership of the infrastructure housing the data.

How Compaction Works in Pi

https://earendil.com/static/og/posts/compaction-in-pi.png
Pi uses compaction to summarize conversation history when context limit is reached, discarding older content while preserving recent work. Compaction can be manually triggered using /compact command or automatically after a turn ends with configurable token budget.

Blog about things you don't understand yet

https://www.seangoedecke.com/og-image.jpg
The author publishes blog posts that represent at least two things they've learned, forcing them to think critically and research topics thoroughly. This process helps clarify their thoughts and leads to more effective writing.

The Library of Ashurbanipal (2025)

Establishing a secure connection... Request ID: d5dedad719ff1e2870bdb0682e67bdfa

Credibility is the barrier to entry in silicon

https://substackcdn.com/image/fetch/$s_!47kV!,w_150,h_150,c_fill,f_auto,q_auto:good,fl_progressive:steep,g_center/https%3A%2F%2Fsubstack-video.s3.amazonaws.com%2Fvideo_upload%2Fpost%2F198703902%2F5f2b81ce-bcdd-477a-af1f-2943380ca4e7%2Ftranscoded-1779374531.png
Darian Domocos, CEO of Fiora5, discussed starting a CPU company and the challenges of credibility in Europe. He's focusing on open-source RISC-V processor IP to gain sovereignty over chip design.

Single log line is 49KB+ (ext4) / 110KB+ (btrfs) of systemd-journald disk writes

https://private-user-images.githubusercontent.com/147394/531625159-41d31a0c-2963-4d22-8d17-73bafa87e08b.png?jwt=eyJ0eXAiOiJKV1QiLCJhbGciOiJIUzI1NiJ9.eyJpc3MiOiJnaXRodWIuY29tIiwiYXVkIjoicmF3LmdpdGh1YnVzZXJjb250ZW50LmNvbSIsImtleSI6ImtleTUiLCJleHAiOjE3ODY2NTM2MDQsIm5iZiI6MTc4NjY1MzMwNCwicGF0aCI6Ii8xNDczOTQvNTMxNjI1MTU5LTQxZDMxYTBjLTI5NjMtNGQyMi04ZDE3LTczYmFmYTg3ZTA4Yi5wbmc_WC1BbXotQWxnb3JpdGhtPUFXUzQtSE1BQy1TSEEyNTYmWC1BbXotQ3JlZGVudGlhbD1BS0lBVkNPRFlMU0E1M1BRSzRaQSUyRjIwMjYwODEzJTJGdXMtZWFzdC0xJTJGczMlMkZhd3M0X3JlcXVlc3QmWC1BbXotRGF0ZT0yMDI2MDgxM1QyMDM1MDRaJlgtQW16LUV4cGlyZXM9MzAwJlgtQW16LVNpZ25hdHVyZT1hYmRkYTlhMDAzYjkyZWE0NWUwN2E0ZmM1Mjc3M2E4N2E5ODFjMDNlM2IxNjVjZTJjMWVhNWRmNjgzMDVkYjFhJlgtQW16LVNpZ25lZEhlYWRlcnM9aG9zdCZyZXNwb25zZS1jb250ZW50LXR5cGU9aW1hZ2UlMkZwbmcifQ.NPr8Z44i9SkV4fVffjwpyhYposdeMtDrMduDRZYlPRs
Using journald in write mode on an XFS file system, it writes to the hard drive but is inefficient and can corrupt during unclean reboots. This shows that iotop's accuracy issue may be due to kernel mechanisms rather than journald itself being slow.
A 17-year-old database backup of the old 0.mk URL shortener was restored and crawled, revealing that only a small percentage of links still load. The majority of links point to non-existent pages or have been blocked by websites.