Topic
Everything filed under Infrastructure, newest first.
RSS · JSON · All topics
OpenAI launches the Agents API in public beta: a managed Codex harness with durable cloud sessions, sandbox compute, context compaction, subagents, and resumable multi-hour agent work for developers.
7 min · 1,539 words
Introducing cf: the agentic CLI for the entire Cloudflare API
Cloudflare releases cf, a new open-source agentic CLI that mirrors the entire Cloudflare API with TypeScript configuration, alongside Forge, their internal SDK generator—aimed at agent-heavy Wrangler usage.
9 min · 1,989 words
How Delhi Cut Electricity Loss from 50 to 5 PercentHere’s How Delhi Achieved Its Epic Power-Grid Fix
IEEE Spectrum's Mini Shaji Thomas recounts how Delhi slashed electricity losses from about 50% to 5%—the privatization, metering, and grid-operations reforms that turned candlelit kitchens into reliable power.
15 min · 3,457 words
EmDash 1.0: the stable CMS with a secure plugin registry
Cloudflare releases EmDash 1.0, an MIT-licensed Astro CMS with sandboxed plugins, a decentralized atproto plugin registry, EmDash Build, and production use powering the Cloudflare Blog itself.
2 min · 479 words
Why I Stopped Defaulting to Next.js and Vercel
How AI coding agents made it practical for me to build and own a different stack with TanStack Start and Cloudflare. The first person who introduced me to Next.js was my friend Haythem Lazaar. We were at university, building Collo , a project management tool for remote teams.
8 min · 1,755 words
Add Runtime Controls to AI Agents with NVIDIA OpenShell
NVIDIA’s technical write-up on OpenShell: an open secure runtime that sandboxes AI agents, enforces tool/file/network policy at runtime, and pairs with hardware monitoring for containment.
7 min · 1,593 words
First Steps of the PLC Organization – Independent Public Ledger of Credentials
A year after Bluesky pledged an independent PLC directory operator, the Public Ledger of Credentials Organization formally exists and takes its first administrative steps.
3 min · 606 words
Accelerated Out of Core Shuffling
Benjamin Zaitlen explains RapidsMPF’s reusable out-of-core shuffler for distributed analytics—how spilling turns shuffle OOM headaches into a budgetable resource, and what it takes to push shuffle bandwidth toward terabytes per second.
15 min · 3,466 words
Introducing Casita: A content-addressed store for source code and build artifacts
Domen Kožar introduces Casita, a content-addressed object store for source code and build artifacts—an argument for rethinking Nix-style packaging from the package manager up.
7 min · 1,723 words
Jordan Petridis introduces Toolpak, a GNOME project for packaging developer tools in an image-based desktop world—why host tooling breaks on immutable systems and how Toolpak aims to fix it.
7 min · 1,644 words
Deploying Guix images on Linode
David Thompson documents migrating to Linode (Akamai Cloud) and deploying Guix system images—image building, boot configuration, and practical notes from moving off DigitalOcean.
4 min · 1,000 words
DeepSeek Elastic Compute (DSec)
# Computer Science > Distributed, Parallel, and Cluster Computing [Submitted on 19 Sep 2026] # Title:DeepSeek Elastic Compute (DSec): A Sandbox Infrastructure for Effective Agentic Training at Scale View PDF HTML (experimental) Abstract:Large-scale agentic training and evaluation with large language models (LLMs) rely on isolated, stateful execution environments in which models inspect repositories, invoke tools, execute commands, and interact with task-specific services. These workloads create sandboxes in large bursts, span heterogeneous functionality and isolation…
16 min · 3,693 words
Mixture of Experts (MoE) for Backend Engineers
A detailed visual guide to token routing, expert batching, weighted combination, and the memory and communication tradeoffs of MoE serving. Suppose a model has dozens of feed-forward subnetworks, but each token uses only two of them. The arithmetic per token can stay modest while the total weight set grows. Now place that model on eight GPUs. If every GPU stores all experts, memory can become the limit; if experts are split across GPUs, token activations must travel to whichever GPU owns their…
15 min · 3,515 words
Crazy as it sounds to say this, since it’s all anybody’s been able to talk about for over a year, but the impact of AI on computing hasn’t yet sunk in.
6 min · 1,349 words
John D. Cook explains dawn-dusk sun-synchronous orbits: why satellites there stay in perpetual sunlight, and what that means for powering orbital compute servers.
1 min · 257 words
How I Could've Accessed 17 Trillion Microsoft Records
A security write-up estimating ~17.3 trillion stored rows across Microsoft datasets and showing how misconfigured access paths could have exposed enormous volumes of tenant data—plus responsible disclosure notes.
9 min · 2,086 words
When to choose x86-64 vs aarch64
Not all cloud vCPUs are created equal. When you create a PlanetScale Postgres or Neki database, you have to choose between `aarch64` (ARM) and `x86-64`. Two clusters on different architectures can have the same vCPU count and RAM, yet perform very differently. It's worth understanding the implications, since you can't easily switch the CPU architecture on your cluster later. x86 grew up on the desktop, prioritizing backward compatibility and performance. It began at Intel in 1978 and IBM’s…
6 min · 1,326 words
Every package is already installed
Farid Zakaria introduces omnibin: a FUSE filesystem that puts every binary Nixpkgs ever shipped on your $PATH—nothing installed upfront, 0 bytes on disk until something is actually run.
4 min · 1,001 words
How Cloudflare addressed a cross-tenant data exposure vulnerability in Containers
Cloudflare details how a Containers/Sandboxes cross-tenant bug let residual dm-thin disk blocks leak between customers, how Oren Yomtov reported it, and the fleet-wide remediation completed by 19 Sep 2026.
7 min · 1,583 words
Behind Project Suncatcher, our moonshot to put AI in space
Google Research outlines Project Suncatcher: a moonshot exploring solar-powered ML infrastructure in space, and the engineering facts behind the idea.
4 min · 846 words