cartero Monday, August 10, 2026 · No. 25855
Developer Tools

Agent platform (Part 1): How we help Grab build and run AI agents at scale

Part 1: From one support bot to a framework At Grab, AI agents have evolved from interesting team prototypes into production services used every day by millions of merchants, drivers, and consumers. Today, more than 500 services run on our internal agent framework, over 50 Model Context Protocol (MCP) servers are registered on our remote MCP framework, and a single Large Language Model (LLM) gateway fronts every model call across the company, handling billions of tokens each month. None of thi

Software Engineering

How to be useful as a software architect

Ever worked with a useful software architect? Yeah me neither. But I sure have tried to be one. Software architects have a weird role: Think about the code tomorrow. Today's code is what it is, how do we make it good tomorrow? Without slowing down, losing business, or falling off a cliff. The worst architects write lots of docs nobody reads, make proclamations of The One True Way, and fall out of touch. Many effective architects put on the Software Janitor hat and burn out in 2 years. Continue r

JavaScript

The AI Boom Made Average People More Interesting

Anybody else noticing this? As an average person, I’m not complaining. I still worry about the long-term cost of genuine human creativity being drowned out by synthetic noise. But there’s been one surprising upside. Over the past year, I’ve been able to start actual, ongoing conversations, often just by sending an email or making a phone call, with people I could never reliably reach before: State representatives VPs of engineering organizations Authors and bloggers Journalists European l

Open Source

The Consensus Weekly, July 25, 2026

Hey folks, This week had three big updates to the platform. First, all locations of jobs and companies are enriched to include their state and country to make it easier to search for regions. Second, all jobs and feed pages get marked with a topic badge if they are particularly close to one of 40 de

Machine Learning and AI

Moving Factory AI from Detection to Guidance

Modern factories already have plenty of systems that can observe. Cameras inspect parts, sensors stream machine state, dashboards summarize yield, and digital twins promise a live view of production. Yet when a station drifts from the expected process, recovery often still depends on someone walking over who has seen the problem before. That gap is where many factory AI efforts fall short. They detect anomalies, but they do not always help the person at the station decide what to do next.

Machine Learning and AI

When Computing Becomes Societal Infrastructure: What Should ACM Become?

ACM’s mission has always been broader than publishing: to advance computing as a science and profession while serving both professional and public interests.1 However, in a recent article, James Larus raises the concern that ACM has become too centered on publishing at the expense of its broader role as a professional society.2 This is the right question, and the AI era makes it even more urgent. As society enters the AI era, computing is no longer only a discipline; it is becoming societal in

Systems Programming

Zig by Example

Machine Learning and AI

Moir: Let the Model Direct Its Own Story for Robust Cross-Domain Knowledge Editing

arXiv:2607.20433v1 Announce Type: new Abstract: While language models remain frozen at their training state, the world evolves continuously. Knowledge editing has emerged as a key alternative to full retraining, but its deployment is bottlenecked by the erosion of core capabilities: mathematical and programmatic reasoning collapse while encyclopedic recall remains intact. We trace this asymmetric degradation to a distributional mismatch. Covariance-based editors preserve only the subspaces span

Machine Learning and AI

Portents Of Doom

Elon Musk is the world champion of totally implausible projections, and Kim Khan reported on a personal best in SpaceX sees total addressable market rivaling size of the U.S. economy: The $28.5T forecast compares to U.S. Q1 2026 nominal GDP of nearly $32T, with the estimate for the market of AI enterprise applications of $22.7T about 70% of total U.S. economic output. Sam Altman and Dario Amodei just aren't this good, but their projections of their Total Available Market (TAM) are still turnin

Machine Learning and AI

Break Through the Compression Bottleneck: From Theory to Practice

arXiv:2607.20434v1 Announce Type: new Abstract: As the parameter size of language models continues to grow, effective model compression is required to reduce their computational and memory overhead. Existing compression methods suffer from bottleneck issues: when the compression ratio is increased, performance degrades significantly. Low-rank decomposition and quantization are two prominent compression methods that have been proven to significantly reduce the computational and memory requiremen

Compilers

What is Good? Extracting and Testing Implicit Theories of Literary Quality from LLM Reasoning Traces

arXiv:2607.20425v1 Announce Type: new Abstract: What makes writing "good" remains a persistent question in literary studies and computational linguistics. We present a two-study investigation of how reasoning-enabled LLMs evaluate literary quality. In Study 1, we construct a benchmark of 30 real texts spanning six quality tiers, from canonical literature to anonymous forum posts, and extract the model's implicit theory of quality from its reasoning traces. Across five DeepSeek replications, t

Linux

ln Command in Linux: Hard Links vs. Symlinks

Hard links and soft links in Linux both use the ln command but behave very differently. This practical guide explains exactly how each works, when to use which, and how to avoid the common pitfalls that catch even experienced sysadmins.Continue reading...

Machine Learning and AI

Claude Opus 5 Is Most Efficient at Medium Effort: FrontierCode Benchmark Data Explained

FrontierCode v1.1 benchmark data reveals that Claude Opus 5 peaks at medium reasoning effort, delivering near-top scores at roughly half the compute cost of higher effort settings. Continue reading Claude Opus 5 Is Most Efficient at Medium Effort: FrontierCode Benchmark Data Explained on SitePoint.

Programming Languages

Position: Natural Language Should Not Fully Replace Formal Languages

arXiv:2607.20432v1 Announce Type: new Abstract: Recent advances in large language models and their widespread adoption have prompted claims that natural language could entirely replace formal languages, such as programming languages for software design. In this position paper, we argue that this perspective overlooks fundamental linguistic properties of natural language, specifically that it is optimized for underspecification in open-ended contexts. We introduce a formal framework centered on

Linux

NPKILL - nuke old node_modules and .venv directories to free up disk space

Got a ton of old side projects, disk spice getting tight? NPKILL is the tool for you! Just run npx npkill and it will scan your system for node_modules folders, list them by age and size and then you can nuke them with a single keystroke. To clean up your old Python virtual environments just run npx npkill --target .venv and off you go.

Rust

Symbolica 2.2: symbolic integration in Rust with 7,000+ pattern-matching rules

For Symbolica 2.2, we ported 7,000+ integration rules from the Rubi project to Rust and tested the implementation against the complete 72,944-problem corpus. The integration engine is available as the MIT-licensed symbolica-integrate crate. The post covers symbolic integration, how the AI-assisted port was organized and kept in check, and other new Symbolica features. Even though AI was used for such a vast but ultimately mechanical port, there was a lot of manual labor, craft and domain knowl

Science

Behind the Blog: No Spoilers

This is Behind the Blog, where we share our behind-the-scenes thoughts about how a few of our top stories of the week came together. This week, we discuss astrology, Flock on the brain, The Odyssey, and more.JASON: This week I wrote pretty much exclusively about Flock, and I am working on a few more articles about automated license plate readers that will probably go in the next week or so. I am well aware that there are other topics in the world, but, basically, there is now so much interest in

Science

Illustrae website review

Illustrae is software for creating scientific illustrations. Let me hit what might be the most relevant feature of this product for many readers right at the top:There’s no free version to try.For that reason, I cannot review the software. I am not able to throw down money for a subscription to try this product right now, not even a temporary subscription. But I wanted to write a post pointing to this product, because it is relevant to this blog’s mission.The main poster-related feature is t

Algorithms

LLM-INSTRUCT at UZH Shared Task 2026: Constraint-Aware Retrieval and Selective Debate for Paragraph-Level Argument Mining

arXiv:2607.20430v1 Announce Type: new Abstract: We present LLM-INSTRUCT, the winning system for the UZH Shared Task at ArgMining 2026 on paragraph-level argument mining in UN and UNESCO resolutions. The task requires paragraph-type classification, prediction of a subset of 141 official tags, and directed relation prediction under a strict JSON schema setting using only open-weight models up to 8B parameters. We frame the task as constrained structured prediction. The system first narrows the ca

Systems Programming

Data locality (sometimes) beats algorithmic complexity

I've been ECS-curious ever since I learned about it in the Bevy game engine documentation.The ECS architecture predictably improves performance in languages that give you low-level control over memory (C, C++, Rust, Zig, and friends). But how does it fare when used in high-level, dynamic, garbage-collected languages such as JavaScript?This is the question Dan Murphy set out to answer in The Physics of Memory: Is it possible to use an ECS-style architecture in Javascript? And for applicable ope

Linux

Discord Patch Notes: July 7, 2026

Check out the finer details of the more technical fixes implemented into Discord recently.

Self-Hosting

Notes from rss.chat land

There are 4 rss.chat servers – neat (rss.chat, demo.rss.chat, rsschat.andysylvester.com and perstitio.us)! The latest (perstitio.us) actually incorporates feeds from other sources besides the rss.chat instance – interesting! The author has posted an essay describing this approach. Two rss.chat forks: xicubed and hal-nine Frank Meeuuwsen shares some thoughts on rss.chat, in addition to John Johnston. Decided to copy the post from the perstitio.us author: This is the question. Blu

Machine Learning and AI

Incredible. Every Single Take on AI is Wrong.

For the past 6 months, I've been reading on-and-off the flurry of blog posts and articles written by very experienced software developers on AI, and the overwhelming feeling I get is, that they are all wrong.I don't have the inclination to write a treatise countering all the things people are saying, but I do have three articles for which I want to simply document agree/disagree on the various points being made. If only to document my own beliefs, so I can check a year or so in the future and ch

Machine Learning and AI

Distilling The Moat

Whisky Still The original function of a Web server was to respond to queries by revealing the appropriate part of their internal data. This necessarily meant that repeated queries, for example from a search engine's or an internet archive's web crawler, could extract the server's entire internal data. Since the extracted data had been published on the Web, it was not trade secret. It was protected by the publisher's copyright. This has led to many lawsuits, for example against the Internet Archi

Machine Learning and AI

Every System Is About to Get a Guard

Every System Is About to Get a Guard In September 2025, a group tracked as GTG-1002 ran an espionage campaign against roughly thirty organizations across technology, finance, chemicals, and government. The unusual part isn’t the target list. It’s that the AI executed an estimated 80 to 90 percent of the operation on its own. Reconnaissance, vulnerability discovery, exploitation, lateral movement, credential harvesting, exfiltration. Human operators broke the work into tasks and stepped in fo

Machine Learning and AI

Open-weights AI models have become good enough

Over the past week I've played around with Kimi K3 by Moonshot AI and Qwen 3.8 Max by Alibaba. Both are large Chinese open-weight models (weights promised to be released soon) and tout benchmarks showing they're as capable as the frontier western models (Fable 5 by Anthropic and GPT-5.5 Sol by OpenAI). I wouldn't go that far, but these are really capable models. In my AI-coding tests, both have performed really well. Compare the test mini-games on my vibe-coding benchmark generated by Fable, S

Databases

Adminer support for DuckDB and Parquet database files

Adminer is the best PHP database management, since PHPMyAdmin has enshittified by a messy jumble of mixed PHP and slow,  Javascript-heavy requests. But Adminer has… The post Adminer support for DuckDB and Parquet database files appeared first on Ambience. Follow the author at: @[email protected]

Machine Learning and AI

No Dumb Questions: What is the AI bottleneck? How does context engineering fix it?​​​​‌‍​‍​‍‌‍‌​‍‌‍‍‌‌‍‌‌‍‍‌‌‍‍​‍​‍​‍‍​‍​‍‌​‌‍​‌‌‍‍‌‍‍‌‌‌​‌‍‌​‍‍‌‍‍‌‌‍​‍​‍​‍​​‍​‍‌‍‍​‌​‍‌‍‌‌‌‍‌‍​‍​‍​‍‍​‍​‍‌‍‍​‌‌​‌‌​‌​​‌​​‍‍​‍​‍‌‍​‌‍‌‌​​‍‍‌​‌‌​‌‍​‌‌‍​‌‍‍‌‍‌‌‍‌‍‌‌‌​‍‌‍‌‍‌‍​‌‍‌‌​‍‍‌‍​‌‍​‍‌‍‍‌‌‍‍‌‌​‌‍‌‌‌‍‍‌‌​​‍‌‍‌‌‌‍‌​‌‍‍‌‌‌​​‍‌‍‌‌‍‌‍‌​‌‍‌‌​‌‌​​‌​‍‌‍‌‌‌​‌‍‌‌‌‍‍‌‌​‌‍​‌‌‌​‌‍‍‌‌‍‌‍‍​‍‌‍‍‌‌‍‌​​‌​‍‌​​‌‍​‍​‌‍​​‍‌‍‌‌​‍‌​​‌​‍‌​‍​‌‍‌‌‌‍​‍​‍‌​‍‌​‌​​‍​‌‍‌​‌‍‌‌​‍‌‌‍​‍‌‍​‌​‌​​​‍​‍‌‌‍​‍​‌‌​‌​‌​​‌‍‌‍​​​​‌‍​​​​‌​​​​‍​‍‌‌​‌‍‌‌​​‌‍‌‌​‌‌‍​‍‌‍​‌‍‌‍‌‌‌​​‌‍‌​‌‌​​‍‌​​‌‍​‌‌‌​‌‍‍​​‌‌‌​‌‍‍‌‌‌​‌‍​‌‍‌‌​‌‍​‍‌‍​‌‌​‌‍‌‌‌‌‌‌‌​‍‌‍​​‌‌‍‍​‌‌​‌‌​‌​​‌​​‍‌‌​​‌​​‌​‍‌‌​​‍‌​‌‍​‍‌‌​​‍‌​‌‍‌‍​‌‍‌‌​​‍‍‌​‌‌​‌‍​‌‌‍​‌‍‍‌‍‌‌‍‌‍‌‌‌​‍‌‍‌‍‌‍​‌‍‌‌​‍‍‌‍​‌‍​‍‌‍‌‍‍‌‌‍‌​​‌​‍‌​​‌‍​‍​‌‍​​‍‌‍‌‌​‍‌​​‌​‍‌​‍​‌‍‌‌‌‍​‍​‍‌​‍‌​‌​​‍​‌‍‌​‌‍‌‌​‍‌‌‍​‍‌‍​‌​‌​​​‍​‍‌‌‍​‍​‌‌​‌​‌​​‌‍‌‍​​​​‌‍​​​​‌​​​​‍​‍‌‍‌‌​‌‍‌‌​​‌‍‌‌​‌‌‍​‍‌‍​‌‍‌‍‌‌‌​​‌‍‌​‌‌​​‍‌‍‌​​‌‍​‌‌‌​‌‍‍​​‌‌‌​‌‍‍‌‌‌​‌‍​‌‍‌‌​‍‌‍‌​​‌‍‌‌‌​‍‌​‌​​‌‍‌‌‌‍​‌‌​‌‍‍‌‌‌‍‌‍‌‌​‌‌​​‌‌‌‌‍​‍‌‍​‌‍‍‌‌​‌‍‍​‌‍‌‌‌‍‌​​‍​‍‌‌

In this No Dumb Questions, Stack's Director of Data Science Michael Foree teaches Phoebe about AI context, context engineering, and what she can do to become a better context engineer. ​​​​‌‍​‍​‍‌‍‌​‍‌‍‍‌‌‍‌‌‍‍‌‌‍‍​‍​‍​‍‍​‍​‍‌​‌‍​‌‌‍‍‌‍‍‌‌‌​‌‍‌​‍‍‌‍‍‌‌‍​‍​‍​‍​​‍​‍‌‍‍​‌​‍‌

Linux

Friday, July 24, 2026

What I should be doing is creating new things rather than constantly rearranging the things I already have. Maybe I like to switch between operating systems because it feels like rearranging my office. A fresh view feels good. Today I'm running Dank Linux as my desktop. It's so weird and I dig it. ✍️ Reply by email

Databases

rss.chat install diary – had some problems

I am writing this post to document the issues that I experienced in installing the rss.chat chat server recently introduced by Dave Winer. I disagree with his assertion that it is “easy to deploy on Node.js”. In following the initial rss.chat server, there are some mentions of changes/updates that are being made. That is good, but some of the issues I experienced have not been addressed, and some of them are significant barriers to entry in starting a new rss.chat server. I would have pos

Developer Tools

The Postmark MCP server, one year later: from 4 tools to 24

About a year ago, we introduced something experimental from Postmark Labs: an MCP server that let an AI assistant send email through Postmark. It shipped with exactly one useful tool, sendEmail, plus three supporting ones (four total). You gave it a recipient, a subject, and a body, and it sent. We said at the time we'd started with a single Postmark server "because we had to start somewhere."A lot has happened since. The project graduated from Labs and became the official @activecampaign/postma

Self-Hosting

A Large eInk Tablet

I had a specific desire: an A4-page-sized way of displaying song sheets, that would work well out of doors.

Science

The Download: NASA’s new space telescope and OpenAI’s autonomous hacker

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Shape-shifting mirrors on NASA’s new space telescope could unveil Jupiters like our own When NASA’s Nancy Grace Roman Space Telescope launches, as early as the end of next month, it will attempt one of astronomy’s most precise disappearing acts to date. It will carry the first space-bound “active” coronagraph, an instrument that effectively

Self-Hosting

YARP and Aspire: "https+http scheme is not supported"

Recently I was wiring up a YARP reverse proxy in front of a couple of Aspire-managed services: an API and an Angular frontend. Aspire gives you service discovery for free, so the obvious move is to point your YARP clusters at the logical service names instead of hardcoded URLs. My first attempt looked like this: "Clusters": { "api-cluster": { "Destinations": { "api-destination": { "Address": "https+http://api" } } }, "frontend-cluster": { "Destinations": {

Machine Learning and AI

AI 2040: Plan A

Databases

Cornelia Biacsics: Building The OAPE PostgreSQL Certification

Building the OAPE PostgreSQL CertificationI’m one of the founders of the Open Alliance for PostgreSQL Education (OAPE).Over the past two years, I’ve had a front-row seat to watching an idea become a community, and a community become a movement.This is the story of how the idea of an open, vendor-neutral PostgreSQL certification turned into an organisation supported by contributors from around the world.The classic “What If” — thoughtThere are moments when an idea seems so obvious

Go

Solod 0.3: Concurrency, JSON, more safety

Solod (So) is a subset of Go that translates to regular C — with zero runtime, manual memory management, and source-level interop. It's designed for two main audiences: Go developers who want low-level control without having to learn another language. C developers who like Go's style. At the end of the v0.2 post, I said the obvious goal for the next release was concurrency, along with the stdlib packages that support it. That's what v0.3 is about. So now has threads, channels, worker pools,

Machine Learning and AI

More Is Not More: What Matters for Diversity in LLM Opinions?

arXiv:2607.20429v1 Announce Type: new Abstract: Large language models are increasingly used to simulate diverse human opinions in open-ended tasks such as synthetic surveys, focus group modeling, and public opinion prediction. However, LLM outputs exhibit systematic opinion homogenization. Practitioners have explored various interventions to increase diversity, but the landscape remains fragmented: different methods are evaluated in isolation with incomparable metrics, and in practice they are

Machine Learning and AI

A Court Reporter Submitted AI-Generated Errors in Official Court Transcript, Judge Says

A judge caught a court reporter making AI-generated errors in a court transcript, and put court reporters everywhere on notice for their use of AI. In a memorandum decision concerning a case about a man who sold drugs to another man who overdosed and died, filed on July 23, Judge Paul Felix wrote in a footnote of the decision that a transcript contained errors that looked a lot like generative AI. The footnote was spotted by attorney Rob Freund on X.

Self-Hosting

Two nasty surprises in Home Assistant's config

Last year, I motorized the rolling shutters on the southern façade of my apartment. My idea was to manage them via Home Assistant. I had a couple of automations in mind: In the evening, roll down the shutters of my bedroomIn the morning:If it’s too hot outside, roll down all shuttersIf it’s too cold outside, roll down all shuttersIn other cases, roll up all shutters but my bedroom’s Living in France, I added the official Météo France integration.

Self-Hosting

Somewhere to stash my mixtapes (Week 28)

TL;DR: I built mixtapes.lmorchard.com to publish playlists on my own domain instead of leaving them trapped in Spotify, which spun off two new repos (byom-sync and byom-player). Plus more starnet gamedev (an exploit-barrage mini-game with a Butterchurn "brain damage" overlay), I finally ditched Disqus for remark42 after almost 20 years, and Minnaloushe is inching toward parole from cat jail.

Machine Learning and AI

Why Your AI System Is Never Done

Antony Evans argues that building an AI system is farming, not hunting: there is no finished, shipped state, only a system you keep tending as models change.

Self-Hosting

Goodbye Discord webhooks, hello Gotify

For a long time I used Discord webhooks for notifications from my services. It was easy, it worked and almost every application knew how to send something to Discord. Create a private channel, copy a webhook URL, paste it into a service and wait until something breaks. Very advanced engineering. The more services I added, however, the

Developer Tools

Entering accented characters in Linux

There are at several distinct ways of getting accented characters and other non-ASCII symbols in Linux. Confusingly, many of them seem to be called called "compose sequences" and none is apparently well-known.

Databases

rainfrog (0.4.1) now has autocomplete!

rainfrog (https://github.com/achristmascarl/rainfrog) is a database terminal tool; the goal is to provide a lightweight, keyboard-first TUI for interacting with databases. It currently supports Postgres, MySQL, SQLite, Oracle, and DuckDB. v0.4.1 introduces a long-awaited (by me, not sure if anyone else was waiting for it...) autocomplete implementation, along with autopairs for quotes/parentheses/brackets. The full list of features and configuration options is in the README! submitted by /

Web Development

Testing Google’s “modern-web-guidance” skill against a real React app

LLM-assisted frontend work has a particular failure mode. The model confidently writes code that was best-practice in 2021. It reaches for 100vh, hand-rolls a dark-mode toggle with a class on <body>, or disables the submit button to “prevent” invalid input. None of it is wrong exactly. It’s just a few years stale, because the training data is a few years stale and the web platform moves faster than that. Google Chrome’s modern-web-guidance skill is a direct attempt to fix that. It’s n

Security and Cryptography

Project ORBITAL

IntroductionThe modern cyber threat landscape has seen a fundamental shift in how threat actors manage and deploy their infrastructure. Advanced persistent threats (APTs) have almost completely moved away from static command-and-control (C2) servers, opting instead to build complex, multi-layered botnets known as Operational Relay Box (ORB) networks. Project ORBITAL (which stands for Operational Relay Box Intelligence, Tracking, & Analysis Lexicon) was established as a centralised intelligence

Rust

My blogs written in Rust

It may load slower (~11Mb) I provided dark/light mode for switching Repo: https://github.com/Cottons29/aimer submitted by /u/Agile_Significance91 [link] [comments]

Security and Cryptography

One Rust crypto core shared by a CLI and a WASM browser client, kept honest by native<->WASM golden vectors in CI

I have been building Sotto, an end to end encrypted secrets manager for developer teams (the one liner is "stop Slacking your .env around"), and the part I would most like r/rust's eyes on is the crypto architecture. There is exactly one crypto implementation: a sotto-core crate. The native CLI links it directly, and the browser client uses the same crate compiled to WASM. So encryption, key wrapping and the vault hierarchy are the same code in both places, rather than two implementations I have

Science

KV the Apostate: Faith-Based Computing Versus the Unnatural Science

Dear KV, The hype cycle that is AI seems as if it will never end. What I was doing as a developer two years ago seems very different to what I am doing now that management has required us to use LLMs in our work. Some of what these things do seems helpful—roughing out a general framework or helping with tasks such as creating unit tests that have often seemed like drudgery. From your writings, I suspect you remain skeptical and continue to be part of the old school that thinks this is a fad th