Hacker News

Latest

GPT-6 Astra Solves a WWI German Radio Cipher

2026-09-19 @ 06:41:44Points: 173Comments: 96

If math is more than proof, we need to better celebrate the rest of it

2026-09-19 @ 06:28:02Points: 175Comments: 126

Apple M6 Pro Achieves the Highest Single-Core CPU Score in Geekbench 7

2026-09-19 @ 06:19:24Points: 95Comments: 107

Human brain is two separate organs, Stanford Medicine-led research finds

2026-09-19 @ 05:48:50Points: 356Comments: 131

NASA-IBM Lunar Foundation open-Source Geospatial AI Model

2026-09-19 @ 04:44:35Points: 43Comments: 4

San Francisco Onion Futures Company

2026-09-19 @ 04:23:30Points: 238Comments: 85

SDCC – Small Device C Compiler

2026-09-19 @ 02:32:36Points: 100Comments: 20

Science Is Open Software

2026-09-19 @ 02:21:35Points: 112Comments: 41

How OpenAI Used Its Own LLMs to Design Its Jalapeño Chip

2026-09-18 @ 23:04:17Points: 150Comments: 100

Android 17 is the first since 3.x to add new APIs without releasing to the AOSP

2026-09-18 @ 19:03:09Points: 898Comments: 484

Cache-to-Cache: Direct Semantic Communication Between LLMs (2025)

2026-09-18 @ 18:55:35Points: 97Comments: 14

Saving another 100TB of RAM

2026-09-18 @ 18:51:46Points: 389Comments: 84

Photon-Emission-Guided Laser Fault Injection Enables RP2350 Secure Debug

2026-09-18 @ 16:54:18Points: 201Comments: 75

Cloudflare Quick Tunnels

2026-09-18 @ 14:18:41Points: 744Comments: 294

OpenJev

2026-09-18 @ 09:42:22Points: 644Comments: 273

Inside ZCode: Silently uploading your Git history to the cloud

2026-09-18 @ 06:11:17Points: 313Comments: 105

Warez: The Infrastructure and Aesthetics of Piracy (2021)

2026-09-18 @ 02:56:48Points: 181Comments: 90

Minimal Phone 2

2026-09-18 @ 02:00:25Points: 294Comments: 239

Show HN: Cactus Needle 3: 8-29MB automation models can match DeepSeek V4 Flash

2026-09-18 @ 00:11:44Points: 206Comments: 89

We submitted Needle 2 here a few weeks ago, and the feedback in the discussion thread was incredibly valuable, thanks! Thanks to all that feedback, we’ve been able to move quickly to release Needle 3 and I'd love to hear what you think again.

The key features:

1) Automation (tool calls & structured JSON output): Needle still doesn't chat by design, its quite challenging to pack general capacity into such small models, so we focus on tool calls and structured JSON. If no tool you declared fits the request, you get an empty list back (note for when playing with the demo).

2) Intelligence Laddering: Every layer (2 to 20) is a deployable subnetwork, so one set of weights, 25 to 121 million parameters at 2-bit, shipping as 8-29MB binaries. On a Raspberry Pi 5 it decodes at up to 4k tokens/sec and prefills at up to 10k.

3) Monarch Hadamard MLP: replaces the dense FFN with three learnable Walsh-Hadamard-initialized Kronecker (Monarch) factor pairs interleaved with per-channel diagonal scales, fixed permutations, a SiLU nonlinearity, and a rank-8 input-conditioned gate, so each token gets a fully mixed nonlinear transform of its d_model channels at O(d√d) parameters and compute instead of the O(d²) a dense 4x-expansion MLP would cost.

4) Performance: On Mobile Actions (phone commands, scored on the exact call) the 20-layer model gets 86.0 through the shipped 2-bit binary; LFM2.5 1.2B is at 82.4, Qwen3.5 0.8B at 76.0, Apple's on-device model at 57.6, all at f16. More results on the link, we do not win everywhere ofc.

5) Multilingual: Needle 3 now supports English, French, Spanish, German, Dutch, Italian, Polish, with more languages coming.

6) Finetuning: You can achieve DeepSeek v4 Flash grade performance on a narrow task with just 4L, stress on "narrow task", we found that production users often prefer tuning before production.

7) Triggers: Grounding is a common challenge for tool call, at least for Needle 2, so we added support case-insensitive regular expressions matched against each request to gate false negatives.

8) Confidence: Every response also carries a calibrated confidence score, the minimum of a judgement on the finished call and its decode probability. Act above your threshold, show the call and ask below it, or escalate to a bigger model.

9) Supported Platforms: macOS, Linux on x86-64, ARM64, ARMv7, RISC-V and MIPS32, Windows x64 and ARM, Android, iOS, watchOS, tvOS, the browser as WebAssembly, and a WASI component.

Thanks for reading and as always, thoughts appreciated!

How to Write with an LLM

2026-09-17 @ 21:48:38Points: 522Comments: 346

The first new cat species discovered in 100 years

2026-09-17 @ 18:31:56Points: 295Comments: 110

Suppress vulnerabilities applying Kubernetes context to scans

2026-09-17 @ 13:39:50Points: 8Comments: 2

Ctenophores: Wonders of Biology

2026-09-17 @ 01:58:25Points: 40Comments: 6

Why building a Rust LSP is hard

2026-09-16 @ 22:51:16Points: 101Comments: 40

Typesafe-computer-use drives a Mac toward a goal for 1/50th of a cent per step

2026-09-16 @ 22:01:49Points: 82Comments: 59

Cyclomatic Complexity in C#

2026-09-16 @ 19:01:09Points: 68Comments: 27

You can run Git on object storage if you re-make packfiles

2026-09-16 @ 17:16:11Points: 74Comments: 19

Veronese's Dogs

2026-09-16 @ 15:32:27Points: 18Comments: 2

Goroutine Leak Profiles

2026-09-16 @ 13:30:30Points: 48Comments: 5

"The Secret Life of Circuits" is here

2026-09-15 @ 23:15:05Points: 114Comments: 31

Archives

2026

2025

2024

2023

2022