Hacker News

Latest

Man and the Computer by John G. Kemeny (1972 book by the co-creator of BASIC)

2026-07-29 @ 22:53:44Points: 14Comments: 4

LLM Honeypot

2026-07-29 @ 22:51:03Points: 23Comments: 11

GitHub is the wrong shape for this new world

2026-07-29 @ 22:27:09Points: 35Comments: 19

AI's top startups are barely publishing their research

2026-07-29 @ 21:25:40Points: 151Comments: 92

The Cold Email

2026-07-29 @ 21:06:42Points: 66Comments: 32

SalesPatriot (YC W25) Is Hiring FDEs

2026-07-29 @ 21:01:01Points: 1

The coolest use for the Vision Pro

2026-07-29 @ 20:39:40Points: 349Comments: 162

A Trampoline

2026-07-29 @ 20:14:36Points: 63Comments: 36

Kimi K3-256k

2026-07-29 @ 19:25:33Points: 326Comments: 92

Commodification of Intelligence: Good, Bad, and Ugly Circular AI Deals

2026-07-29 @ 18:57:10Points: 52Comments: 28

How to think about software quality (2022)

2026-07-29 @ 18:42:10Points: 47Comments: 27

Turning a dumb AC unit smart (without losing my security deposit)

2026-07-29 @ 18:28:51Points: 99Comments: 84

Show HN: CheapFoodMap – A map of good meals under $10

2026-07-29 @ 16:59:52Points: 117Comments: 142

It's inspried by 거지맵 (Begger's Map) a Korean crowdsourced map students use to find cheap eats.

Ocverage is heaviest in Texas, since I live in Dallas, but have 1200 meals across 15 US cities. Seed data came from Google Review, 4.2 star or higher with at least 500 reviews, and verified price under $10 per menu item.

Things I would love feedback on : whether the price-freshness model makes sense, and what would make you trust the price on a site like this. How to encourage people to update prices, since inflation is making food price very frequent.

https://cheapfoodmap.com

Any and all suggestion will be super helpful. Thank you!

Some thoughts about Anthropic's new cryptanalysis results

2026-07-29 @ 16:42:20Points: 98Comments: 52

Keychron announces first open-source firmware for gaming mice

2026-07-29 @ 16:36:59Points: 276Comments: 101

Launch HN: Tokenless (YC S26) – Automatic model switching to save money

2026-07-29 @ 15:55:27Points: 52Comments: 42

https://usetokenless.com/), which I’m building alongside co-founders Andrew and Kev. We’re building an API gateway which routes agent traffic dynamically turn-by-turn between different models to save on AI spend.

The cost of AI tokens is top-of-mind for many. Companies like Uber and Salesforce have been complaining about blowing their yearly AI spend faster than expected.

Frontier models are amazing for dev work, but are so expensive. Open-source models are cheap and rapidly improving, closing the gap with frontier models, but aren’t quite there yet.

Tokenless gets you the best of both worlds–routing harder turns to smarter models only when needed, which keeps costs low.

Before Tokenless, I was doing a PhD at Princeton. While using coding/other agents, I constantly agonized over model choice, to make sure my AI spend was going as far as possible on my academic Cursor account.

At the same time, I was doing LLM research, and a small technique I developed while in recovery from NeurIPS submission season seemed to hit SOTA pretty fast. I was surprised that such simple ideas could do routing well.

We’ve been able to develop a version of the router that matches the performance of Claude Fable 5 at half the cost. The blog post on our website explores the technical details on how we did this (https://usetokenless.com/blog/building-tokenless/).

Highlights: - Our approach queries multiple models at once and uses their progress to make decisions (this technique is novel AFAIK, let us know if you know anyone else doing this). - Switching models doesn’t destroy the cache if the routing algorithm is aware of when the cache is hot/cold.

To come: - Adding Kimi K3, all other GPT efforts and more to the router

Go ahead and sign up on usetokenless.com and try using Tokenless with your agent, you’ll get $20 of free credit. Here’s a demo on how to use it: https://youtu.be/sjZWriclcls

Tokenless provides frontier-level intelligence for cheaper, so we’d love some feedback on how it feels to use, any corner cases that the router routes incorrectly, and whether you find the routing problem interesting!

Superlogical

2026-07-29 @ 15:41:33Points: 506Comments: 315

Show HN: Open-source engine running Gemma 4 26B in 2 GB RAM on any M-series Mac

2026-07-29 @ 15:05:43Points: 632Comments: 223

I built a specialized inference engine for running 4-bit Gemma 4 26B-A4B-IT on any M-series Mac using about 2 GB of RAM. It is called TurboFieldfare and is written in Swift and Metal.

I have always adored on-device AI. It feels like magic that you can run a powerful NN on your Mac or iPhone. So I wanted to push the limits a bit and run a model whose weights don’t fit in memory.

The model’s 4-bit quantized weights occupy roughly 14 GB, which makes running it with conventional inference tools almost impossible on an 8 GB or even 16 GB Mac once the OS, applications, and KV cache are included.

The trick is to keep the shared part of the model and the KV cache in RAM, then stream only the routed experts needed for each token from SSD. An SSD is way slower than RAM, so the runtime uses a small expert cache and bounded parallel `pread`. While those reads are in flight, the GPU runs the shared part of the layer.

I ran more than 100 experiments. Most didn’t work. A few got me here. The experiments are described in the GitHub repo.

It currently generates 5–6 tok/s on an 8 GB M2 MacBook Air and 31–35 tok/s on an M5 MacBook Pro.

I also added an experimental OpenAI-compatible local server. It supports streaming and tool calls, and reuses one prompt prefix from the KV cache.

Try it! The Mac app is easy to install. On the first run, it will download 15 GB of weights from Hugging Face. The model is surprisingly capable.

I would love any kind of feedback!

A.I. companies are recruiting electricians and carpenters by the thousands

2026-07-29 @ 14:43:47Points: 206Comments: 260

Self-hosting Kimi K3: 20% more hardware cost, 20% better task resolution

2026-07-29 @ 14:38:35Points: 121Comments: 41

Handbook.md shows that long policy documents do not reliably govern agents

2026-07-29 @ 13:01:57Points: 286Comments: 181

Darktable

2026-07-29 @ 12:33:02Points: 287Comments: 142

Document-borne AI worms can self-propagate through Copilot for Word

2026-07-29 @ 11:44:33Points: 337Comments: 256

KOReader

2026-07-29 @ 11:05:08Points: 654Comments: 209

Anatomy of a Frontier Lab Agent Intrusion: A Timeline of the July 2026 Incident

2026-07-28 @ 20:28:33Points: 278Comments: 168

Hamburg's Stadtpark: A Park Built to Be Used

2026-07-27 @ 06:23:46Points: 112Comments: 27

Interconverting std::function with copyable_function

2026-07-27 @ 05:39:33Points: 6Comments: 0

Refactoring cuisine: how an Iraqi stew sailed to Singapore

2026-07-26 @ 22:18:42Points: 21Comments: 0

Staging patches with git add (2024)

2026-07-25 @ 15:44:34Points: 31Comments: 39

The Rust on ESP Book

2026-07-25 @ 13:15:41Points: 122Comments: 10

Archives

2026

2025

2024

2023

2022