Hacker News

Latest

Bluesky draws its logo on screenshots

2026-08-17 @ 22:20:40Points: 68Comments: 40

Nation's Largest Reservoirs Are Drying Up, Threatening Life in the Southwest

2026-08-17 @ 22:15:38Points: 16Comments: 20

Quake Shareware, a CD-ROM just a little too full

2026-08-17 @ 22:06:14Points: 28Comments: 3

Fairphone 6 and PostmarketOS working main camera

2026-08-17 @ 22:01:17Points: 25Comments: 4

How do functions like alloca allocate memory from the stack?

2026-08-17 @ 21:06:10Points: 6Comments: 0

The Origin of Consciousness (2008)

2026-08-17 @ 20:12:58Points: 49Comments: 49

AI;DR (AI; Didn't Read)

2026-08-17 @ 19:47:15Points: 470Comments: 288

India has paved the way for charging merchants a fee on UPI transactions

2026-08-17 @ 19:25:11Points: 77Comments: 69

Los Puesteros, solitary men who look after ranches and livestock in Patagonia

2026-08-17 @ 18:34:09Points: 93Comments: 35

Roboflow Playground: Try and Compare 30 Computer Vision Models

2026-08-17 @ 18:28:07Points: 33Comments: 3

We Are Forking dotenvy into dotenv-ng

2026-08-17 @ 17:55:24Points: 34Comments: 34

GPU Offload in Rust: Portable, Safe, and Fast

2026-08-17 @ 17:54:59Points: 133Comments: 26

Qwen3.8 27B scores 52 on Artificial Analysis

2026-08-17 @ 17:25:17Points: 267Comments: 122

An update on leaving Gmail for Fastmail

2026-08-17 @ 17:15:20Points: 81Comments: 71

Sun Clock

2026-08-17 @ 16:37:54Points: 151Comments: 48

Judge sets framework for Nine PBS to retrieve archival data

2026-08-17 @ 16:11:37Points: 109Comments: 42

Launch HN: Speko (YC S26) – OpenRouter for Voice AI

2026-08-17 @ 15:36:18Points: 84Comments: 51

Demo: https://www.youtube.com/watch?v=no2LY2gRh-c

Typical production voice agent is an ensemble of three models: STT, an LLM, and TTS.

Each of those layers offers a dozen credible vendors, and each month there are new models on the market. Almost everyone evaluates once, picks a stack of their choice, and never rechecks because switching from a vendor to another involves yet another integration and arguments about the numbers.

The result is that you use voice agents running last quarter's models while better and cheaper options are available.

Before founding Speko, I spent four years as cofounder and CTO building voice agents for enterprises across Asia in 10+ languages. Each time a new speech model would arrive, we repeated the same ritual: hire native-speaking raters, benchmark it against our existing stack, and update production if it improved. Speko turns this process into an API. A team running thousands of calls a day told us: "we can literally go to this dashboard, switch the model, and it will do it for us."

How it works: you send a request with your optimization criteria (accuracy, latency, cost or balanced), language and region. The router filters to models which we measured for the given combination of constraints, benchmarks them, selects the winner, and returns a response with headers containing provider, model names, and the scores. The gateway prefetches signed session plans, so a new session dials the provider straight from memory; no control-plane round trip while a caller waits.

Failover happens only during connection setup stage: if the provider refuses the connection attempt, we start connecting to the runners-up.

Some of the customer stories: one founder came to us not knowing what to pick at all: he gave us his use case and now routes everything through the platform. A property management AI runs LiveKit in Python and had not updated STT or TTS since launch: they did not know their STT had high error rates on their calls, better options existed, and swapping always looked like an R&D project. One team did not know which models to pick for Spanish. A medical team did not know which STT handles medical vocabulary best. In every case we helped find the right stack from the benchmarks, and now they route through us.

The measuring part is public: we pass the same inputs to every model in one region in different dated runs and we publish the boards, including those where our selections perform worse than alternatives. A launch demo answers which 30-second clip sounds better; production asks which model survives minute eight, so we test spontaneous speech, money and dates, ten-minute takes, and the rankings change. We trained an automatic scorer for TTS naturalness on our blind head-to-head listening votes; on providers it has never seen a vote for, it picks the same winner our raters do about as often as raters agree with each other.

We don't train or sell models ourselves, that's precisely how we keep our rankings impartial.

We also open sourced the gateway for teams who want to avoid an extra network hop on the audio path and don't want to share keys with our cloud (https://github.com/SpekoAI/gateway, MIT): one Go binary, which is running as a sidecar in your agent's container, speaks one local protocol over Unix socket, pins provider hosts and attaches your keys. In BYOK mode it doesn't communicate with us at all.

Notice that the anonymous, content-free telemetry is enabled by default, and one env var disables it.

Cost: the gateway and BYOK setup will be free forever, we charge for the hosted router and managed keys with consolidated billing. Since we started the batch in late June, external usage has grown about 25 percent per week on average, front-loaded toward the launch weeks.

I would love feedback from the community: how do you pick speech models now, and what makes you trust the third-party benchmark?

https://speko.ai/

How to put 170 atoms in an atom

2026-08-17 @ 14:21:25Points: 96Comments: 20

AI-Generated GitHub Copilot “Autofix” Allowed Compromise of Snowflake's Jira

2026-08-17 @ 14:18:38Points: 295Comments: 120

How to disable or avoid intrusive AI

2026-08-17 @ 14:07:56Points: 233Comments: 128

Ask HN: Alternatives to GitHub

2026-08-17 @ 13:59:17Points: 462Comments: 293

Github has been down consistently over the last few months - does it make sense to switch to alternatives?

A Preview of DuckDB v2.0

2026-08-17 @ 13:46:27Points: 497Comments: 86

Incident with Github.com

2026-08-17 @ 13:35:06Points: 482Comments: 867

Edit: at the time of posting there was not an incident on githubstatus.com. Now there is. https://www.githubstatus.com/incidents/zkxwbgr0cnmx

Original Title: "Tell HN: GitHub Is Overloaded"

GPT 5.6 Sol is the best "vision" model OpenAI ever released

2026-08-17 @ 12:09:42Points: 287Comments: 149

A simple fix for LLM tail latency

2026-08-14 @ 05:58:16Points: 25Comments: 11

The oldest bar in every state

2026-08-13 @ 18:14:06Points: 50Comments: 29

Marketers are Addicted to Bad Data (2020)

2026-08-13 @ 16:47:40Points: 35Comments: 41

Intriguing Stories in Computer Science

2026-08-12 @ 21:56:40Points: 15Comments: 1

A particle made of force: physicists say they've found mysterious 'glueball'

2026-08-12 @ 14:26:41Points: 86Comments: 5

Olo (Color)

2026-08-12 @ 10:26:14Points: 265Comments: 61

Archives

2026

2025

2024

2023

2022