Hacker News

Latest

Show HN: Macros with a Behringer FCB1010 MIDI Pedalboard in macOS

2026-09-14 @ 23:01:36Points: 54Comments: 11

Charts built for Chat

2026-09-14 @ 21:22:54Points: 196Comments: 61

Amazon vs. Perplexity – U.S. Court of Appeals for the Ninth Circuit

2026-09-14 @ 21:05:22Points: 201Comments: 200

GPT-5.6 Luna vs. GPT-6 Astra: Is a $1.20 Model Good Enough for Code Review?

2026-09-14 @ 19:56:20Points: 140Comments: 127

Backprop Alternative: Augmented Lagrangian Predictive Coding

2026-09-14 @ 18:03:12Points: 78Comments: 22

iOS 27, iPadOS 27, and macOS 27

2026-09-14 @ 17:50:29Points: 533Comments: 579

Steam Frame starts at $1059

2026-09-14 @ 17:27:55Points: 623Comments: 464

Pion, an agent designed to run any company autonomously

2026-09-14 @ 17:16:06Points: 361Comments: 415

Cloudflare AKE cuts origin HelloRetryRequests from 52% to 3.7%

2026-09-14 @ 17:02:34Points: 104Comments: 28

Cua (YC P25) Is Hiring a Founding Technical GTM Lead

2026-09-14 @ 17:00:53Points: 1

Why don't machine learning research agents overfit?

2026-09-14 @ 16:32:23Points: 119Comments: 66

How my e-reader lost its stripes

2026-09-14 @ 16:23:16Points: 188Comments: 30

Microsoft patches Windows and Excel – breaks audio, remote access, and paste

2026-09-14 @ 16:09:45Points: 233Comments: 149

Show HN: Nari Qwen3-TTS and Qwen3-ASR – High accuracy, low latency and cost

2026-09-14 @ 16:07:58Points: 76Comments: 28

We've been working on making OSS speech models super-fast. Last year, we built Dia, the first OSS text-to-speech model capable of doing natural dialogue. Since then, so many more great speech models have been released to the public.

But the market is still dominated by closed source models. We think that's an inference problem. Existing systems such as vLLM / SGLang are not well suited for multimodal inference. To prove this, we built an inference engine specialized for Qwen3-TTS and open-sourced it (https://github.com/nari-labs/nari-qwen3-tts). Running at sub-50 ms latency at 10 RPS, this showed open models can be run much faster and cheaper.

Since then, we've been working hard to bring cheap, fast, and high quality serving to all. And we've even beat closed models at their game!

Measured on the highly cited Coval (YC S24) voice AI benchmarks, our Qwen3-TTS endpoint not just is #2 in latency, but #1 in accuracy (WER) compared to 11Labs, Cartesia etc. while being the cheapest endpoint. Our Qwen3-ASR endpoint has the lowest latency and #2 accuracy, just 0.1% away from #1. It is the second cheapest model on the list.

It took a lot of clever inference engineering to make these models quick, perform well while keeping costs low. Interestingly, Alibaba's official endpoints seem to perform worse in terms of accuracy and latency compared to ours. But nonetheless, much love to the Qwen team for OSS-ing these amazing speech models.

We want to continue to push prices down to make speech technology a commodity - so that every app can have great TTS and STT without worrying about unit costs. We're also working on other parts of audio such as diarization - as well as video and world model inference. More to come!

Show HN: Neobrutalism.dev – Just added Base UI support and added new color theme

2026-09-14 @ 16:02:02Points: 151Comments: 72

Distributed Systems Classics (2017)

2026-09-14 @ 16:02:01Points: 276Comments: 60

A Beginning for Mathematics

2026-09-14 @ 15:33:07Points: 196Comments: 109

Principles for Fast Tokio Applications

2026-09-14 @ 15:27:56Points: 192Comments: 47

Dario, Please

2026-09-14 @ 14:50:32Points: 437Comments: 207

Dropping eBPF CPU Cost by About 90% with Memoization (Not AI Gen)

2026-09-14 @ 14:29:23Points: 89Comments: 25

Notes on gotchas while migrating 35kb preprompts from Opus to self-hosted Ollama

2026-09-14 @ 13:59:09Points: 130Comments: 70

People who can't picture anything are rewriting the science of imagination

2026-09-14 @ 13:23:53Points: 147Comments: 208

Largest known Roman mosaic, beneath Baths of Trajan, opens to the public

2026-09-14 @ 13:01:04Points: 53Comments: 8

OpenAI bots knew about the RubyGems caching vulnerability

2026-09-14 @ 12:40:57Points: 426Comments: 343

XCancel service is suspended until further notice

2026-09-14 @ 09:51:50Points: 574Comments: 843

Ask HN: What are you working on? (September 2026)

2026-09-13 @ 17:31:38Points: 318Comments: 985

What are you working on? What have you been curious about lately?

4,400-Year-Old Tomb of Egyptian Judge Found at Saqqara with Colors on Walls

2026-09-12 @ 18:51:40Points: 96Comments: 19

Compressing a Flag to 11 Bits

2026-09-12 @ 15:57:43Points: 115Comments: 49

Optimizing a Spin-Lock

2026-09-12 @ 09:43:44Points: 59Comments: 29

Trying to Make a Loop Auto-Vectorize

2026-09-10 @ 04:50:25Points: 77Comments: 18

Archives

2026

2025

2024

2023

2022