Posts

New ask Hacker News story: Laguna S 2.1:118B-a9B better than Qwen3.5:122B-a10B? So far, yes

Laguna S 2.1:118B-a9B better than Qwen3.5:122B-a10B? So far, yes 2 by spottedmarley | 0 comments on Hacker News. I just found out about this new American (San Francisco) model today and I'm currently benchmarking it and, so far, it is outperforming my standard go-to model Qwen3.5:122b and even Sonnet in my benchmarking test. It's personality is much better than Qwen and it is more naturally creative when it is designing UIs. I have about 6 more tests to go though. If you're interested you can look at my benchmark dashboard here: https://ift.tt/tL8ia74

New ask Hacker News story: Ask GitHub SRE: How serious is the situation there?

Ask GitHub SRE: How serious is the situation there? 6 by laxk | 0 comments on Hacker News.

New ask Hacker News story: Anybody tried Laguna S 2.1 (by Poolside)?

Anybody tried Laguna S 2.1 (by Poolside)? 4 by spottedmarley | 0 comments on Hacker News. I just found out about it but I hadn't ever seen it mentioned and it sounds really interesting. Im pulling laguna-s-2.1:q4_K_M right now

New ask Hacker News story: Why Fireworks doesn't support Voice AI

Why Fireworks doesn't support Voice AI 3 by kushalpatil07 | 0 comments on Hacker News. I started thinking over why doesn't fireworks support voice models. There are really good opensource models available now, like parakeet, kokoro, Qwen ASR etc but no way to use it without managing a bunch of GPUs yourself. Even LLMs like Gemma 4 used by voice agents are not supported. Then I figured that the inference platform needs to be optimized differently for the kind of usecase you are using. Lets take an example for LLMs, not even STT and TTS. - Coding agents -> lot of cached input, needs to optimize for KV cache - Creation slides/blogs -> lots of output, needs to optimize for speculative decoding - Voice LLMs -> Cached input small output, not yet figured out on how to optimize this. So TTS and STT is a completely different ballgame. What I don't know is the timing, do people want to use open source models like kokoro, parakeet, Qwen etc RIGHT NOW?

New ask Hacker News story: Ask HN: Who Has This Pain?

Ask HN: Who Has This Pain? 3 by swk-phil | 2 comments on Hacker News. has to open media files from untrusted sources in sensitive environments as part of their daily workflow + fear that a zero-day exploit could be in one of those media files + cyber attack would have huge impact on the business

New ask Hacker News story: Ask HN: What was your big failure? How did you get around it?

Ask HN: What was your big failure? How did you get around it? 2 by jspann | 0 comments on Hacker News. Sincerely a guy who took a gamble on the last few years and is about to lose