Hacker Newsnew | past | comments | ask | show | jobs | submit | coder543's commentslogin

There is also a new, larger voice to text model available... but not for an old iPhone like the 13 mini. Only like the iPhone 17 Pro and above.

The best workaround is a third party app that lets you run Parakeet V3 or Whisper on your phone. There are quite a few.


Dictus is one that I’ve been playing around with.

This is GREAT! It’s not possible to enable local-only dictation on iOS with the OS alone. :(

It’s by far the best speech to text I’ve ever used so I’m a fan

Can you clarify, you like third party? Or you like the larger Apple model

Some OpenRouter providers do not implement reasoning levels for these models correctly at all: https://www.reddit.com/r/DeepSeek/comments/1vdqjwr/openroute...

If you're going to use OpenRouter to test reasoning levels, always make sure you are locking to the official provider instead of third party providers.


That's a useful tip, thanks.


"Dictionary gains are mostly effective in the first few KB."

https://facebook.github.io/zstd/index.html

Pretrained dictionaries have never been intended to help with book sized or bigger compression. zstd automatically learns the most efficient dictionary it can within a few kilobytes. Pretrained dictionaries are only useful when you're independently compressing very small records.


Please run GLM-5.3 and GLM-5.3-Flash. I would love to see how they do. On the smaller end of things, Qwen3.8-27B and Ling-3.0-Flash would also be interesting.

In the benchmark, have you considered instructing the models to build their own SPICE simulations to test their work? Simply asking them to write and run simulations could improve performance, even without telling them what to simulate.


FunctionGemma never worked well for me (without fine tuning). Liquid has released 230M and 350M models that work far, far better in my testing: https://huggingface.co/LiquidAI/LFM2.5-230M

I really look forward to a hypothetical LFM3-230M, because LFM2.5-230M is so close to being usable, while FunctionGemma is miles away from being usable.

But, yes, still tangential to TERMy.


Novita does not offer Hy4-preview on either OpenRouter or their own model list. Maybe you confused it with Hy3.


You are correct.


I haven't tried it, but this looked promising for that exact task: https://huggingface.co/superwhisper/s1-mini


The spark can easily run UD-Q4_K_XL on this model... using IQ1_S doesn't make much sense.


No. What else could it reasonably be named? Hard to imagine.

The rule has always been intended to cover types that have another word in them but still choose to pointlessly repeat the package name.

`uuid.UUIDGenerator` is a hypothetical example of the anti-pattern that would instead be better named as `uuid.Generator`.


> No. What else could it reasonably be named? Hard to imagine.

Ocaml often just has it be T

so uuid.T


uuid.Entifier


The post does not imply the 5090 is needed, that is just a common reference point.

A single six year old RTX 3090 works great: https://www.reddit.com/r/LocalLLaMA/comments/1vkm42m/muse_gl...

I fully expect Meta will release other, smaller Muse models in the near future too.

The 5090 is also supposed to be a $2000 GPU, not a $5000 one. The entire market is utterly distorted right now, which will impact cloud inference more and more over time too. They are not immune to the absurdly high RAM prices, so their prices will have to go up over time too until the RAM supply chain goes back to normal.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: