This is awesome. I really want to get something like this integrated directly into my agent IDE (https://getness.dev).
The main thing I need is some way of auto-sending a "end of message" button (like a newline would do) at the end of the message so that it auto-sends the chat.
Not sure the exact right way to implement that, but it seems like something that could be general purpose. Maybe a setting per-active app?
If folks like local-AI dictation, but want it optimized for meetings recordings (like Granola/Otter/Notion) check out Biscotti. Free, local, separates voices, voice identification, AI summaries, etc.
To be honest, I used Wspr so much I only recently started realizing just HOW bad it gets trying to 'improve' your diction, especially coding. I have seen totally opperate directives. I had no idea it was potentially occuring, but it gets past a point where it is WAY to confident in guessing what you meant, sometimes to the extent of changing DON'T to DO, for example.
This is awesome. I really want to get something like this integrated directly into my agent IDE (https://getness.dev).
The main thing I need is some way of auto-sending a "end of message" button (like a newline would do) at the end of the message so that it auto-sends the chat.
Not sure the exact right way to implement that, but it seems like something that could be general purpose. Maybe a setting per-active app?
Push-to-talk?
I expected dictation to be much more useful on the computer, but since I can type decently fast I find myself using it far less than I expected.
If folks like local-AI dictation, but want it optimized for meetings recordings (like Granola/Otter/Notion) check out Biscotti. Free, local, separates voices, voice identification, AI summaries, etc.
https://github.com/scosman/Biscotti
Been using Handy I am pretty happy but good to see more options, specially OSS.
Seems everyone is cobbling these together. Here's a local-only one from Google: https://apps.apple.com/us/app/google-ai-edge-eloquent/id6756...
This reminds me very much in terms of features as https://github.com/cjpais/Handy https://handy.computer , what's the use case for this over that?
I recently switched to Voiceink which also does similar. I'm trying to find the edge this has over that but can't find it
VoiceInk took a little time compiling but it's been working flawlessly for the three people that use it now for dictating consultation notes.
They both use parakeet so I guess they would perform similar.
How good is parakeet (assuming it’s Sota oss) compared to wisprflow? Wisprflow is magical ime
ye parakeet redux has been my new fav, that and tdt v2 both working well for my own clone app that was made.
I'm a bit confused, isn't "Wispr" a trademarked term? Why would you pick a name that could easily be shut down?
This came up a month ago - I forked from something now gone - https://github.com/NickJLange/parrot
And if MacOS 26 does it right - all of this is no longer necessary for us to vibe fork/ code on weekends :-)
P.S. Handy looks slick. May be able to ditch mine
The native / built-in API SpeechAnalyzer (introduced last generation on iOS and MacOS) works quite well.
Are folks finding it lacking / needing alternatives?
Voice shortcuts etc
This is pretty neat, I built something similar a couple months ago. I'd say try to support Windows or Linux next.
To be honest, I used Wspr so much I only recently started realizing just HOW bad it gets trying to 'improve' your diction, especially coding. I have seen totally opperate directives. I had no idea it was potentially occuring, but it gets past a point where it is WAY to confident in guessing what you meant, sometimes to the extent of changing DON'T to DO, for example.
How does it compare to handy ?