Hacker Newsnew | past | comments | ask | show | jobs | submit | macinjosh's commentslogin

This comment is literally off topic. You opened the comment saying as much. So I have dutifully downvoted it.


Right now we use a special algorithm to rephrase the submitted text without changing or losing meaning. Once Anthropic documents how to detect the watermark itself work to remove it more surgically can commence.


Good for the EU. I am not subject to their authoritarian decrees, thank God.


This is a privacy oriented web utility for removing claude's watermarks from its text outputs and image outputs.

Image files are processed locally in the browser. It removes the C2PA signals from them directly.

Text inputs use an open model (Nemotron) on OpenRouter with Zero Data Retention enabled, no logging of inputs or outputs is done on this site or by OpenRouter or their providers.

It is 100% free to use and is ad-network free.


I feel like there will be a minimal UI/no screen but it will have 9 far field microphones. 2-3 cameras , an iris scanner, GPS, and lidar. This is because it is a data collection device for use in OpenAI's future models, more than it is a an assistant.


I make a podcast client for iOS that detects and skips ads for you using an on device model. This allows me to sell without a subscription. I am working on v2 which will bring support for video podcasts and YouTube podcasts.


This one hurts. We will miss you Mr. Dvorak! RIP


This project is open source and use the excellent SpeechAnalyzer API on macOS.

100% free and local.

https://github.com/leftouterjoins/voicewrite


But this will only ever be as good as the iOS / macOS dictation prediction. Which is OK but not really the point of what I was trying to do.

If you wanted that, you could just use accessibility mode built in to your mac.


I make an iOS app that uses this API heavily for transcribing diverse audio of varying bitrate and recording quality. The audio often contains music, multiple speakers, sound effects. SpeechAnalyzer almost always gets it.

It can struggle with proper nouns but will return something phonetically similar.

My main gripe is that it requires a separate model download per language. I understand the why they did this (to save disk space). But it makes multi-lingual audio hard to transcribe unless you know ahead of time the languages in the audio.

As an app developer the biggest win from using Apple's model is I don't have to bundle it in my app so my app looks much smaller. If a user has many transcription apps each one could have their own model. If Apple's model is used only one copy is needed.


> This is a prime example of why programmers are not seriously considered engineers.

The civil engineer who builds a great suspension bridge probably looks down on the one who builds a bridge over a irrigation ditch in a rural county using a big metal pipe covered with dirt.

Much like you may look down on train builders who make the novelty trains for kids parks.

Software engineering happens to be useful everywhere and most stuff in life is low stakes and the economics do not exist to make it perfect.

However, in aerospace, banking, and other high stakes industries software engineering projects are met with the rigor that is called for.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: