voice
43 stories and discussions about voice, aggregated from every source we track.
It sends encrypted signals with no satellites or relay towers required
Large Language Models (LLMs) tend to add disclaimers like "I'm just an AI" when asked about something related to themselves. The self-reports from such responses are used in debates about AI safety or…
Yesterday, we released new Gemini Live models in the Gemini API and Google AI Studio, expanding our...
Smart headphones, better voice recognition, language models, and contextual data have set the stage for audio-centric augmented reality.
This video (the script, the voice timings, the source code, the storyboard, the briefs the subagents...
The voice-first physical AI interface for modern productivity — push-to-talk dictation into any app, replies from your agents, and rate limits right on the glass.
This post was originally posted on X, by Annie Wang, Developer Relations Engineer, Google Cloud and...
You think faster than you type. Hold a key, say it, and Fluentry types the words — with the speech model running on your own machine.
Kenjiro Tsuda wins case against TikTok account in landmark Japanese verdict to protect publicity rights
Your coding agents, side by side, by voice. Claude, Codex & Grok in one desktop app, with Jev for the fast decisions. - reindent/jauvex
open-source speech AI platform for organizations that cannot send sensitive conversations to a third party - nanosamurai/nanosamurai
The 2026 Ruby on Rails Community Survey is now open. Add your voice to the community data. Brought to you by Planet Argon.
Windows desktop voice dictation app that types transcribed text directly into whatever app you are focused in. Local/offline speech-to-text via faster-whisper, no cloud API. - ahmedhmam1994/voxscri...
ASTRO — self-hosted, self-improving voice AI agent (PL): Raspberry Pi + Hailo NPU, wake word, STT/TTS, vision, LoRA self-training. Offline-first, open-core. - mobium-app/ai_astro_public
We benchmarked 11 speech-to-text engines on 265 real recordings with a competing voice in the room. See where voice isolation helps.
Self-hosted voice agent runtime: speech-to-text, an LLM and text-to-speech streaming into each other in one process. - SamarthUrs18/fusion-runtime
With AssemblyAI's industry-leading Speech AI models, transcribe speech to text and extract insights from your voice data.
Talk to Claude Code while it works. A full-duplex voice companion powered by GPT Live 1. - CakeCrusher/full_duplex_code
Jev reads the message, Gradium Voice Design gives the agent a voice that fits. Every step on the critical path is traced.
Find files, send emails, schedule meetings, and more with your voice. Explore Action Mode, now in beta, with Voice Cursor.
Read articles, PDFs, ebooks, and stories aloud with a voice that fits. Follow along with highlighting. Try free; 200K characters/month with a free account.
Sage The Great & The Ancient Observatory. Authoritative classical Pythagorean Numerology, Western Astrology, and Live NASA Celestial Patterns.
The American Radio Relay League (ARRL) is the national association for amateur radio, connecting hams around the U.S. with news, information and resources.
Why voice is the next capability overhang.
Robot writing has no voice. That's why you find it annoying. Like it or not, your brain is currently one of the best organic pattern matchers in the known solar
Multilingual push-to-talk dictation for Omarchy and macOS (experimental). Choose your primary and secondary languages, or enable automatic language detection. Powered by ElevenLabs Scribe; transcri...
Transcribe audio, translate languages in real time, and bridge spoken thoughts with code and team docs—100% free.
Your coding agents, side by side, by voice. Claude, Codex & Grok in one desktop app, with Jev for the fast decisions. - reindent/jauvex
macOS menu bar app that translates selected text in place with a hotkey: up to three language pairs with automatic direction, your writing style, and Claude, ChatGPT or Grok via your subscription o...
Build a deterministic Python commit gate that combines finalized speech segments, rejects duplicate callbacks, invalidates work after reconnects, and keeps stale AI replies out of your voice companion.
Extract voice style embeddings from any WAV for SupertonicTTS — no style encoder needed. - kdrkdrkdr/supertonic.embed
Practice presentations with a voice-activated teleprompter, automatic scrolling and speaking pace alerts. Write or import your script. Try Storia without an account.
Speak a thought and hear it continued in your own cloned voice. Speech recognition, AI-generated continuations, and voice synthesis run locally in your browser.
Hear a practice before you record it, read along in a voice you choose, or send a tape to someone. Recordings stay encrypted on your device, with no account.
reotoi turn your voice into original visual art.
This is a fictional short story and no claim made reflects any real world event.
Deepgram Speak brings the people building voice AI together for one day, in one room. San Francisco, October 29, 2026. Reserve your spot.
Late last year we shared Resenha, an experiment that added Discord-style voice and video rooms directly into Discourse. No external apps required, just peer-to-peer WebRTC in your sidebar. The experiment worked. We’ve…
Gladia joins OVH Groupe's AI Lab, combining its cutting-edge voice AI platform with OVH Groupe's sovereign cloud infrastructure.
EssilorLuxottica, the company that Meta partnered with for its Ray-Ban smart glasses, announced a new version of its Nuance Audio glasses that are designed to double as discreet over-the-counter hearing aids. The new Nuance Audio Plus still manages to squeeze their FDA cleared hearing aid tech into temple arms that don't look much thicker than what you'll find on a standard pair of glasses. But the new Plus option introduces improved sound boosting, better voice clarity, extra on-device controls, and up to 10 hours of battery life. The Nuance Audio Plus are now available in the US in three frame styles including Square, Pathos, and Upturn i … Read the full story at The Verge.