Launching today

MosMos
Voice writing that works before, during, and after meetings
205 followers
Voice writing that works before, during, and after meetings
205 followers
MosMos goes beyond voice dictation by turning both individual thoughts and group conversations into usable writing. Speak naturally in any app, get fast and accurate text in the style you choose, or ask MosMos to search the web for current information. Its personal glossary remembers specialized terms after you add them once. In multi-speaker meetings, it tracks precise timestamps, distinguishes speakers, and creates structured notes, summaries, decisions, and action items for easier follow-up.







MosMos
NexaSDK for Mobile
I understand speech recognition runs locally.
For meeting summaries and text polishing, I wonder what content is sent to external services?
MosMos
@power_valsha Thanks for asking!
For text polishing, we send the text or transcript being refined, your instructions, language or formatting preferences, and the limited context needed to produce the result.
For meeting processing, this can include meeting audio and transcripts, along with the meeting title and processing instructions. Local speech recognition applies to regular dictation; meeting features can involve temporary cloud processing.
MosMos
@power_valsha Thanks for asking! We offer a local ASR option for regular dictation. With local ASR selected Voice Input stays on-device. While Meeting uploads the recording for speaker-separated transcription and processes the resulting transcript in the cloud.
JobJump- AI Interview Copilot
Hi, can users choose between verbatim transcription and polished writing?
MosMos
@stan_xu1 Hi, Stan! Thanks for asking!
Yes, you can choose this under “Polish level” in Voice Typing.
The lowest setting gives you the original transcription without rewriting, while higher settings refine the wording and structure.
You control how much polishing is applied.
MosMos
@stan_xu1 Hi Stan—just to clarify, by “verbatim,” do you mean the raw ASR transcript without AI polishing, or a strict word-for-word transcript that also preserves fillers, repetitions, and self-corrections?
MosMos
Hey Product Hunt,
I’m Alex, the product lead at MosMos and a full-stack developer. My day moves between product discussions, meetings, writing, and building software.
I do a lot of vibe coding, often with several coding agents running at once. Typing instructions into every window quickly becomes the bottleneck. Voice lets me explain requirements, add context, and iterate with multiple agents much faster.
That’s why we built MosMos: I could think faster than I could type.
MosMos is a native macOS voice workspace. Press Fn anywhere to speak, and it turns your thoughts into polished text. It adapts the writing style to the app you’re using and remembers personal vocabulary and technical terms.
You can also use Ask MosMos for quick questions and web searches with reviewable sources.
For meetings, MosMos records multi-person conversations and produces speaker-separated transcripts, summaries, and action items. Processing runs in the background, so you can keep working.
We’re building MosMos for product managers, developers, operators, and anyone who needs to turn conversations into documents, instructions, and next steps.
MosMos is still early, and we’d especially value feedback on:
Voice input speed and writing quality
Multi-agent coding workflows
Meeting summaries and action items
Thanks for checking out MosMos. I’ll be here throughout the launch to answer questions and learn from your feedback.
How much do you actually have to clean up after speaking something out?
MosMos
@sukumar_sukumar1 Thanks for asking!
It depends on how rough the spoken draft is and the polish level you choose.
Higher settings help clean up the wording and structure, while the lowest keeps the original transcription.
I’d still give the result a quick read before sending, especially names, numbers, and technical details.
Being able to check the sources it finds is a good touch. I’d definitely want that before using anything for research.
MosMos
@vhim_jana Absolutely!
You can revisit the source links in History to read the original pages and check the context. We wanted that verification step to be easy, especially when you’re using an answer for research.
MosMos
Hi Product Hunt 👋 Lego here, technical lead at MosMos.
Building a voice-input demo is easy. Making it fast, accurate, and dependable all day is much harder.
We focused on low latency, polished output, and privacy-conscious processing. Speech recognition, personal vocabulary matching, and active-window detection run locally. MosMos uses the active app or window as a cue for polishing, and we do not retain users’ audio.
MosMos can also search the web for current information and return answers with sources.We see dictation, meetings, and tasks as one workflow—whether you’re speaking to write, think together, or capture an action.
Press Fn+Space and say, “Meeting at 3 PM.” MosMos creates a to-do and reminds you at 3. You can review tasks in Records and check them off when done.
Multi-person meetings are a core part of MosMos. It separates speakers, adds precise timestamps, and turns conversations into structured notes. You can trace any point back to the right person and moment, then create minutes, project docs, follow-up emails, or next steps without replaying the whole meeting.
Try MosMos with a niche term, unusual app, or chaotic meeting—and tell me what breaks. I’ll be here all day for technical questions.