Hi AInauts,
Welcome to a new edition of your favorite newsletter!
Last week, we mentioned in passing that we had replaced Wispr Flow with our own app. We received more messages about that one sentence than almost anything in recent memory.
Nearly everyone asked the same thing: βHow? And can I have it?β
Yes, you can. But today's issue is really about something bigger.
Anyone can build now. Or generate: images, voices, videos, and even someone to talk to on a live video call. The tools are no longer the limit. The real limit is whether we come up with smart ways to use them.
Here is what we are covering today:
ποΈ We rebuilt Wispr Flow. Get our version.
π The old avatar fooled one in 41 people. The new one fooled almost half.
π§ From user to builder: How to create your first app
Let's get started!
How 2M+ Professionals Stay Ahead on AI
AI is moving fast and most people are falling behind.Β
The Rundown AI keeps you ahead of the curve.Β
It's a free AI newsletter that keeps you up-to-date on the latest AI news, and teaches you how to apply it in just 5 minutes a day.
Plus, complete the quiz after signing up and theyβll recommend the best AI tools, guides, and courses β tailored to your needs.
ποΈ We rebuilt Wispr Flow. Get our version.
A quick reminder: Wispr Flow is a dictation app. Hold a key, speak, and release it. Cleanly written text appears wherever your cursor is. We loved the app, used it every day, and dictated more than a million words with it.
Today we use our own version: AInauten Voice.
It does exactly what we need:
Entirely local: Your voice never leaves your Mac. NVIDIA's Parakeet handles speech recognition, and a small Qwen3 model cleans up the text.
No subscription: Wispr Flow costs $180 a year, and it was absolutely worth it. Our version is free.
Your setup comes with you: It imports your dictionary, replacements, and keyboard shortcuts from Wispr Flow.
How to get it
You have two options: Download the installer from voice.ainauten.com and install it directly on your Mac. Or use the source code and let your agent install the app for you.
We recommend the second route. We have only tested the app on our own Macs, so your agent can inspect the code first and adjust it if anything causes trouble.
Use the ChatGPT or Claude desktop app for this. Do not use a normal chat. Use Work or Codex mode in ChatGPT, or Claude Code.
Those environments can create files and build programs on your computer. You will generally need a paid subscription.
Option 1: You have a Mac with Apple silicon. You need an M1 or newer, macOS 14 or later, at least 8 GB of RAM, and about 3 GB of space for the speech models.
Give your agent this prompt:
Install the AInauten Voice dictation app (https://voice.ainauten.com) from this repository: https://github.com/MediaPublishing/ainauten-voice
Inspect the code first and briefly explain what the app will be allowed to do on my Mac. Ask me before installing anything.A setup assistant then walks you through the rest: downloading the models, importing Wispr Flow settings, granting permissions, and running a test dictation. Wispr Flow is only closed, not deleted, so you can switch back at any time.
You should never install an app that listens and types into other programs without inspecting it first. That includes ours. This is why the prompt says: inspect first, then install.
Option 2: You use Windows, an older Mac, or want your own version. Ask your agent to build one:
I want an open alternative to Wispr Flow, similar to the AInauten Voice dictation app (https://voice.ainauten.com) in this repository: https://github.com/MediaPublishing/ainauten-voice
Build it for my device ([Windows / Mac with an Intel chip / ...]). Ask me the questions you need first, then create a clear specification and plan.Is the app perfect? Probably not. It is a weekend workshop project, not a polished commercial product. It works for us, but you may still hit a snag. If that happens, send your AI a screenshot and ask it to fix the problem, or send us feedback through the app.
How we built it
The app did not appear in 30 minutes, especially because the agent needed time to write the code. But the process was not difficult.
Specification. Start with a brain dump: What should the app do, and what should it explicitly not do? The AI turns that into a clean specification and a clear goal. This is the most important step.
Build. ChatGPT Work/Codex spent a few hours building while we worked on other things.
Test and give feedback. Then came several rounds of testing, with shorter and shorter intervals. The dictation bar was initially three times larger than Wispr Flow's. German speech was recognized as English. We took screenshots, described the current and desired result, and sent the feedback. Once the prototype worked, Claude reviewed the code and suggested improvements. Then we did one more refinement pass.
The third story below shows you how to build your own app.
Here are a few practical tips that are useful well beyond coding:
Tip 1: Put feedback in the queue. You do not need to wait until the AI has finished before giving feedback. But you also do not need to interrupt it and push it in a new direction.
We like ChatGPT's queue feature. In the desktop app, open Settings, then General, Composer, and Follow-up message behavior. Choose Queue.
The AI then processes your feedback in order. To send one message immediately, hold the Command key while sending or use the arrow's Steer option.
Some people even keep their system running overnight by queuing several βContinue and improve itβ instructions. That can work, but it may also produce surprises the next morning. π
Tip 2: Use the right model for each step. We use the strongest model for the specification and review, such as GPT Astra or Claude Fable/Opus. A faster model, such as GPT Sol or Claude Sonnet, is usually enough for the actual implementation.
This is not just our intuition. Anthropic recommends the same pattern in its Advisor strategy, and Claude Code even offers an opusplan setting for it.
Tip 3: Run the work as a goal. An AI normally stops when it decides it is finished. Often, the app is only half built at that point. A goal lets you define what finished actually means. Use the /goal command.
Before stopping, the AI checks whether the goal has been met. It keeps cycling through build, test, and fix until every requirement passes or it reaches a decision only you can make. This produces better results, but it also uses more tokens.
P.S. We understand that dictating out loud is not ideal in every environment. But there is a solution for that too, at least as long as nobody is looking directly at you. π
P.P.S. Yes, that is in our app as a beta too. Unfortunately, it does not work in German yet.
The agentic era needs a different CRM. Thatβs Attio.
Parallel, Turbopuffer, and Wordsmith run their entire GTM motion on Attio, with agents that chase every buying signal, build pipeline, and move deals forward, 24/7.
π The old avatar fooled one in 41 people. The new one fooled almost half.
We have seen a lot by now. But this genuinely impressed us.
A few days ago, Tavus introduced Griffin, an AI counterpart for live video calls. This is not a prerecorded video. Griffin listens while speaking, nods, says βmm-hmm,β and does not interrupt just because you pause to think.
All it needs is one photo and about ten seconds of your voice.
The test
Fifty-four people were sent into a one-minute video call. They were told they were speaking with another test participant. Only at the end were they asked whether the person on the other side might have been AI.
Twenty-six of 54 were convinced it was a human. That is 48%.
For comparison, Tavus's previous system fooled only one out of 41 people in the same test. This was Tavus's own study, the sample was small, and the call lasted only one minute. But watch the videos and you will immediately see how realistic they are.
First text, then voice, now the face
Live avatars already existed, including HeyGen's version. Griffin still feels like another level. Tavus is therefore releasing the first version only to selected testers. We applied and will report back if we get access.
Griffin did not arrive alone. Several releases in the past few days have made AI output even more realistic:
Voice: ElevenLabs v4 adds more expression, supports 90 languages, and can clone a voice from ten seconds of audio. We also used it in this promo video.
Images: FLUX 3 Image from Black Forest Labs, with a demo here, and Ideogram 4.5. Both can change individual details without shifting the rest of the image.
Video: HeyGen Video creates five- to 15-second clips with sound. During October, it costs one cent per second, so a ten-second clip costs ten cents.
These tools now cost almost nothing, and your agent can set them up for you. You still need to decide what they should be used for. That brings us to the next story.
π§ From user to builder: How to create your first app
To finish, here is an idea that has been on our minds over the past few days.
We used to learn software. Which button does what? Where is the feature in the menu? Traditional software has a feature list. You can count the items.
AI does not have a feature list. It is an open field. You have to imagine what it could do for you.
From user to builder
Ralf from our community put it well last week. He had given an introductory AI session, and nine of the 11 participants had barely used AI before.
One question dominated at the end: βSo how exactly do I use AI in my own workplace?β His conclusion: βFor the first time, mainstream software expects people to imagine the features themselves.β
There is something to that. We were trained to be users. If the tool cannot do something, it cannot be done. A builder asks a different question: not βWhat can this tool do?β, but βWhat could I do with it?β
That is exactly how our dictation app came to exist. We did not search for the right tool. We built one. You can do the same without programming skills, often with the ChatGPT subscription you already have.
You may be thinking: βI already build my own apps.β If so, well done. You are already a builder and far ahead of many people around you. Take the prompt below anyway. It might contain your next idea.
For everyone else: You are one prompt away from joining them. Give the first step 15 minutes. The AI can build the rest.
What you can build
An HTML page: a redesigned homepage, a quiz for your team, or an interactive learning element. One file, double-click, and it runs in the browser.
A landing page or web app: for your offer, club, or event, complete with text and images. You can publish it with ChatGPT Sites or Claude Artifacts.
A browser extension: a small helper that supports a process in your browser. Rebuild something you already use, then improve it.
An app inside ChatGPT: The new Plugin Extensions let your own app run directly in the ChatGPT window. We tested them, and they work beautifully.
A local app: like our dictation app, running directly on your computer. Use ChatGPT Work/Codex or Claude Code for this.
Rule of thumb: Start with the smallest format that solves your problem.
How to start: One prompt for almost anything
The sequence matters. Start with the idea, then create images or mockups, then write a clear description, and only then start building.
Paste the following prompt into ChatGPT Work/Codex, not a normal chat, or Claude Code/Cowork in the desktop app, also not a normal chat.
Use the best available model to begin. You can switch to a faster model for the implementation later.
You are my builder coach. I cannot code, and I want to build my first small app today.
1. Idea: Independently review the chats and memories you can access without asking me first. Look for recurring manual work, concrete frustrations, postponed projects, and personal interests. For relevant findings, read the context and check which solutions I already have. If you cannot access chats or memories, ask me three short questions about my work and everyday life instead.
Suggest the five to ten strongest and most distinct app ideas: HTML page, web app, extension, local app, and so on. Do not suggest generic ideas without evidence of a real need. At least one idea should surprise me.
For each idea, provide:
- the specific trigger from my chats or memories, with the source;
- the benefit: what do I do today, and what would the app take over?
- a realistic example of the input and output;
- the format and effort: small or medium.
2. Mockups: Once I choose an idea, create a few mockup images showing what the app or system could look like.
3. Specification: Ask no more than three questions per round and include suggested answers. Then write a short specification: what the app does, what it deliberately does not do, and three to five acceptance criteria that can be tested. Plan a first version that runs quickly. Extras come later.
4. Goal: Write a goal that is complete only when every acceptance criterion has been tested. Wait for my βGoβ before building. Then start with the smallest version and explain how I can open and test it.Do not give up too soon. The first attempt is almost never very good. Give feedback, take a screenshot, describe the current and desired result, and start another round.
One warning: It has never been easier to build something nobody needs. Ask yourself: useful or merely new?
Do not let that stop you. If it saves you time, money, or frustration, you have already won. And once you build something for the first time, an entirely new universe of possibilities opens up.
Tell us what you built. We are curious!
P.S. We are considering a step-by-step Deep Dive on vibe coding. If that interests you, use the rating below and leave us a comment.
That's it for today. See you in the next issue!
Reto & Fabian from AInauten






