Changelog
v0.0.116
Improvements
- Better keyboard shortcut display on macOS – On Mac, modifier keys in shortcut hints now show the correct macOS icons (like
⌘,⌥,⌃) instead of their Windows counterparts. Hotkey icons throughout the interface are also slightly larger for improved readability.
Bug Fixes
-
Clearer auto-submit description – Updated the instruction text in the settings to better explain the auto-submit behavior. The text now reads "Automatically submit after inserting the transcription" instead of the previous wording.
-
Tighter spacing for auto-submit instruction – Adjusted the alignment of the auto-submit explanation so it sits more comfortably next to its label.
v0.0.115
New Features
- Auto Submit section added to profile settings. You can now configure a key to press after transcription that automatically submits the text. This setting has been moved from the hotkey grid to its own dedicated section on the profile page, making it easier to find and manage.
Improvements
-
Hotkey labels updated for clarity. "Press & Hold" is now labeled "Push to talk", and "Toggle" is now "Press to start/stop" — making it clearer what each hotkey does.
-
Shift keys now show with an arrow icon. Instead of displaying "LShift" and "RShift" as text, the Shift keys are now shown with a compact "L" or "R" label paired with an up-arrow icon, matching the visual style of other modifier keys.
-
Better responsive layout for hotkey and language selection. The grid now adapts more gracefully across different screen sizes, ensuring the language dropdown and hotkey options remain well-organized on tablets and smaller screens.
v0.0.114
Improvements
- Refined how operating-system–specific configuration settings are handled. Settings like the whisper backend, injection method, hotkeys and auto-submit now reliably use their platform-specific keys (for example,
whisperBackend.windowson Windows orhotkeyHold.macoson macOS). If you edit your configuration file manually on Windows or macOS, make sure you are using the OS-specific key names – generic keys without the platform suffix are no longer read on those systems.
v0.0.113
New Features
- Toggle shortcuts for profiles. Each profile now supports two types of shortcuts: a "Press & Hold" shortcut (start recording while held, stop when released) and a separate "Toggle" shortcut (press once to start recording, press again to stop). You can set both or just one, depending on how you prefer to work.
- Auto-submit after dictation. After your dictated text is inserted, you can now choose to have the app automatically send it — just like pressing Enter, Ctrl+Enter, Shift+Enter, or other common submit key combinations. Configure this per profile in the profile settings.
- Press Escape to stop a toggle recording. If you started a recording with a toggle shortcut, pressing Escape will now stop it without needing to press the toggle shortcut again.
- Cancel button on hotkey fields. Each hotkey field now shows a small X button when you hover over it — click it to quickly clear the shortcut.
- Hotkey validation is now clearer. The Escape key can no longer be set as a shortcut. The app also prevents you from assigning the same key combination to both the hold and toggle shortcuts for the same profile.
- New language option: "Translate to English". The previous "Auto English" option has been renamed to make its purpose more obvious.
Improvements
- OiPer model names and descriptions refreshed. ASR models are now called "Pro" (most accurate) and "Flash" (fastest). LLM models have been renamed to a celestial theme: Auto, Earth, Moon, Mars, Jupiter, and Pluto. Each model now shows a short tagline beneath its name and a detailed description in a tooltip on hover.
- Model ratings now use a 1–5 dot scale. The quality, speed, and cost indicators for OiPer models have been simplified to show a clearer 1-to-5 rating.
- Nova model cost updated. The Nova (Flash) ASR model's cost rating has been reduced, making it a more budget-friendly option for near real-time transcription.
- Profile header simplified. Profile names now sit next to a one-line description of the mode — the mode badge icons and badges have been removed for a cleaner look.
- Doctor now warns about missing hotkeys. The Settings doctor will alert you if a profile has no hotkey assigned at all, so you never lose track of an unusable profile.
- Duplicate hotkey detection covers both shortcut types. The doctor now checks for duplicates across both hold and toggle shortcuts, and its messages clearly indicate which profile uses which type.
v0.0.112
New Features
-
Application context tracking for OiPer mode: When using an OiPer profile, the app now captures which application and window you're currently using (app name, window title, and visible text). This context is sent along with your recording for smarter processing. Context capture runs in the background so it doesn't slow down recording.
-
Application icons across the app: Application icons are now shown in several places — the Overview page's "Most Used" card displays the icon of your most frequently used app, and the History details dialog now shows the app icon next to the application name. Icon support has also been added for Windows (previously only macOS was supported).
-
Revamped Overview cards: The overview page now has a four‑column layout on larger screens. A new "Most Used" card shows which application you use most and what percentage of sessions it appears in. Typing speed now displays as "X× faster than typing" (or "Warming up") instead of a percentage comparison. All "today" stats and the "…/transcription" labels have been replaced with cleaner "avg per session" summaries.
-
OiPer profile model and template selectors: Profiles in OiPer mode now offer two ways to choose models:
- Simple mode: Pick a predefined template (Default, Fast, Accurate, Slow) — each has a description and a unique icon.
- Advanced mode: Independently choose an ASR (speech recognition) model and an LLM (enhancement) model. Each model shows quality, speed, and cost ratings as easy‑to‑read dot meters.
- The enhancement model selector is only shown when formatting is enabled, keeping the interface tidy.
-
History details improvements: The details dialog now shows:
- The application you were using (with its icon) as a dedicated field.
- How long context collection took ("Context duration").
- The Whisper backend field now appears earlier in the metadata list.
- For OiPer mode recordings, transcription and enhancement are shown as a single "Processing" step in the timeline instead of two separate steps.
- A "Preparation" step has been added to the timeline, and all timeline colors have been refreshed for better readability.
- Duration formatting is now more human‑readable (e.g. "1 minute 30 seconds").
v0.0.111
New Features
-
New text enhancement presets: Casual and Romantic. You can now choose "Casual" to rewrite your text in a friendly, conversational tone, or "Romantic" to give it a warm, affectionate feel. All existing presets (Normal, Prompt, Formal, Email) are still available.
-
Task-based custom formatting. Instead of writing a full system prompt, you can now describe a simple task — for example, "translate to French" — and the AI will automatically apply a formatting template to produce the result. A new "Task" option appears alongside "Custom" in the formatting selector.
-
Toggle between Task and Advanced mode. When using custom formatting, a checkbox lets you switch between Task mode (describe what you want) and Advanced mode (write the full system prompt directly). A tooltip explains the difference at a glance.
Improvements
-
Better save feedback. The Save button now shows a floppy disk icon, and changes to a green checkmark with "Saved!" after you save — so you always know your changes have been stored.
-
Cleaner formatting selector. The preset buttons now use consistent icons throughout, making it easier to identify each option at a glance.
-
Sharper AI prompts. The text sent to the enhancement LLM is now structured more clearly, which helps the AI follow your formatting instructions more accurately.
Bug Fixes
- Fixed an issue where the enhancement system would fall back to the transcription model when no enhancement model was set. It now correctly uses only the enhancement model you selected.
v0.0.110
New Features
- Delete transcription records from the details dialog – You can now remove a transcription from your history directly from the details view. A confirmation prompt asks before deleting, so accidental removals are avoided.
- Navigate to a profile from the details dialog – Profile names in the metadata section are now clickable links that take you directly to the corresponding profile page.
Improvements
- Timeline redesigned as a visual progress bar – The processing steps (resampling, transcription, enhancement, injection) are now shown as a horizontal, color-coded progress bar with a legend underneath. This makes it easier to see at a glance how long each step took relative to the total.
- Smarter output display – The details dialog now shows the final output (the best available result). Separate transcription and enhancement sections appear only when they differ from the final output, reducing clutter when they are the same.
- New metadata fields – You’ll now see Recording duration and Processing time directly in the details. The Whisper backend used for transcription is also shown.
- Cleaner metadata layout – Information is presented in a compact two-column grid with clear labels, making it easier to scan.
- Renamed “Input Audio” to “Recording” – The audio player section now has a clearer label.
- Close button in the dialog footer – A dedicated Close button has been added to the bottom of the details dialog for convenience.
Removed
- The following fields are no longer shown in the details dialog: Status, Started, Completed, Duration (overall), Enhancement Method, and Enhancement Tech. This information was either duplicative or not commonly needed.
v0.0.109
No notable changes in this release.
v0.0.108
After carefully reviewing all commits, these changes are all internal build system and developer tooling adjustments:
- Commenting out CMake environment variables in the Makefile and release workflow (affects compiler flags during build, not runtime behavior)
- Simplifying a build log message in the Makefile
- Removing a custom build variable that was only set during local development (never used in release builds)
- Adding a
.envrcfile for developer environment automation - Reordering and adding CPU feature flags for local Windows development builds (release builds already had these)
None of these changes introduce new features, fix user-facing bugs, or alter the application's behavior for end users.
No notable changes in this release.
v0.0.107
Improvements
- Faster performance on Windows: Windows x64 builds now enable AVX, AVX2, FMA, and F16C CPU optimizations, delivering noticeably better speed on compatible hardware.
v0.0.106
No notable changes in this release.
v0.0.105
Improvements
-
Sound effects settings are now more reliable. If you had never changed a sound effect option before, it will still behave correctly with sensible defaults — no more unexpected silence or missing error sounds.
-
Whisper backend selector simplified. The dropdown for choosing your speech‑to‑text engine (Auto, CPU, GPU) no longer shows hint labels like “Recommended” or “Accelerated”, making the list cleaner and easier to scan at a glance. The icons have also been updated for a more consistent look.
-
Settings dropdowns are more compact. The width of dropdown menus in the General and Whisper Backend sections has been reduced, giving the settings panel a tighter, more balanced layout.
v0.0.104
New Features
- GPU-accelerated transcription — Speech recognition can now use your graphics card (GPU) for faster processing. The app supports Vulkan on Windows and Linux, and Metal on macOS, giving you accelerated performance when a compatible GPU is available.
- Simplified backend selection — The speech recognition backend setting has been streamlined. Instead of separate options for CUDA, Vulkan, Metal, and CPU, you now simply choose Auto, CPU, or GPU. The app automatically detects what acceleration your system supports, so you don't need to worry about the technical details.
Bug Fixes
- Fixed an issue where the recording start time could be reported incorrectly in certain scenarios.
v0.0.103
Improvements
- Recording status overlay now stays visible more reliably on macOS. The overlay window's display level has been adjusted so it appears above most windows, including full-screen apps, making it easier to see when a recording is active. The overlay still won't interfere with your work — it ignores mouse clicks and doesn't steal focus.
v0.0.102
Improvements
- macOS recording overlay overhaul: The recording status indicator now uses a proper system panel instead of custom window manipulation. This means the overlay reliably stays on top of other windows, works correctly in full-screen spaces, and no longer interferes with your keyboard focus or mouse clicks.
- Cleaner overlay appearance: The overlay window now has proper transparency and no drop shadow, for a more polished look during recording.
New Features
- Custom build identifier: If your app was built with a custom build environment variable (e.g., "staging"), a badge now appears next to the version number in Settings so you can easily tell which build you're running.
Internal Changes
- Removed several obsolete configuration files (
.npmrc,.gitattributes,AGENTS.md) that were no longer needed.
v0.0.101
Improvements
- Settings page redesigned – The Settings page has been reorganized into clear sections (General, Input Device, Advanced, About) making it much easier to find and adjust the options you need.
- Audio input device handling – You can now choose whether to always use your system's default microphone or manually select and prioritize specific devices. A new "Use system default" toggle makes switching between the two modes quick.
- Individual sound effect controls – Instead of a single "disable sound effects" switch, you can now turn each sound on or off independently: start recording sound, stop recording sound, error sound, and no-speech (silent audio) sound. Find them under the new "Sound effects" section in Advanced settings.
- New start/stop recording sounds – You can optionally enable a sound that plays when recording begins and another when recording ends. Both are disabled by default.
- Sound feedback timing improved – The start recording sound now plays after recording actually starts, and the stop recording sound plays after audio processing is complete, making the feedback more accurate.
- Logo shown in About section – The app logo now appears in the Settings page's About section alongside the version number.
- App name and version shown dynamically – The Settings page now displays the actual application name and version, correctly reflecting the current build instead of using fixed text.
- "Whisper Hardware" label – The speech-recognition backend setting has been renamed from "Whisper backend" to "Whisper Hardware" for clarity.
- New profiles default to English – When adding a new profile, the language is automatically set to English, saving you an extra step.
- Profile deletion navigates to overview – Deleting a profile now takes you back to the overview page, providing a clearer flow.
Bug Fixes
- Fixed sound playing before recording starts – The start-recording sound no longer plays before the recording has actually begun, ensuring accurate audio feedback.
- Simplified config directory setup – Removed the old migration logic that moved configuration from a legacy folder. The app now creates its configuration directory directly without relying on deprecated paths.
v0.0.100
Improvements
- Updated application icons and branding. Icons are now automatically generated during each build, ensuring crisp, consistent logos across all platforms. The development build's logo color has been updated from red to amber to better match the brand identity.
- Config directory has moved to a standard location. Your settings are now stored in the OS‑standard app config directory. Old settings from
~/.oiperor~/.oiper.devare automatically migrated on first launch — no action needed. - Downloads now show progress and can be cancelled. When the app downloads files (such as whisper models), you'll see real‑time progress updates, and you can cancel an in‑progress download if needed.
- Refined recording pipeline. The app now uses a single, more reliable vocal‑detection check to decide whether a recording should be processed, making the flow faster and more consistent.
Bug Fixes
- A "stop" sound now plays when a recording is skipped. If the app detects no voice in a recording and discards it, you'll hear a subtle sound so you know what happened.
- Fixed sound file paths on macOS. Sound effects now resolve correctly on macOS, preventing missing‑sound errors.
- Release workflows now generate logo assets correctly. Logo generation is now a proper step in the release pipeline, ensuring release builds include the correct icons.
Breaking Changes
- The configuration directory has changed. Your settings are now stored in the platform‑standard app config directory (e.g.,
~/Library/Application Support/com.oiper.desktop/on macOS,~/.config/oiper-desktop/on Linux). If you were relying on the old~/.oiperor~/.oiper.devpaths, the app will automatically migrate your data on first launch. No user action is required.
v0.0.99
New Features
- Choose and prioritize your audio input devices. You can now see all your connected microphones and audio input devices directly in Settings under a new "Audio Input" section. Drag to reorder them — the first working device in your list will be used for recording. Devices that are currently disconnected are shown with a "Not connected" badge but stay in your list so they're ready when you plug them back in. A "System Default" option is always available if you'd rather let your operating system decide.
v0.0.98
Bug Fixes
- Fixed audio playback in history details – Audio recordings can now be played back and revealed in your file explorer reliably, as the app now looks for them in the correct location.
v0.0.97
No notable changes in this release.