FPVSIM Timer 5.2 brings more natural AI event broadcasting to the local computer: different sounds can be switched in Workspace Settings, and prompts such as name, next group, and game abort are all generated by the machine; the first synthesis requires the creation of a cache, and the playback will be significantly accelerated thereafter.
What does this voice update fix?
The voice of the timing system must not only read out the fixed prompts, but also clearly announce the player's name, round and event status. Timer 5.2 introduces a new AI voice engine focused on making these high-frequency cues sound more natural while retaining the clarity needed on the field. The video successively plays system phrases such as "Racers up next" and "Race aborted", and uses names such as JR7, Kyosanim, Ramon, and Toothless to demonstrate the broadcast effect.
Voice computing is completely done locally and does not rely on continuous networking. For live events, this means that broadcasts can still be generated when the network is unstable, and it also avoids sending player names to external voice services. When the version was released, it was first targeted at the desktop. macOS and Linux support were still marked as coming soon at that time; before installation, you should refer to the system support listed on the current release page.
Select Sound in Workspace Settings
Speech is selectable in the settings menu of Workspace Settings. The video continuously demonstrates multiple switches starting from about 00:50. The system will first read out "Voice switch to..." and then use a new voice to announce the next group of players or the suspension of the game. After hearing the switch confirmation, and then playing an example containing the name, it is easier to judge whether the pronunciation, rhythm and recognition are suitable for a noisy venue than just listening to a fixed short sentence.
Different sounds will not change the timing or event arrangement logic, only the broadcast output. Before the official competition, it is recommended to listen to a trial round with the actual list of contestants to confirm that the nicknames, letters and numbers can be heard clearly by the staff and players, and then fix the sound settings for the event.
The first slowness belongs to the local cache process
The first time the AI voice generates a piece of content, you may need to wait because the program needs to complete the synthesis locally and write it to the cache. The cache can be reused when the same prompt or name appears again, so subsequent playback is usually faster. Don’t just assume that the latency on the first listen is the same latency that will occur throughout the game; it’s safer to pre-hear the roster and common tips before the game starts.
Official release notes state that computers with available GPUs will run approximately ten times faster. Computers without a GPU can still be used, but will require more preparation time when generating a large number of new names for the first time. After upgrading, complete a complete audition in a non-competition environment and retain the original manual callout or display screen as on-site redundancy.
Operation steps
- Enter voice settings
Open Workspace Settings and find the voice selections available with this version. The release notes only confirm this menu level. Please refer to the current version of the interface for specific control names.
- Switch and listen to the announcement containing names
Examples such as playing the next group of players after selecting a sound; the video starts from about 00:50 to show switching confirmation and name announcement, which is suitable for comparing the clarity of different sounds.
- Create cache before game
Generate contestant names and common event tips in advance. It may be slow the first time, but repeated broadcasting will be faster after the cache is completed.
FAQ
Does AI voice need to be connected to the Internet?
Not required. Timer 5.2's speech calculations are done natively; they are also cached after the first generation, and subsequent identical content is generally faster.
Why is the first playback slower?
For the first time, you need to synthesize it locally and create a cache. It is recommended to listen to the player names and common tips before the game, and do not wait until the official start to generate them for the first time.
Is a GPU required?
Not required, but officially a computer with an available GPU can be about ten times faster. The CPU runtime should allow more preparation time for the first build.
Does switching the sound in the video change the event settings?
No. It only changes the broadcast sound; players, rounds and game status are still determined by Timer's event data.
Full timeline transcript
Transcripts are arranged according to video time, making it easy to quickly locate the explanation content. Transcript language: English.
Welcome to FPVSIM timer. Racers up next.
Voice switch to heart. Voice switch to AIID. Voice switch to Nicole. Racers up next. JR7. Kyosanim R7. Race aborted. Ramon R7. Toothless R7. Voice switch to Daniel. Racers up next. JR7. Kyosanim R7. Race aborted.
Toothless R7



