A local neural voice studio for creating character voices, natural speech, and procedural ambient sound effects. Fine‑tune performances and design immersive soundscapes for games, films, and AI projects.
Release Date:
Coming soon
Developer:
Publisher:
Tags

Sign in to add this item to your wishlist, follow it, or mark it as ignored

This game is not yet available on Steam

Coming soon

Interested?
Add to your wishlist and get notified when it becomes available.
 

About This Software

Status – Awaiting Steam’s final review for release scheduling. Build is complete and ready for launch.

Mars Ind – Voice is a next‑generation synthetic‑voice creation suite built for developers, creators, and storytellers. It uses generative artificial intelligence to transform text into expressive, natural‑sounding speech with precise control over tone, pacing, and delivery.

Choose from a wide range of voice profiles or design your own using adjustable parameters for vocal character, age, and style. Add expressive elements such as laughter, emphasis, breath, and emotional cues. The system can interpret math, symbols, and structured text, and includes contextual tools for quickly inserting voice changes, effects, and modifiers.

A built‑in FX engine provides presets for shaping and refining audio, while the timeline editor allows basic sequencing and adjustment of generated lines. Advanced users can enable Server Mode to expose a local endpoint for integration with external tools, pipelines, or game engines.

Mars Ind – Voice is designed for flexibility, creativity, and production‑ready output, giving developers a powerful way to bring characters, dialogue, and interactive systems to life.

Interface Overview The suite includes dedicated tabs for every stage of voice creation and sound design:

  • Studio – Type, speak, and direct a whole cast in one script. Highlight any passage, right-click, and assign it a voice; each line is spoken by its own Replicant. Add laughter, hesitation, or spoken physics equations, then render, replay, and export layered conversations and broadcast-ready audio.

  • Voice Designer – Build a voice from nothing. Powered by Piper, start from an age preset; young child through elderly, then sculpt pitch, pace, warmth, and expressiveness with live sliders across a broad tonal range. Switch on cloning to bridge into XTTS-v2 and turn a sample of a real human voice into a reusable clone.

  • Sound FX – A full ambience engine. Generate rain, storms, and atmospheric effects with fine-grained intensity, density, and brightness controls. Seamless looping crossfades the clip end-to-start so it can play forever with no click, ideal for game ambience. 

  • Audio Timeline - Mix it all together. The Audio tab layers your rendered voices with sound effects on a visual timeline; drag clips to place them, trim the edges, snap them together, and combine everything into one file. Overlapping clips blend, gaps become silence, and the finished mix exports in a click.

  • Sound Tuning – Dial in the human details. Tune how non-word sounds; laughter, sighs, vocal fillers, and spoken math, are paced and pitched. Set a global default or save bespoke tuning per voice for believable, lifelike delivery everywhere.

  • Relay Server – Enables local voice streaming and integration. When activated, the app hosts a lightweight HTTP relay on your machine, allowing connected tools to stream the audio.

  • Log – displays system activity and diagnostic output.

Together, these modules form a complete local neural‑voice environment for creators and developers, with optional network connectivity for advanced workflows and multi‑tool integration.

AI Generated Content Disclosure

The developers describe how their game uses AI Generated Content like this:

This software uses generative artificial intelligence to create custom voice output based on user‑provided text or prompts. The AI system analyzes the input and produces natural‑sounding speech in real time.

Mature Content Description

The developers describe the content like this:

This software may produce audio with strong language or adult themes when the user chooses to enter that type of text. No mature content is included by default, and non‑adult voices have built‑in safety filters that prevent explicit or inappropriate output.

System Requirements

    Minimum:
    • Requires a 64-bit processor and operating system
    • OS: Windows 10 version 22H2 (64-bit)
    • Processor: Intel Core i5-8400 or AMD Ryzen 5 2600
    • Memory: 16 GB RAM
    • Graphics: NVIDIA GeForce RTX 2060 with 6 GB VRAM
    • DirectX: Version 12
    • Storage: 15 GB available space
    • Sound Card: Windows-compatible audio device
    • Additional Notes: Requires an NVIDIA Turing-generation or newer GPU with current NVIDIA drivers. AMD and Intel GPUs are not supported. A microphone is required only for recording new voice-cloning samples.
    Recommended:
    • Requires a 64-bit processor and operating system
    • OS: Windows 11 (64-bit)
    • Processor: Intel Core i7-10700 or AMD Ryzen 7 3700X
    • Memory: 32 GB RAM
    • Graphics: NVIDIA GeForce RTX 3060 with 12 GB VRAM or better
    • DirectX: Version 12
    • Storage: 20 GB available space
    • Sound Card: Windows-compatible audio device
    • Additional Notes: Requires an NVIDIA Turing-generation or newer GPU with current NVIDIA drivers. AMD and Intel GPUs are not supported. An SSD and current NVIDIA drivers are recommended. A microphone is required only for recording new voice-cloning samples.
There are no reviews for this product

You can write your own review for this product to share your experience with the community. Use the area above the purchase buttons on this page to write your review.

Review Filters