NovaEar Standalone v43an RD NovaSphere application · reliable presenter video · smart transcription · richer music
RD NovaSphere ↗
Unified Library

One recording. Everything attached.

The original audio remains the permanent source. Transcript, speakers, notes, NovaAI insights, melody analysis, arrangements and exports belong to that same record.

Quick actions

Start where you are

NovaEar status

Foundation at a glance

Audio EngineChecking…
Speech EngineChecking…
Music EngineChecking…
Local StorageChecking…
Status here is lightweight. Tools → Diagnostics performs the deeper independent engine tests.

No recordings yet

Your first saved recording will appear here with its transcript, notes, music analysis and exports attached.
Audio Engine · NovaCapture

What are you recording?

Choose the recording purpose and capture source independently. NovaEar can record audio, webcam video, or screen video. The original media is saved first; Speech and Music only run when you request them.

Microphone mode records only your selected microphone.
1 · Capture2 · Save3 · Open4 · Transcribe5 · Analyse6 · Export
Ready.
Original media is saved locally before any transcript, caption, or music artifact becomes authoritative.
Capture

New personal note

00:00
Input: not activeSource: microphoneIdle
Audio Engine · First-class import

Upload audio or music

Bring existing recordings, songs, instrumentals, lectures, meetings or interviews into NovaEar. The original file is preserved first; Speech and Music remain independent.

Choose the content type, then select an audio/music file.
Imported files stay local in IndexedDB. NovaEar does not need a cloud upload to preserve or analyse them.
Recording Workspace

Recording

Original media preserved
Balanced uses a stronger everyday model and automatically rechecks uncertain sections.
Language availability is verified against the active ASR backend.
Speech

Transcript

Speaker map

Speakers

NovaAI

Insights attached to this recording

Music

Analysis & arrangements

Notes

Your notes for this recording

Portable record

Export without lock-in

The JSON export contains record metadata and attached artifacts. Original audio/video remains separately downloadable in its captured format.

Speech Engine

Transcript

NovaStudio

Build in layers. Keep the voice central.

Voice → pitch contour → notes → rhythm → key → arrangement. v34 adds NovaReference style learning and NovaStem preparation while preserving the three-engine architecture, performance instruments, accent-aware Speech and NovaAI context correction.

Music Engine

My Music Sources

No song or imported music sources yet. Record a song or upload audio/music.
NovaReference · Style intelligence

Learn from an uploaded instrumental or song

Choose a saved music source and let NovaEar measure its musical character. The resulting profile guides a new arrangement without copying the reference melody or audio.

Import a music/instrumental file, then analyse it here.
Copyright-safe design: NovaReference stores measurements and descriptors—not the reference audio pattern, melody or exact chord sequence—as generation guidance.
NovaStem · Source-separation preparation

Vocals ↔ instrumental pipeline

NovaEar now reserves stem artifacts on each music record so a local separation model can attach Vocals and Instrumental stems without creating a fourth core engine.

Preparation does not pretend to separate stems without a real source-separation model. It creates the safe pipeline target for the next local backend.
NovaInstrument Lab · v3

Create and perform your own instrument

Build a reusable local synth patch, then play it from your computer keyboard or a compatible MIDI controller. Nothing is uploaded.

Computer keyboard: A W S E D F T G Y H U J K play chromatic notes · Z/X change octave · Space = sustain · play several keys together for chords. Octave 0 · Sustain Off
MIDI not connected. Enable the computer keyboard first.
Press “Enable computer keyboard”, then use A/W/S/E/D/F… to play.
Create a sound, play the keyboard, then save it.
Foundation Tools

Inspect before expanding.

Diagnostics, architecture and sound tools are separated from the everyday recording flow.

Recommended first

Diagnostics

Run browser checks plus independent Audio, Speech and Music engine smoke-tests.

Engineering

Architecture

Review engine ownership, boundaries and the permanent record contract.

Music Engine

Instrument Vault

Audition generated instruments without coupling synthesis code to the UI.

Music Engine · Generated Vault

Playable code-generated sounds

Each card directly calls the Music Engine.

Researched Open Layer

Preserved research catalogue

    Speech Quality Architecture

    Three engines. Hard boundaries.

    The orchestrator may ask an engine to work. No engine is allowed to own another engine’s responsibility.

    1 · Audio Engine

    Capture, pause/resume/stop, imports, decode/resample, IndexedDB and canonical record schema.

    2 · Speech Engine

    Worker lifecycle, audio conditioning, adaptive VAD, quality-tier model routing, uncertain-section retry, repetition rejection and quality scoring.

    3 · Music Engine

    Smoothed pitch contour, note segmentation, key confidence, chord-candidate scoring, multiple instrumental palettes, custom synth patches and Web Audio rendering.

    Permanent Record Contract · schema v2
    record
    ├── original audio blob + metadata
    ├── transcript + timestamped segments
    ├── speakers[]
    ├── notes[]
    ├── novaInsights[]
    ├── musicAnalysis
    └── arrangements[]

    Existing v25 records are normalised into this schema when read; the database name is deliberately retained so upgrading does not discard the user’s local library.

    v43 Presenter Video Reliability Quality Gate

    Engine isolation diagnostics

    These tests do more than detect APIs: Audio performs an IndexedDB write/read/delete round-trip, Speech boots and responds through its worker, and Music verifies its analysis/arrangement contract. The ASR model itself still loads only when transcription is requested. v31 keeps Quick, Balanced and Best tiers, adds accent-aware routing plus NovaLexicon polishing, first-class uploaded music, computer/MIDI instrument performance, and safer voice/backing preview mixing while preserving hard engine boundaries.

    Audio EngineNot tested
    Speech EngineNot tested
    Music EngineNot tested
    No test has run yet.