Audio to text
A candidate speech model would separate speech from silence, transcribe each turn, and preserve speaker, time, and words. Vocero's own model is not trained yet.
Vocero proposes a workflow from audio to transcript and from transcript to minutes a person can verify. Every decision, task, or answer should retain a link to what was said.
In development for internal use. Our own model is not trained yet.
The architecture separates speech recognition from meeting analysis so each stage can be evaluated on its own.
A candidate speech model would separate speech from silence, transcribe each turn, and preserve speaker, time, and words. Vocero's own model is not trained yet.
A candidate language model would extract entities, decisions, tasks, owners, and open issues from the transcript. This stage is also part of the plan.
The planned output organizes findings into readable minutes and retains references to transcript lines for review.
The minutes assist the person responsible for the meeting; they do not make an automatic decision.
Every point should identify the line or lines that support it. If the meeting does not contain an answer, the system should say so instead of inventing one.
A person confirms names, numbers, decisions, owners, and dates before saving or sending the minutes.
Names, numbers, or decisions absent from the transcript count as inventions and must be recorded in evaluation.
The interactive presentation uses fictional meetings with known correct output to explain the format.
A synthetic customer-support script shows a duplicate charge, a priority case, a credit note, and confirmation by email.
Scripted answers cite the relevant lines and identify missing information, such as the case number.
Timings, waveforms, entities, and answers are programmed. The example is not a test of speech recognition or Vocero's own model.
The application is used internally with a dictation engine that is not Vocero's own model. The program still needs to create data, train candidates, and measure results.
The site and its synthetic demos run, and CareCompile internally uses a meeting and dictation application.
There is no trained Vocero model, no accuracy figure, and no completed sealed evaluation.
Vocero investigates dictation for Salvadoran and Central American Spanish. Try the implemented command rules and explore the regional speech plan.
Explore the scopeCareCompile seeks an institutional counterpart in El Salvador to explore consented speech and text data, limited purpose, and independent evaluation.
Explore the scopeReview Vocero's proposed architecture, training plan, and Vocero-Bench design for measuring speech, dictation, and minutes by country. There are no results yet.
Explore the scopeWe are looking for an institution willing to discuss regional meetings, voices, and texts. We explain purpose, consent, and governance before recording or receiving data.