Communications·production

IAM.Speech

Transcription, diarization, QA scoring and conversation reports.

IAM.Speech interface
Current public IAM.Speech surface. The screenshot was verified when the documentation was published.

Key capabilities

  • ASR and speaker diarization
  • Quality assurance with keyword and intent analysis
  • Reports, fact tables and result export

Role in the ecosystem

IAM.Speech processes audio and recorded communications, turning them into transcriptions, speaker segments, facts and quality control reports. Source could be a file upload, API, or a write from IAM.Comm/IAM.Voice.

Processing pipeline

ingest → normalize → VAD/ASR → diarization → QA/extraction → report

Each stage publishes state and maintains a connection to the source. Partial the result is clearly marked: the absence of diarization should not look like confirmed assignment of cues to specific people.

Possibilities

  • timestamps, punctuation and domain vocabulary;
  • speaker diarization and manual role correction;
  • keywords, intents, required phrases and QA score;
  • structured facts and report export;
  • batch API and reprocessing with a new configuration.

Quality and reproducibility

Metrics are divided by language, channel, noise and number of speakers. Model version, the processing profile and source checksum are stored next to the result. Reprocessing does not silently overwrite the previous version.

Safety

Retention of the source audio and text is set separately. No access to recording follows automatically from access to the aggregated report. Export and deletion leave an audit receipt.