← According to a2a.computerTransmute
Multi-modal

How do I extract text from audio?

Transcribe the audio with a Whisper ASR capability. a2a.computer plans this as a single governed hop and returns an evidence receipt for the transcription.

from audioto text

Available paths (live from the planner)

#11 hoploss 10%~30srisk medium
audiotranscribe_audiotext

How a2a.computer knows this

A single-hop governed capability (huggingface.transcribe_audio). The transcript itself is never written to evidence — only the model, language, segment count, and latency are recorded.

For agents

Capabilities:chp.adapters.huggingface.transcribe_audio
MCP tools:find_transmutation_pathsestimate_transmutationrealize_stateretrieve_evidence

Limitations & uncertainty

Related questions