diff options
Diffstat (limited to 'README.md')
| -rw-r--r-- | README.md | 5 |
1 files changed, 3 insertions, 2 deletions
@@ -519,8 +519,9 @@ short default instruction so the request always carries a text part. Capability comes from the router: llama-server's model router reports each model's accepted inputs in `/v1/models` (`architecture.input_modalities`), and -a model listing `"audio"` there gets the control. Cloud models, whose -endpoints do not report modalities, can be marked by hand with the +a model listing `"audio"` there gets the control and a microphone (🎙) beside +its name in the picker, the way a vision model gets an eye. Cloud models, +whose endpoints do not report modalities, can be marked by hand with the **Accepts audio** box in the model settings dialog. Recording is 16 kHz mono WAV, the format speech encoders expect. The clip is |
