* feat: support multimodal LLM inputs - add model capability config for vision, audio, and video inputs - wire multimodal flags through simple app, chat agent, and workflow LLM nodes - normalize uploaded files and file links into image/audio/video request parts - adapt chat, agent, and toolcall runtime context for multimodal messages - add capability tags, i18n entries, design docs, and related tests * perf: code * test: update agent file prompt assertion --------- Co-authored-by: archer <545436317@qq.com>
| Name |
Last commit
|
Last Update |
|---|---|---|
| .. | ||
| image | Loading commit data... | |
| s3 | Loading commit data... | |
| s3TTL | Loading commit data... | |
| api.ts | Loading commit data... | |
| constants.ts | Loading commit data... | |
| icon.ts | Loading commit data... | |
| tools.ts | Loading commit data... | |
| type.ts | Loading commit data... | |
| utils.ts | Loading commit data... |