Seems I'm missing something. Does this model support other inputs?
Image outputs are supported, videos I'm not sure but I don't think that's an output, just a preview of the equirectangular example, so, same question here, what does this model outputs that isn't supported?
it's multimodal, see https://github.com/ggml-org/llama.cpp/blob/master/docs/multi...
Multimodal doesn't guarantee input and output.
> Currently, we support image, audio and video input.
Seems I'm missing something. Does this model support other inputs?
Image outputs are supported, videos I'm not sure but I don't think that's an output, just a preview of the equirectangular example, so, same question here, what does this model outputs that isn't supported?
7 replies →