← Back to context

Comment by tharkun__

8 hours ago

Not impressed. I asked it how to run itself (giving it the Huggingface link) on limited RAM i.e. less than stated as needed and on llama.cpp and true to what we read about "it will tell you when it doesn't know" that's almost all I got: It doesn't know, it told me I should go click on tabs in the Huggingface interface for more information. This was with extended thinking on.

No, I'm not gonna do that, I asked you to do that Mr Kolibri.

Also feedback on that interface: It's very annoying while answering. It almost immediately shows a list of sources, which on my screen fill up all the space and then when it starts answering it keeps those in view but also scrolls down the tiny part of actual text its outputting but I can't scroll up to start reading from the top, coz it keeps scrolling. I have to wait until it's completely done generating its output.

It can do search, but it doesn't seem to have a tool that pulls URLs into context. I get why you would expect that tho, as most productized LLMs do it.

  • Correct, search is added as a tool.

    Uses Brave search, the idea is to test how well the Model can decided when to leverage search or not.

    • And that's my point when I said I wasn't impressed: It apparently can't do that properly. It can't seem to think things through by itself and then make more tool calls to fetch more information.

      What I couldn't tell but maybe you can tell us: it also complained that it couldn't read the full text as something was cut off. Is that because of the tool you gave it, of brave itself or is it the model?

Thank you for the feedback on UI, improved the streaming to make it less frustrating.