• leanleft@lemmy.ml
    link
    fedilink
    English
    arrow-up
    6
    ·
    edit-2
    3 days ago

    im not bothered by it.
    but also… i wouldn’t use it as my sole model to perform every task. [waves hand toward huggingface.] ( a gazillion bajillion models, specialized for all types of tasks).

    edit: i checked. and there aren’t as many models as i expected, for this task(didnt test them). and… small multimodal model failed badly(see picture of the iguana … or whatever the hell that is)
    it’s a strange benchmark. but it seems reasonable to expect LLMs to be able to do this.

      • leanleft@lemmy.ml
        link
        fedilink
        English
        arrow-up
        1
        ·
        2 days ago

        i dont want to share it. but it was easy to add a little more description for basic criteria in the request. although, i did not detail a step-by-step process for drawing.
        i had this idea to feed it an actual image, to base the drawing on. but i wasnt able to. llama.cpp seems to be messed up[esp vulkan pkg](and also i probably dont know the proper way to utilize features(if they even exist for a model))

        • Domi@lemmy.secnd.me
          link
          fedilink
          arrow-up
          5
          ·
          2 days ago

          Here’s the result by Qwen3.8-27b Q8 for “Create an SVG of an Iguana.”:

          For your image with “Create an SVG of an Iguana based on the provided image.”:

          (Using llama.cpp with Opencode)