I wish you delivered what you say you delivered

#3
by NezTheNaughty - opened

This is hot garbage.

I wish you delivered what you say you delivered
This is hot garbage.

Could you expand with few words on the hot garbage and missing delivery? What are your findings?

Institute of Foundation Models org

Hi, my friends,

I am from the IFM team. Could you let me know what specifics are you looking for?

Btw, our team do find that some of our artifacts are not finished uploading yet (e.g., we mess up on some ckpt uploading due to simple index confusion). Here is our rough follow-up plan:

  • code will be ready soon, some cleanup on branches will be needed.
  • checkpoints are a bit messy, they are there as git references but we need to provide a table explaining it better
  • data are halfway uploaded.

The full paper is probably gonna take longer for us to polish.

Excellent. Good to know! Thank you.
Literal garbage was showing up in the results, and I wasn't the only one. Benchmarks were not even coming close in most areas.
If you can straighten this out to what you have claimed, this will be an amazing piece of work.

this is some hard core benchmaxxing

It is currently very hard to get this model actually running locally. I gave up trying to figure out how to compile the llama.cpp fork etc. and just used another model which benchmarked 2x worse than K2-Horizon-7B while having more parameters. It would be nice if the setup process was elaborated on more because I would love to use this model.

Could you please make a commit with this model architecture into llama.cpp.
I am getting Failed to load model: The installed llama.cpp does not recognise this GGUF's model architecture ('k2-horizon'). The file is valid. If the model is newer than this llama.cpp build, updating llama.cpp may add support for it.

image

Institute of Foundation Models org

Got it my friends. We have llama.cpp ports but haven't merged them yet. Will follow up. @mvillmow-mbzuai @lmlmcat

Institute of Foundation Models org

@inikishev Can you explain more what your issue is with the llama.cpp fork? We are working with the llama.cpp team to get the code integrated here: https://github.com/ggml-org/llama.cpp/discussions/28308

I have also put up a PR against lemonade-server.ai here: https://github.com/lemonade-sdk/lemonade/pull/3530

It shows how to build the fork and install it with lemonade-server. The commands from the PR are:

git clone --branch model/K2Horizon --single-branch https://github.com/MBZUAI-IFM/llama.cpp.git llama.cpp-ifm
cmake -S llama.cpp-ifm -B llama.cpp-ifm/build -DGGML_METAL=ON -DCMAKE_BUILD_TYPE=Release
cmake --build llama.cpp-ifm/build --config Release --target llama-server -j

If you don't have a Mac, then you will need to set the correct GGML_ flags instead of GGML_METAL=ON from the documentation found here: https://github.com/ggml-org/llama.cpp/blob/master/docs/build.md#blas-build.

Sign up or log in to comment