Congratulations for the team and the contributors 🗡️
I was testing the limits of this project, for that I used Deepseek Harness along with ornith 1.5 35B A3B and this FreeToken.
Got a good cache 262144 and 15 to 19 tokens per seconds (RTX 5070ti 16GB, 96GB system ram)
The project https://github.com/rodjjo/pimbalgame
I validated using this #354
Not related to this project the only thing I saw was this funny thing (related to the model, not the project):
At the end the model just created a python script to concatenate "set" + "Scale" to fix the C++ files :D
Thanks, Hope we have support for vision and MTP for this model :)
Congratulations for the team and the contributors 🗡️
I was testing the limits of this project, for that I used Deepseek Harness along with ornith 1.5 35B A3B and this FreeToken.
Got a good cache 262144 and 15 to 19 tokens per seconds (RTX 5070ti 16GB, 96GB system ram)
The project https://github.com/rodjjo/pimbalgame
I validated using this #354
Not related to this project the only thing I saw was this funny thing (related to the model, not the project):
At the end the model just created a python script to concatenate "set" + "Scale" to fix the C++ files :D
Thanks, Hope we have support for vision and MTP for this model :)