Fix bofenghuang/vigogne-2-70b-chat on Windows #22

ochafik · 2025-01-18T17:41:15Z

Now that gated models are run on CI (#21), time to unskip this one & promote previously commented-out fix.

* Copy minja from google/minja@58f0ca6 * Add --jinja and --chat-template-file flags * Add missing <optional> include * Avoid print in get_hf_chat_template.py * No designated initializers yet * Try and work around msvc++ non-macro max resolution quirk * Update test_chat_completion.py * Wire LLM_KV_TOKENIZER_CHAT_TEMPLATE_N in llama_model_chat_template * Refactor test-chat-template * Test templates w/ minja * Fix deprecation * Add --jinja to llama-run * Update common_chat_format_example to use minja template wrapper * Test chat_template in e2e test * Update utils.py * Update test_chat_completion.py * Update run.cpp * Update arg.cpp * Refactor common_chat_* functions to accept minja template + use_jinja option * Attempt to fix linkage of LLAMA_CHATML_TEMPLATE * Revert LLAMA_CHATML_TEMPLATE refactor * Normalize newlines in test-chat-templates for windows tests * Forward decl minja::chat_template to avoid eager json dep * Flush stdout in chat template before potential crash * Fix copy elision warning * Rm unused optional include * Add missing optional include to server.cpp * Disable jinja test that has a cryptic windows failure * minja: fix vigogne (google/minja#22) * Apply suggestions from code review Co-authored-by: Xuan Son Nguyen <thichthat@gmail.com> Co-authored-by: Georgi Gerganov <ggerganov@gmail.com> * Finish suggested renamings * Move chat_templates inside server_context + remove mutex * Update --chat-template-file w/ recent change to --chat-template * Refactor chat template validation * Guard against missing eos/bos tokens (null token otherwise throws in llama_vocab::impl::token_get_attr) * Warn against missing eos / bos tokens when jinja template references them * rename: common_chat_template[s] * reinstate assert on chat_templates.template_default * Update minja to google/minja@b8437df * Update minja to google/minja#25 * Update minja from google/minja#27 * rm unused optional header --------- Co-authored-by: Xuan Son Nguyen <thichthat@gmail.com> Co-authored-by: Georgi Gerganov <ggerganov@gmail.com>

ochafik added 2 commits January 18, 2025 17:41

Unskip bofenghuang/vigogne-2-70b-chat from tests

e11b7fb

Fix strip on Windows

48db570

ochafik changed the title ~~Unskip bofenghuang/vigogne-2-70b-chat from tests~~ Fix bofenghuang/vigogne-2-70b-chat on Windows Jan 18, 2025

Merge branch 'main' into vigogne

6e1dd07

ochafik added a commit to ochafik/llama.cpp that referenced this pull request Jan 18, 2025

minja: fix vigogne (google/minja#22)

cc50356

ochafik merged commit 5d57c98 into main Jan 18, 2025
7 checks passed

ochafik deleted the vigogne branch January 18, 2025 17:59

ochafik mentioned this pull request Jan 18, 2025

Add Jinja template support ggerganov/llama.cpp#11016

Merged

3 tasks

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Fix bofenghuang/vigogne-2-70b-chat on Windows #22

Fix bofenghuang/vigogne-2-70b-chat on Windows #22

ochafik commented Jan 18, 2025 •

edited

Loading

Fix bofenghuang/vigogne-2-70b-chat on Windows #22

Fix bofenghuang/vigogne-2-70b-chat on Windows #22

Conversation

ochafik commented Jan 18, 2025 • edited Loading

ochafik commented Jan 18, 2025 •

edited

Loading