Skip to content

Conversation

@heyyymonth
Copy link

Added two new configuration presets to simplify command-line usage: 1. --chat-llama3-8b-default for running a chat server with Llama3 8B model, 2. --rerank-bge-default for running a reranking server with the BGE model. These presets configure appropriate model paths, server ports, GPU settings, and other parameters. Refs: #10932

Added two new configuration presets to simplify command-line usage: 1. --chat-llama3-8b-default for running a chat server with Llama3 8B model, 2. --rerank-bge-default for running a reranking server with the BGE model. These presets configure appropriate model paths, server ports, GPU settings, and other parameters. Refs: ggml-org#10932
@heyyymonth heyyymonth changed the title common: Add configuration presets for chat and reranking servers ggml: Add configuration presets for chat and reranking servers May 12, 2025
@heyyymonth heyyymonth changed the title ggml: Add configuration presets for chat and reranking servers llama: Add configuration presets for chat and reranking servers May 12, 2025
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant