Skip to content

Commit 795016d

Browse files
feat: tongyi plugin add glm4.7 model (#2362)
Co-authored-by: zhaojiangang-it <[email protected]>
1 parent afa872f commit 795016d

File tree

3 files changed

+97
-1
lines changed

3 files changed

+97
-1
lines changed

models/tongyi/manifest.yaml

Lines changed: 1 addition & 1 deletion
Original file line numberDiff line numberDiff line change
@@ -25,5 +25,5 @@ resource:
2525
model:
2626
enabled: false
2727
type: plugin
28-
version: 0.1.16
28+
version: 0.1.17
2929
created_at: "2024-12-10T16:13:50.29298939+08:00"

models/tongyi/models/llm/_position.yaml

Lines changed: 1 addition & 0 deletions
Original file line numberDiff line numberDiff line change
@@ -96,5 +96,6 @@
9696
- qwen-coder-turbo-0919
9797
- qwen-coder-turbo
9898
- qwen-flash-2025-07-28
99+
- glm-4.7
99100
- qwen-flash
100101
- farui-plus
Lines changed: 95 additions & 0 deletions
Original file line numberDiff line numberDiff line change
@@ -0,0 +1,95 @@
1+
# this model corresponds to qwen-flash, for more details
2+
# please refer to (https://help.aliyun.com/zh/model-studio/getting-started/models)
3+
model: glm-4.7
4+
label:
5+
en_US: glm-4.7
6+
model_type: llm
7+
features:
8+
- multi-tool-call
9+
- agent-thought
10+
- stream-tool-call
11+
model_properties:
12+
mode: chat
13+
context_size: 202752
14+
parameter_rules:
15+
- name: temperature
16+
use_template: temperature
17+
type: float
18+
default: 1.0
19+
min: 0.0
20+
max: 2.0
21+
help:
22+
zh_Hans: 用于控制随机性和多样性的程度。具体来说,temperature值控制了生成文本时对每个候选词的概率分布进行平滑的程度。较高的temperature值会降低概率分布的峰值,使得更多的低概率词被选择,生成结果更加多样化;而较低的temperature值则会增强概率分布的峰值,使得高概率词更容易被选择,生成结果更加确定。
23+
en_US: Used to control the degree of randomness and diversity. Specifically, the temperature value controls the degree to which the probability distribution of each candidate word is smoothed when generating text. A higher temperature value will reduce the peak value of the probability distribution, allowing more low-probability words to be selected, and the generated results will be more diverse; while a lower temperature value will enhance the peak value of the probability distribution, making it easier for high-probability words to be selected. , the generated results are more certain.
24+
- name: max_tokens
25+
use_template: max_tokens
26+
type: int
27+
default: 16384
28+
min: 1
29+
max: 16384
30+
help:
31+
zh_Hans: 用于指定模型在生成内容时token的最大数量,它定义了生成的上限,但不保证每次都会生成到这个数量。
32+
en_US: It is used to specify the maximum number of tokens when the model generates content. It defines the upper limit of generation, but does not guarantee that this number will be generated every time.
33+
- name: top_p
34+
use_template: top_p
35+
type: float
36+
default: 0.8
37+
min: 0.01
38+
max: 0.99
39+
help:
40+
zh_Hans: 生成过程中核采样方法概率阈值,例如,取值为0.8时,仅保留概率加起来大于等于0.8的最可能token的最小集合作为候选集。取值范围为(0,1.0),取值越大,生成的随机性越高;取值越低,生成的确定性越高。
41+
en_US: The probability threshold of the kernel sampling method during the generation process. For example, when the value is 0.8, only the smallest set of the most likely tokens with a sum of probabilities greater than or equal to 0.8 is retained as the candidate set. The value range is (0,1.0). The larger the value, the higher the randomness generated; the lower the value, the higher the certainty generated.
42+
- name: top_k
43+
type: int
44+
min: 0
45+
max: 99
46+
label:
47+
zh_Hans: 取样数量
48+
en_US: Top k
49+
help:
50+
zh_Hans: 生成时,采样候选集的大小。例如,取值为50时,仅将单次生成中得分最高的50个token组成随机采样的候选集。取值越大,生成的随机性越高;取值越小,生成的确定性越高。
51+
en_US: The size of the sample candidate set when generated. For example, when the value is 50, only the 50 highest-scoring tokens in a single generation form a randomly sampled candidate set. The larger the value, the higher the randomness generated; the smaller the value, the higher the certainty generated.
52+
- name: seed
53+
required: false
54+
type: int
55+
default: 1234
56+
label:
57+
zh_Hans: 随机种子
58+
en_US: Random seed
59+
help:
60+
zh_Hans: 生成时使用的随机数种子,用户控制模型生成内容的随机性。支持无符号64位整数,默认值为 1234。在使用seed时,模型将尽可能生成相同或相似的结果,但目前不保证每次生成的结果完全相同。
61+
en_US: The random number seed used when generating, the user controls the randomness of the content generated by the model. Supports unsigned 64-bit integers, default value is 1234. When using seed, the model will try its best to generate the same or similar results, but there is currently no guarantee that the results will be exactly the same every time.
62+
- name: enable_thinking
63+
required: false
64+
type: boolean
65+
default: false
66+
label:
67+
zh_Hans: 思考模式
68+
en_US: Thinking mode
69+
help:
70+
zh_Hans: 是否开启思考模式。
71+
en_US: Whether to enable thinking mode.
72+
- name: response_format
73+
label:
74+
zh_Hans: 回复格式
75+
en_US: Response Format
76+
type: string
77+
help:
78+
zh_Hans: 指定模型必须输出的格式
79+
en_US: specifying the format that the model must output
80+
required: false
81+
options:
82+
- text
83+
- json_object
84+
pricing:
85+
input: '0.003'
86+
output: '0.014'
87+
unit: '0.001'
88+
currency: RMB
89+
tiered_pricing:
90+
- input: '0.003'
91+
output: '0.014'
92+
max_tokens: 32768
93+
- input: '0.004'
94+
output: '0.016'
95+
max_tokens: 169984

0 commit comments

Comments
 (0)