Skip to content

Commit 10fa95a

Browse files
techpro-aimlapigitbook-bot
authored andcommitted
GITBOOK-401: docs: add hermes-4-405b, qwen3-next-80b-a3b (instruct and thinking), grok-code-fast-1
1 parent 019a74b commit 10fa95a

14 files changed

Lines changed: 641 additions & 12 deletions

File tree

‎docs/README.md‎

Lines changed: 2 additions & 0 deletions
Original file line numberDiff line numberDiff line change
@@ -115,6 +115,8 @@ If you've already made your choice and know the model ID, use the [Search panel]
115115

116116
**Mistral AI**: [Text/Chat](api-references/text-models-llm/Mistral-AI/) [Vision(OCR)](api-references/vision-models/ocr-optical-character-recognition/mistral-ai/)
117117

118+
**NousResearch**: [Text/Chat](api-references/text-models-llm/nousresearch/)
119+
118120
**NVIDIA**: [Text/Chat](api-references/text-models-llm/NVIDIA/)
119121

120122
<mark style="background-color:green;">**OpenAI**</mark>: [Text/Chat](api-references/text-models-llm/OpenAI/) [Image](api-references/image-models/OpenAI/) [Speech-To-Text](api-references/speech-voice-models/stt/OpenAI/) [Embedding](api-references/embedding-models/OpenAI/)

‎docs/SUMMARY.md‎

Lines changed: 5 additions & 0 deletions
Original file line numberDiff line numberDiff line change
@@ -23,6 +23,8 @@
2323
* [qwen3-32b](api-references/text-models-llm/alibaba-cloud/qwen3-32b.md)
2424
* [qwen3-coder-480b-a35b-instruct](api-references/text-models-llm/alibaba-cloud/qwen3-coder-480b-a35b-instruct.md)
2525
* [qwen3-235b-a22b-thinking-2507](api-references/text-models-llm/alibaba-cloud/qwen3-235b-a22b-thinking-2507.md)
26+
* [qwen3-next-80b-a3b-instruct](api-references/text-models-llm/alibaba-cloud/qwen3-next-80b-a3b-instruct.md)
27+
* [qwen3-next-80b-a3b-thinking](api-references/text-models-llm/alibaba-cloud/qwen3-next-80b-a3b-thinking.md)
2628
* [Anthracite](api-references/text-models-llm/Anthracite/README.md)
2729
* [magnum-v4](api-references/text-models-llm/Anthracite/magnum-v4.md)
2830
* [Anthropic](api-references/text-models-llm/Anthropic/README.md)
@@ -73,6 +75,8 @@
7375
* [Mixtral-8x7B-Instruct](api-references/text-models-llm/Mistral-AI/Mixtral-8x7B-Instruct-v0.1.md)
7476
* [Moonshot](api-references/text-models-llm/moonshot/README.md)
7577
* [kimi-k2-preview](api-references/text-models-llm/moonshot/kimi-k2-preview.md)
78+
* [NousResearch](api-references/text-models-llm/nousresearch/README.md)
79+
* [hermes-4-405b](api-references/text-models-llm/nousresearch/hermes-4-405b.md)
7680
* [NVIDIA](api-references/text-models-llm/NVIDIA/README.md)
7781
* [llama-3.1-nemotron-70b](api-references/text-models-llm/NVIDIA/llama-3.1-nemotron-70b.md)
7882
* [OpenAI](api-references/text-models-llm/OpenAI/README.md)
@@ -108,6 +112,7 @@
108112
* [grok-3-beta](api-references/text-models-llm/xai/grok-3-beta.md)
109113
* [grok-3-mini-beta](api-references/text-models-llm/xai/grok-3-mini-beta.md)
110114
* [grok-4](api-references/text-models-llm/xai/grok-4.md)
115+
* [grok-code-fast-1](api-references/text-models-llm/xai/grok-code-fast-1.md)
111116
* [Zhipu](api-references/text-models-llm/zhipu/README.md)
112117
* [glm-4.5-air](api-references/text-models-llm/zhipu/glm-4.5-air.md)
113118
* [glm-4.5](api-references/text-models-llm/zhipu/glm-4.5.md)

‎docs/api-references/model-database.md‎

Lines changed: 1 addition & 1 deletion
Large diffs are not rendered by default.

‎docs/api-references/text-models-llm/Alibaba-Cloud/qwen-turbo.md‎

Lines changed: 6 additions & 6 deletions
Original file line numberDiff line numberDiff line change
@@ -41,6 +41,12 @@ If you need a more detailed walkthrough for setting up your development environm
4141

4242
</details>
4343

44+
## API Schema
45+
46+
{% openapi-operation spec="qwen-turbo" path="/v1/chat/completions" method="post" %}
47+
[OpenAPI qwen-turbo](https://raw.githubusercontent.com/aimlapi/api-docs/refs/heads/main/docs/api-references/text-models-llm/Alibaba-Cloud/qwen-turbo.json)
48+
{% endopenapi-operation %}
49+
4450
## Code Example
4551

4652
{% tabs %}
@@ -117,9 +123,3 @@ main();
117123
{% endcode %}
118124

119125
</details>
120-
121-
## API Schema
122-
123-
{% openapi-operation spec="qwen-turbo" path="/v1/chat/completions" method="post" %}
124-
[OpenAPI qwen-turbo](https://raw.githubusercontent.com/aimlapi/api-docs/refs/heads/main/docs/api-references/text-models-llm/Alibaba-Cloud/qwen-turbo.json)
125-
{% endopenapi-operation %}

‎docs/api-references/text-models-llm/README.md‎

Lines changed: 1 addition & 1 deletion
Large diffs are not rendered by default.
Lines changed: 147 additions & 0 deletions
Original file line numberDiff line numberDiff line change
@@ -0,0 +1,147 @@
1+
# qwen3-next-80b-a3b-instruct
2+
3+
<table data-header-hidden data-full-width="true"><thead><tr><th width="546.4443969726562" valign="top"></th><th width="202.666748046875" valign="top"></th></tr></thead><tbody><tr><td valign="top"><div data-gb-custom-block data-tag="hint" data-style="info" class="hint hint-info"><p>This documentation is valid for the following model: </p><p><code>alibaba/qwen3-next-80b-a3b-instruct</code></p></div></td><td valign="top"><a href="https://aimlapi.com/app/?model=alibaba/qwen3-next-80b-a3b-instruct&#x26;mode=chat" class="button primary">Try in Playground</a></td></tr></tbody></table>
4+
5+
## Model Overview
6+
7+
An instruction-tuned chat model optimized for fast, stable replies without reasoning traces, designed for complex tasks in reasoning, coding, knowledge QA, and multilingual use, with strong alignment and formatting.
8+
9+
## How to Make a Call
10+
11+
<details>
12+
13+
<summary>Step-by-Step Instructions</summary>
14+
15+
### :digit\_one: Setup You Can’t Skip
16+
17+
:black\_small\_square: [**Create an Account**](https://aimlapi.com/app/sign-up): Visit the AI/ML API website and create an account (if you don’t have one yet).\
18+
:black\_small\_square: [**Generate an API Key**](https://aimlapi.com/app/keys): After logging in, navigate to your account dashboard and generate your API key. Ensure that key is enabled on UI.
19+
20+
### &#x20;:digit\_two: Copy the code example
21+
22+
At the bottom of this page, you'll find [a code example](qwen3-next-80b-a3b-instruct.md#code-example) that shows how to structure the request. Choose the code snippet in your preferred programming language and copy it into your development environment.
23+
24+
### :digit\_three: Modify the code example
25+
26+
:black\_small\_square: Replace `<YOUR_AIMLAPI_KEY>` with your actual AI/ML API key from your account.\
27+
:black\_small\_square: Insert your question or request into the `content` field—this is what the model will respond to.
28+
29+
### :digit\_four: <sup><sub><mark style="background-color:yellow;">(Optional)<mark style="background-color:yellow;"><sub></sup> Adjust other optional parameters if needed
30+
31+
Only `model` and `messages` are required parameters for this model (and we’ve already filled them in for you in the example), but you can include optional parameters if needed to adjust the model’s behavior. Below, you can find the corresponding [API schema](qwen3-next-80b-a3b-instruct.md#api-schema), which lists all available parameters along with notes on how to use them.
32+
33+
### :digit\_five: Run your modified code
34+
35+
Run your modified code in your development environment. Response time depends on various factors, but for simple prompts it rarely exceeds a few seconds.
36+
37+
{% hint style="success" %}
38+
If you need a more detailed walkthrough for setting up your development environment and making a request step by step — feel free to use our [Quickstart guide](../../../quickstart/setting-up.md).
39+
{% endhint %}
40+
41+
</details>
42+
43+
## API Schema
44+
45+
{% openapi-operation spec="qwen3-next-80b-a3b-instruct" path="/v1/chat/completions" method="post" %}
46+
[OpenAPI qwen3-next-80b-a3b-instruct](https://raw.githubusercontent.com/aimlapi/api-docs/refs/heads/main/docs/api-references/text-models-llm/Alibaba-Cloud/qwen3-next-80b-a3b-instruct.json)
47+
{% endopenapi-operation %}
48+
49+
## Code Example
50+
51+
{% tabs %}
52+
{% tab title="Python" %}
53+
{% code overflow="wrap" %}
54+
```python
55+
import requests
56+
import json # for getting a structured output with indentation
57+
58+
response = requests.post(
59+
"https://api.aimlapi.com/v1/chat/completions",
60+
headers={
61+
# Insert your AIML API Key instead of <YOUR_AIMLAPI_KEY>:
62+
"Authorization":"Bearer <YOUR_AIMLAPI_KEY>",
63+
"Content-Type":"application/json"
64+
},
65+
json={
66+
"model":"alibaba/qwen3-next-80b-a3b-instruct",
67+
"messages":[
68+
{
69+
"role":"user",
70+
"content":"Hello" # insert your prompt here, instead of Hello
71+
}
72+
],
73+
"enable_thinking": False
74+
}
75+
)
76+
77+
data = response.json()
78+
print(json.dumps(data, indent=2, ensure_ascii=False))
79+
```
80+
{% endcode %}
81+
{% endtab %}
82+
83+
{% tab title="JavaScript" %}
84+
{% code overflow="wrap" %}
85+
```javascript
86+
async function main() {
87+
const response = await fetch('https://api.aimlapi.com/v1/chat/completions', {
88+
method: 'POST',
89+
headers: {
90+
// insert your AIML API Key instead of <YOUR_AIMLAPI_KEY>
91+
'Authorization': 'Bearer <YOUR_AIMLAPI_KEY>',
92+
'Content-Type': 'application/json',
93+
},
94+
body: JSON.stringify({
95+
model: 'alibaba/qwen3-next-80b-a3b-instruct',
96+
messages:[
97+
{
98+
role:'user',
99+
content: 'Hello' // insert your prompt here, instead of Hello
100+
}
101+
],
102+
}),
103+
});
104+
105+
const data = await response.json();
106+
console.log(JSON.stringify(data, null, 2));
107+
}
108+
109+
main();
110+
```
111+
{% endcode %}
112+
{% endtab %}
113+
{% endtabs %}
114+
115+
<details>
116+
117+
<summary>Response</summary>
118+
119+
{% code overflow="wrap" %}
120+
```json5
121+
{
122+
"id": "chatcmpl-a944254a-4252-9a54-af1b-94afcfb9807e",
123+
"system_fingerprint": null,
124+
"object": "chat.completion",
125+
"choices": [
126+
{
127+
"index": 0,
128+
"finish_reason": "stop",
129+
"logprobs": null,
130+
"message": {
131+
"role": "assistant",
132+
"content": "Hello! How can I help you today? 😊"
133+
}
134+
}
135+
],
136+
"created": 1758228572,
137+
"model": "qwen3-next-80b-a3b-instruct",
138+
"usage": {
139+
"prompt_tokens": 9,
140+
"completion_tokens": 46,
141+
"total_tokens": 55
142+
}
143+
}
144+
```
145+
{% endcode %}
146+
147+
</details>
Lines changed: 151 additions & 0 deletions
Original file line numberDiff line numberDiff line change
@@ -0,0 +1,151 @@
1+
# qwen3-next-80b-a3b-thinking
2+
3+
<table data-header-hidden data-full-width="false"><thead><tr><th width="546.4443969726562" valign="top"></th><th width="202.666748046875" valign="top"></th></tr></thead><tbody><tr><td valign="top"><div data-gb-custom-block data-tag="hint" data-style="info" class="hint hint-info"><p>This documentation is valid for the following model: </p><p><code>alibaba/qwen3-next-80b-a3b-thinking</code></p></div></td><td valign="top"><a href="https://aimlapi.com/app/?model=alibaba/qwen3-next-80b-a3b-thinking&#x26;mode=chat" class="button primary">Try in Playground</a></td></tr></tbody></table>
4+
5+
## Model Overview
6+
7+
The model may take longer to generate reasoning content than its predecessor. Alibaba Cloud strongly recommends its use for highly complex reasoning tasks.
8+
9+
## How to Make a Call
10+
11+
<details>
12+
13+
<summary>Step-by-Step Instructions</summary>
14+
15+
### :digit\_one: Setup You Can’t Skip
16+
17+
:black\_small\_square: [**Create an Account**](https://aimlapi.com/app/sign-up): Visit the AI/ML API website and create an account (if you don’t have one yet).\
18+
:black\_small\_square: [**Generate an API Key**](https://aimlapi.com/app/keys): After logging in, navigate to your account dashboard and generate your API key. Ensure that key is enabled on UI.
19+
20+
### &#x20;:digit\_two: Copy the code example
21+
22+
At the bottom of this page, you'll find [a code example](qwen3-next-80b-a3b-thinking.md#code-example) that shows how to structure the request. Choose the code snippet in your preferred programming language and copy it into your development environment.
23+
24+
### :digit\_three: Modify the code example
25+
26+
:black\_small\_square: Replace `<YOUR_AIMLAPI_KEY>` with your actual AI/ML API key from your account.\
27+
:black\_small\_square: Insert your question or request into the `content` field—this is what the model will respond to.
28+
29+
### :digit\_four: <sup><sub><mark style="background-color:yellow;">(Optional)<mark style="background-color:yellow;"><sub></sup> Adjust other optional parameters if needed
30+
31+
Only `model` and `messages` are required parameters for this model (and we’ve already filled them in for you in the example), but you can include optional parameters if needed to adjust the model’s behavior. Below, you can find the corresponding [API schema](qwen3-next-80b-a3b-thinking.md#api-schema), which lists all available parameters along with notes on how to use them.
32+
33+
### :digit\_five: Run your modified code
34+
35+
Run your modified code in your development environment. Response time depends on various factors, but for simple prompts it rarely exceeds a few seconds.
36+
37+
{% hint style="success" %}
38+
If you need a more detailed walkthrough for setting up your development environment and making a request step by step — feel free to use our [Quickstart guide](../../../quickstart/setting-up.md).
39+
{% endhint %}
40+
41+
</details>
42+
43+
## API Schema
44+
45+
{% openapi-operation spec="qwen3-next-80b-a3b-thinking" path="/v1/chat/completions" method="post" %}
46+
[OpenAPI qwen3-next-80b-a3b-thinking](https://raw.githubusercontent.com/aimlapi/api-docs/refs/heads/main/docs/api-references/text-models-llm/Alibaba-Cloud/qwen3-next-80b-a3b-thinking.json)
47+
{% endopenapi-operation %}
48+
49+
## Code Example
50+
51+
{% tabs %}
52+
{% tab title="Python" %}
53+
{% code overflow="wrap" %}
54+
```python
55+
import requests
56+
import json # for getting a structured output with indentation
57+
58+
response = requests.post(
59+
"https://api.aimlapi.com/v1/chat/completions",
60+
headers={
61+
# Insert your AIML API Key instead of <YOUR_AIMLAPI_KEY>:
62+
"Authorization":"Bearer <YOUR_AIMLAPI_KEY>",
63+
"Content-Type":"application/json"
64+
},
65+
json={
66+
"model":"alibaba/qwen3-next-80b-a3b-thinking",
67+
"messages":[
68+
{
69+
"role":"user",
70+
"content":"Hello" # insert your prompt here, instead of Hello
71+
}
72+
],
73+
"enable_thinking": False
74+
}
75+
)
76+
77+
data = response.json()
78+
print(json.dumps(data, indent=2, ensure_ascii=False))
79+
```
80+
{% endcode %}
81+
{% endtab %}
82+
83+
{% tab title="JavaScript" %}
84+
{% code overflow="wrap" %}
85+
```javascript
86+
async function main() {
87+
const response = await fetch('https://api.aimlapi.com/v1/chat/completions', {
88+
method: 'POST',
89+
headers: {
90+
// insert your AIML API Key instead of <YOUR_AIMLAPI_KEY>
91+
'Authorization': 'Bearer <YOUR_AIMLAPI_KEY>',
92+
'Content-Type': 'application/json',
93+
},
94+
body: JSON.stringify({
95+
model: 'alibaba/qwen3-next-80b-a3b-thinking',
96+
messages:[
97+
{
98+
role:'user',
99+
content: 'Hello' // insert your prompt here, instead of Hello
100+
}
101+
],
102+
}),
103+
});
104+
105+
const data = await response.json();
106+
console.log(JSON.stringify(data, null, 2));
107+
}
108+
109+
main();
110+
```
111+
{% endcode %}
112+
{% endtab %}
113+
{% endtabs %}
114+
115+
<details>
116+
117+
<summary>Response</summary>
118+
119+
{% code overflow="wrap" %}
120+
```json5
121+
{
122+
"id": "chatcmpl-576aaaf9-f712-9114-b098-c1ee83fbfb6b",
123+
"system_fingerprint": null,
124+
"object": "chat.completion",
125+
"choices": [
126+
{
127+
"index": 0,
128+
"finish_reason": "stop",
129+
"logprobs": null,
130+
"message": {
131+
"role": "assistant",
132+
"content": "Hello! 😊 How can I assist you today?",
133+
"reasoning_content": "Okay, the user said \"Hello\". I need to respond appropriately. Let me think.\n\nFirst, I should acknowledge their greeting. A simple \"Hello!\" would be good. Maybe add a friendly emoji to keep it warm.\n\nWait, but maybe they want to start a conversation. I should ask how I can help them. That way, I'm being helpful and opening the door for them to ask questions.\n\nLet me check the standard response. Typically, for \"Hello\", the assistant says something like \"Hello! How can I assist you today?\" or \"Hi there! What can I do for you?\"\n\nYes, that's right. Keep it friendly and open-ended. Maybe add a smiley emoji to make it approachable.\n\nSo the response should be: \"Hello! How can I assist you today?\"\n\nThat's good. Let me make sure there's no mistake. Yes, that's standard. No need for anything complicated here. Just a simple, welcoming reply.\n\nAlternatively, sometimes people use \"Hi\" instead of \"Hello\", but since they said \"Hello\", responding with \"Hello\" is fine. Maybe \"Hi there!\" could also work, but sticking to \"Hello\" matches their greeting.\n\nYes, \"Hello! How can I assist you today?\" is perfect. It's polite, friendly, and offers assistance. That should be the response."
134+
}
135+
}
136+
],
137+
"created": 1758229078,
138+
"model": "qwen3-next-80b-a3b-thinking",
139+
"usage": {
140+
"prompt_tokens": 9,
141+
"completion_tokens": 7182,
142+
"total_tokens": 7191,
143+
"completion_tokens_details": {
144+
"reasoning_tokens": 277
145+
}
146+
}
147+
}
148+
```
149+
{% endcode %}
150+
151+
</details>

‎docs/api-references/text-models-llm/moonshot/kimi-k2-preview.md‎

Lines changed: 1 addition & 1 deletion
Original file line numberDiff line numberDiff line change
@@ -1,6 +1,6 @@
11
# kimi-k2-preview
22

3-
<table data-header-hidden data-full-width="true"><thead><tr><th width="546.4443969726562" valign="top"></th><th width="202.666748046875" valign="top"></th></tr></thead><tbody><tr><td valign="top"><div data-gb-custom-block data-tag="hint" data-style="info" class="hint hint-info"><p>This documentation is valid for the following model: <br> <code>moonshot/kimi-k2-preview</code></p></div></td><td valign="top"><a href="https://aimlapi.com/app/?model=moonshot/kimi-k2-preview&#x26;mode=chat" class="button primary">Try in Playground</a></td></tr></tbody></table>
3+
<table data-header-hidden data-full-width="true"><thead><tr><th width="546.4443969726562" valign="top"></th><th width="202.666748046875" valign="top"></th></tr></thead><tbody><tr><td valign="top"><div data-gb-custom-block data-tag="hint" data-style="info" class="hint hint-info"><p>This documentation is valid for the following model: <br> <code>moonshot/kimi-k2-preview</code></p></div></td><td valign="top"><a href="https://aimlapi.com/app/?model=moonshot/kimi-k2-preview&#x26;mode=chat" class="button primary">Try in Playground</a></td></tr><tr><td valign="top"></td><td valign="top"></td></tr></tbody></table>
44

55
## Model Overview
66

Lines changed: 2 additions & 0 deletions
Original file line numberDiff line numberDiff line change
@@ -0,0 +1,2 @@
1+
# NousResearch
2+

0 commit comments

Comments
 (0)