Instructions to use jdopensource/JoyAI-LLM-Flash-GGUF with libraries, inference providers, notebooks, and local apps. Follow these links to get started.

Libraries

How to use jdopensource/JoyAI-LLM-Flash-GGUF with Transformers:

# Use a pipeline as a high-level helper
from transformers import pipeline

pipe = pipeline("text-generation", model="jdopensource/JoyAI-LLM-Flash-GGUF")
messages = [
    {"role": "user", "content": "Who are you?"},
]
pipe(messages)

# Load model directly
from transformers import AutoModel
model = AutoModel.from_pretrained("jdopensource/JoyAI-LLM-Flash-GGUF", dtype="auto")

llama-cpp-python

How to use jdopensource/JoyAI-LLM-Flash-GGUF with llama-cpp-python:

# !pip install llama-cpp-python

from llama_cpp import Llama

llm = Llama.from_pretrained(
	repo_id="jdopensource/JoyAI-LLM-Flash-GGUF",
	filename="JoyAI-LLM-Flash-IQ3_XS.gguf",
)

llm.create_chat_completion(
	messages = [
		{
			"role": "user",
			"content": "What is the capital of France?"
		}
	]
)

Notebooks
Google Colab
Kaggle
Local Apps Settings

llama.cpp

How to use jdopensource/JoyAI-LLM-Flash-GGUF with llama.cpp:

Install (macOS, Linux)

curl -LsSf https://llama.app/install.sh | sh
# Start a local OpenAI-compatible server with a web UI:
llama serve -hf jdopensource/JoyAI-LLM-Flash-GGUF:IQ3_XS
# Run inference directly in the terminal:
llama cli -hf jdopensource/JoyAI-LLM-Flash-GGUF:IQ3_XS

Install from WinGet (Windows)

winget install llama.cpp
# Start a local OpenAI-compatible server with a web UI:
llama serve -hf jdopensource/JoyAI-LLM-Flash-GGUF:IQ3_XS
# Run inference directly in the terminal:
llama cli -hf jdopensource/JoyAI-LLM-Flash-GGUF:IQ3_XS

Use pre-built binary

# Download pre-built binary from:
# https://github.com/ggerganov/llama.cpp/releases
# Start a local OpenAI-compatible server with a web UI:
./llama-server -hf jdopensource/JoyAI-LLM-Flash-GGUF:IQ3_XS
# Run inference directly in the terminal:
./llama-cli -hf jdopensource/JoyAI-LLM-Flash-GGUF:IQ3_XS

Build from source code

git clone https://github.com/ggerganov/llama.cpp.git
cd llama.cpp
cmake -B build
cmake --build build -j --target llama-server llama-cli
# Start a local OpenAI-compatible server with a web UI:
./build/bin/llama-server -hf jdopensource/JoyAI-LLM-Flash-GGUF:IQ3_XS
# Run inference directly in the terminal:
./build/bin/llama-cli -hf jdopensource/JoyAI-LLM-Flash-GGUF:IQ3_XS

Use Docker

docker model run hf.co/jdopensource/JoyAI-LLM-Flash-GGUF:IQ3_XS

LM Studio
Jan

vLLM

How to use jdopensource/JoyAI-LLM-Flash-GGUF with vLLM:

Install from pip and serve model

# Install vLLM from pip:
pip install vllm
# Start the vLLM server:
vllm serve "jdopensource/JoyAI-LLM-Flash-GGUF"
# Call the server using curl (OpenAI-compatible API):
curl -X POST "http://localhost:8000/v1/chat/completions" \
	-H "Content-Type: application/json" \
	--data '{
		"model": "jdopensource/JoyAI-LLM-Flash-GGUF",
		"messages": [
			{
				"role": "user",
				"content": "What is the capital of France?"
			}
		]
	}'

Use Docker

docker model run hf.co/jdopensource/JoyAI-LLM-Flash-GGUF:IQ3_XS

SGLang

How to use jdopensource/JoyAI-LLM-Flash-GGUF with SGLang:

Install from pip and serve model

# Install SGLang from pip:
pip install sglang
# Start the SGLang server:
python3 -m sglang.launch_server \
    --model-path "jdopensource/JoyAI-LLM-Flash-GGUF" \
    --host 0.0.0.0 \
    --port 30000
# Call the server using curl (OpenAI-compatible API):
curl -X POST "http://localhost:30000/v1/chat/completions" \
	-H "Content-Type: application/json" \
	--data '{
		"model": "jdopensource/JoyAI-LLM-Flash-GGUF",
		"messages": [
			{
				"role": "user",
				"content": "What is the capital of France?"
			}
		]
	}'

Use Docker images

docker run --gpus all \
    --shm-size 32g \
    -p 30000:30000 \
    -v ~/.cache/huggingface:/root/.cache/huggingface \
    --env "HF_TOKEN=<secret>" \
    --ipc=host \
    lmsysorg/sglang:latest \
    python3 -m sglang.launch_server \
        --model-path "jdopensource/JoyAI-LLM-Flash-GGUF" \
        --host 0.0.0.0 \
        --port 30000
# Call the server using curl (OpenAI-compatible API):
curl -X POST "http://localhost:30000/v1/chat/completions" \
	-H "Content-Type: application/json" \
	--data '{
		"model": "jdopensource/JoyAI-LLM-Flash-GGUF",
		"messages": [
			{
				"role": "user",
				"content": "What is the capital of France?"
			}
		]
	}'

Ollama
How to use jdopensource/JoyAI-LLM-Flash-GGUF with Ollama:
```
ollama run hf.co/jdopensource/JoyAI-LLM-Flash-GGUF:IQ3_XS
```

Unsloth Studio

How to use jdopensource/JoyAI-LLM-Flash-GGUF with Unsloth Studio:

Install Unsloth Studio (macOS, Linux, WSL)

curl -fsSL https://unsloth.ai/install.sh | sh
# Run unsloth studio
unsloth studio -H 0.0.0.0 -p 8888
# Then open http://localhost:8888 in your browser
# Search for jdopensource/JoyAI-LLM-Flash-GGUF to start chatting

Install Unsloth Studio (Windows)

irm https://unsloth.ai/install.ps1 | iex
# Run unsloth studio
unsloth studio -H 0.0.0.0 -p 8888
# Then open http://localhost:8888 in your browser
# Search for jdopensource/JoyAI-LLM-Flash-GGUF to start chatting

Using HuggingFace Spaces for Unsloth

# No setup required
# Open https://huggingface.co/spaces/unsloth/studio in your browser
# Search for jdopensource/JoyAI-LLM-Flash-GGUF to start chatting

How to use jdopensource/JoyAI-LLM-Flash-GGUF with Pi:

Start the llama.cpp server

# Install llama.cpp:
brew install llama.cpp
# Start a local OpenAI-compatible server:
llama serve -hf jdopensource/JoyAI-LLM-Flash-GGUF:IQ3_XS

Configure the model in Pi

# Install Pi:
npm install -g @mariozechner/pi-coding-agent
# Add to ~/.pi/agent/models.json:
{
  "providers": {
    "llama-cpp": {
      "baseUrl": "http://localhost:8080/v1",
      "api": "openai-completions",
      "apiKey": "none",
      "models": [
        {
          "id": "jdopensource/JoyAI-LLM-Flash-GGUF:IQ3_XS"
        }
      ]
    }
  }
}

Run Pi

# Start Pi in your project directory:
pi

Hermes Agent new

How to use jdopensource/JoyAI-LLM-Flash-GGUF with Hermes Agent:

Start the llama.cpp server

# Install llama.cpp:
brew install llama.cpp
# Start a local OpenAI-compatible server:
llama serve -hf jdopensource/JoyAI-LLM-Flash-GGUF:IQ3_XS

Configure Hermes

# Install Hermes:
curl -fsSL https://hermes-agent.nousresearch.com/install.sh | bash
hermes setup
# Point Hermes at the local server:
hermes config set model.provider custom
hermes config set model.base_url http://127.0.0.1:8080/v1
hermes config set model.default jdopensource/JoyAI-LLM-Flash-GGUF:IQ3_XS

Run Hermes

hermes

Atomic Chat new

OpenClaw new

How to use jdopensource/JoyAI-LLM-Flash-GGUF with OpenClaw:

Start the llama.cpp server

# Install llama.cpp:
brew install llama.cpp
# Start a local OpenAI-compatible server:
llama serve -hf jdopensource/JoyAI-LLM-Flash-GGUF:IQ3_XS

Configure OpenClaw

# Install OpenClaw:
npm install -g openclaw@latest
# Register the local server and set it as the default model:
openclaw onboard --non-interactive --mode local \
  --auth-choice custom-api-key \
  --custom-base-url http://127.0.0.1:8080/v1 \
  --custom-model-id "jdopensource/JoyAI-LLM-Flash-GGUF:IQ3_XS" \
  --custom-provider-id llama-cpp \
  --custom-compatibility openai \
  --custom-text-input \
  --accept-risk \
  --skip-health

Run OpenClaw

openclaw agent --local --agent main --message "Hello from Hugging Face"

Docker Model Runner
How to use jdopensource/JoyAI-LLM-Flash-GGUF with Docker Model Runner:
```
docker model run hf.co/jdopensource/JoyAI-LLM-Flash-GGUF:IQ3_XS
```

Lemonade

How to use jdopensource/JoyAI-LLM-Flash-GGUF with Lemonade:

Pull the model

# Download Lemonade from https://lemonade-server.ai/
lemonade pull jdopensource/JoyAI-LLM-Flash-GGUF:IQ3_XS

Run and chat with the model

lemonade run user.JoyAI-LLM-Flash-GGUF-IQ3_XS

List all available models

lemonade list

Mingke977 commited on Feb 27

Commit

2e2822f

verified ·

1 Parent(s): d1b0e89

Add files using upload-large-folder tool

Browse files

Files changed (7) hide show

.gitattributes +5 -0
JoyAI-LLM-Flash-IQ3_XS.gguf +3 -0
JoyAI-LLM-Flash-IQ4_XS.gguf +3 -0
JoyAI-LLM-Flash-Q4_K_M.gguf +3 -0
JoyAI-LLM-Flash-Q8_0.gguf +3 -0
README.md +376 -0
figures/joyai-logo.png +3 -0

.gitattributes CHANGED Viewed

@@ -33,3 +33,8 @@ saved_model/**/* filter=lfs diff=lfs merge=lfs -text
 *.zip filter=lfs diff=lfs merge=lfs -text
 *.zst filter=lfs diff=lfs merge=lfs -text
 *tfevents* filter=lfs diff=lfs merge=lfs -text

 *.zip filter=lfs diff=lfs merge=lfs -text
 *.zst filter=lfs diff=lfs merge=lfs -text
 *tfevents* filter=lfs diff=lfs merge=lfs -text
+figures/joyai-logo.png filter=lfs diff=lfs merge=lfs -text
+JoyAI-LLM-Flash-IQ3_XS.gguf filter=lfs diff=lfs merge=lfs -text
+JoyAI-LLM-Flash-IQ4_XS.gguf filter=lfs diff=lfs merge=lfs -text
+JoyAI-LLM-Flash-Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
+JoyAI-LLM-Flash-Q8_0.gguf filter=lfs diff=lfs merge=lfs -text

JoyAI-LLM-Flash-IQ3_XS.gguf ADDED Viewed

	@@ -0,0 +1,3 @@

+version https://git-lfs.github.com/spec/v1
+oid sha256:9057e9015838e778cf20c7e44a4b898ad49c345b179d611e5dcbcd666c3516b1
+size 20086884320

JoyAI-LLM-Flash-IQ4_XS.gguf ADDED Viewed

	@@ -0,0 +1,3 @@

+version https://git-lfs.github.com/spec/v1
+oid sha256:31e5aa805870a085e88684524fb36c5c6f411c5ac4eca6422c72352fc9bb23e0
+size 26411919104

JoyAI-LLM-Flash-Q4_K_M.gguf ADDED Viewed

	@@ -0,0 +1,3 @@

+version https://git-lfs.github.com/spec/v1
+oid sha256:144d086e5b74ab5e08b060eba9af0cc4ec2efadc68535bfaa261557f68fc5115
+size 29669242624

JoyAI-LLM-Flash-Q8_0.gguf ADDED Viewed

	@@ -0,0 +1,3 @@

+version https://git-lfs.github.com/spec/v1
+oid sha256:b4a95b59201f76b3f1a26eacac541574427a993e008adc5a92e8e4cfda9e9a0b
+size 52067555072

README.md ADDED Viewed

	@@ -0,0 +1,376 @@

+---
+language:
+- zh
+- en
+pipeline_tag: text-generation
+library_name: transformers
+---
+<div align="center">
+  <picture>
+      <img src="figures/joyai-logo.png" width="30%" alt="JoyAI-LLM Flash">
+  </picture>
+</div>
+<hr>
+<div align="center" style="line-height: 1;">
+  <a href="https://huggingface.co/jdopensource" target="_blank"><img alt="Hugging Face" src="https://img.shields.io/badge/%F0%9F%A4%97%20Hugging%20Face-JD-ffc107?color=ffc107&logoColor=white"/></a>
+  <a href="https://huggingface.co/jdopensource/JoyAI-LLM-Flash/blob/main/LICENSE"><img alt="License" src="https://img.shields.io/badge/License-Modified_MIT-f5de53?&color=f5de53"/></a>
+</div>
+## 1. Model Introduction
+JoyAI-LLM-Flash is a state-of-the-art medium-sized instruct language model with 3 billion activated parameters and 48 billion total parameters. JoyAI-LLM-Flash was pretrained on 20 trillion text tokens using Muon optimizer, followed by large-scale supervised fine-tuning (SFT), direct preference optimization (DPO), and reinforcement learning (RL) across diverse environments. JoyAI-LLM-Flash achieves strong performance across frontier knowledge, reasoning, coding tasks and agentic capabilities.
+### Key Features
+- Fiber Bundle RL: Introduces fiber bundle theory into reinforcement learning, proposing a novel optimization framework, FiberPO. This method is specifically designed to handle the challenges of large-scale and heterogeneous agent training, improving stability and robustness under complex data distributions.
+- Training-Inference Collaboration: apply Muon optimizer with dense MTP, develop novel optimization techniques to resolve instabilities while scaling up, delivering 1.3× to 1.7× the throughput of the non-MTP version.
+- Agentic Intelligence: designed for tool use, reasoning, and autonomous problem-solving.
+## 2. Model Summary
+|                                             |                          |
+| :-----------------------------------------: | :----------------------: |
+|              **Architecture**               | Mixture-of-Experts (MoE) |
+|            **Total Parameters**             |           48B            |
+|          **Activated Parameters**           |            3B            |
+| **Number of Layers** (Dense layer included) |            40            |
+|         **Number of Dense Layers**          |            1             |
+|       **Attention Hidden Dimension**        |           2048           |
+|    **MoE Hidden Dimension** (per Expert)    |           768            |
+|        **Number of Attention Heads**        |            32            |
+|            **Number of Experts**            |           256            |
+|       **Selected Experts per Token**        |            8             |
+|        **Number of Shared Experts**         |            1             |
+|             **Vocabulary Size**             |           129K           |
+|             **Context Length**              |           128K           |
+|           **Attention Mechanism**           |           MLA            |
+|           **Activation Function**           |          SwiGLU          |
+|                   </div>                    |                          |
+## 3. Evaluation Results
+<table>
+<thead>
+<tr>
+<th align="center">Benchmark</th>
+<th align="center"><sup>JoyAI-LLM Flash</sup></th>
+<th align="center"><sup>Qwen3-30B-A3B-Instuct-2507</sup></th>
+<th align="center"><sup>GLM-4.7-Flash<br>(Non-thinking)</sup></th>
+</tr>
+</thead>
+<tbody>
+<tr>
+<td align="center" colspan=8><strong>Knowledge &amp; Alignment</strong></td>
+</tr>
+<tr>
+<td align="center" style="vertical-align: middle">MMLU</td>
+<td align="center" style="vertical-align: middle"><strong>89.50</strong></td>
+<td align="center" style="vertical-align: middle">86.87</td>
+<td align="center" style="vertical-align: middle">80.53</td>
+</tr>
+<tr>
+<td align="center" style="vertical-align: middle">MMLU-Pro</td>
+<td align="center" style="vertical-align: middle"><strong>81.02</strong></td>
+<td align="center" style="vertical-align: middle">73.88</td>
+<td align="center" style="vertical-align: middle">63.62</td>
+</tr>
+<tr>
+<td align="center" style="vertical-align: middle">CMMLU</td>
+<td align="center" style="vertical-align: middle"><strong>87.03</strong></td>
+<td align="center" style="vertical-align: middle">85.88</td>
+<td align="center" style="vertical-align: middle">75.85</td>
+</tr>
+<tr>
+<td align="center" style="vertical-align: middle">GPQA-Diamond</td>
+<td align="center" style="vertical-align: middle"><strong>74.43</strong></td>
+<td align="center" style="vertical-align: middle">68.69</td>
+<td align="center" style="vertical-align: middle">39.90</td>
+</tr>
+<tr>
+<td align="center" style="vertical-align: middle">SuperGPQA</td>
+<td align="center" style="vertical-align: middle"><strong>55.00</strong></td>
+<td align="center" style="vertical-align: middle">52.00</td>
+<td align="center" style="vertical-align: middle">32.00</td>
+</tr>
+<tr>
+<td align="center" style="vertical-align: middle">LiveBench</td>
+<td align="center" style="vertical-align: middle"><strong>72.90</strong></td>
+<td align="center" style="vertical-align: middle">59.70</td>
+<td align="center" style="vertical-align: middle">43.10</td>
+</tr>
+<tr>
+<td align="center" style="vertical-align: middle">IFEval</td>
+<td align="center" style="vertical-align: middle"><strong>86.69</strong></td>
+<td align="center" style="vertical-align: middle">83.18</td>
+<td align="center" style="vertical-align: middle">82.44</td>
+</tr>
+<tr>
+<td align="center" style="vertical-align: middle">AlignBench</td>
+<td align="center" style="vertical-align: middle"><strong>8.24</strong></td>
+<td align="center" style="vertical-align: middle">8.07</td>
+<td align="center" style="vertical-align: middle">6.85</td>
+</tr>
+<tr>
+<td align="center" style="vertical-align: middle">HellaSwag</td>
+<td align="center" style="vertical-align: middle"><strong>91.79</strong></td>
+<td align="center" style="vertical-align: middle">89.90</td>
+<td align="center" style="vertical-align: middle">60.84</td>
+</tr>
+<tr>
+<td align="center" colspan=8><strong>Coding</strong></td>
+</tr>
+<tr>
+<td align="center" style="vertical-align: middle">HumanEval</td>
+<td align="center" style="vertical-align: middle"><strong>96.34</strong></td>
+<td align="center" style="vertical-align: middle">95.12</td>
+<td align="center" style="vertical-align: middle">74.39</td>
+</tr>
+<tr>
+<td align="center" style="vertical-align: middle">LiveCodeBench</td>
+<td align="center" style="vertical-align: middle"><strong>65.60</strong></td>
+<td align="center" style="vertical-align: middle">39.71</td>
+<td align="center" style="vertical-align: middle">27.43</td>
+</tr>
+<tr>
+<td align="center" style="vertical-align: middle">SciCode</td>
+<td align="center" style="vertical-align: middle"><strong>3.08/22.92</strong></td>
+<td align="center" style="vertical-align: middle"><strong>3.08/22.92</strong></td>
+<td align="center" style="vertical-align: middle">3.08/15.11</td>
+</tr>
+<tr>
+<td align="center" colspan=8><strong>Mathematics</strong></td>
+</tr>
+<tr>
+<td align="center" style="vertical-align: middle">GSM8K</td>
+<td align="center" style="vertical-align: middle"><strong>95.83</strong></td>
+<td align="center" style="vertical-align: middle">79.83</td>
+<td align="center" style="vertical-align: middle">81.88</td>
+</tr>
+<tr>
+<td align="center" style="vertical-align: middle">AIME2025</td>
+<td align="center" style="vertical-align: middle"><strong>65.83</strong></td>
+<td align="center" style="vertical-align: middle">62.08</td>
+<td align="center" style="vertical-align: middle">24.17</td>
+</tr>
+<tr>
+<td align="center" style="vertical-align: middle">MATH 500</td>
+<td align="center" style="vertical-align: middle"><strong>97.10</strong></td>
+<td align="center" style="vertical-align: middle">89.80</td>
+<td align="center" style="vertical-align: middle">90.90</td>
+</tr>
+<tr>
+<td align="center" colspan=8><strong>Agentic</strong></td>
+</tr>
+<tr>
+<td align="center" style="vertical-align: middle">SWE-bench Verified</td>
+<td align="center" style="vertical-align: middle"><strong>60.60</strong></td>
+<td align="center" style="vertical-align: middle">24.44</td>
+<td align="center" style="vertical-align: middle">51.60</td>
+</tr>
+<tr>
+<td align="center" style="vertical-align: middle">Tau2-Retail</td>
+<td align="center" style="vertical-align: middle"><strong>67.55</strong></td>
+<td align="center" style="vertical-align: middle">53.51</td>
+<td align="center" style="vertical-align: middle">62.28</td>
+</tr>
+<tr>
+<td align="center" style="vertical-align: middle">Tau2-Airline</td>
+<td align="center" style="vertical-align: middle"><strong>54.00</strong></td>
+<td align="center" style="vertical-align: middle">32.00</td>
+<td align="center" style="vertical-align: middle">52.00</td>
+</tr>
+<tr>
+<td align="center" style="vertical-align: middle">Tau2-Telecom</td>
+<td align="center" style="vertical-align: middle">79.83</td>
+<td align="center" style="vertical-align: middle">4.39</td>
+<td align="center" style="vertical-align: middle"><strong>88.60</strong></td>
+</tr>
+<tr>
+<td align="center" colspan=8><strong>Long Context</strong></td>
+</tr>
+<tr>
+<td align="center" style="vertical-align: middle">RULER</td>
+<td align="center" style="vertical-align: middle"><strong>95.60</strong></td>
+<td align="center" style="vertical-align: middle">89.66</td>
+<td align="center" style="vertical-align: middle">56.12</td>
+</tr>
+</tbody>
+</table>
+## 4. Deployment
+> [!Note]
+> You can access JoyAI-LLM Flash API on https://docs.jdcloud.com/cn/jdaip/chat and we provide OpenAI/Anthropic-compatible API for you.
+> Currently, JoyAI-LLM-Flash-GGUF is recommended to run on the following inference engines:
+* Llama.cpp
+* Ollama
+## 5. Model Usage
+The usage demos below demonstrate how to call our official API.
+For third-party APIs deployed with vLLM or SGLang, please note that:
+> [!Note] Recommended sampling parameters: `temperature=0.6`, `top_p=1.0`
+### Chat Completion
+This is a simple chat completion script which shows how to call JoyAI-Flash API.
+```python
+from openai import OpenAI
+client = OpenAI(base_url="http://IP:PORT/v1", api_key="EMPTY")
+def simple_chat(client: OpenAI):
+    messages = [
+        {
+            "role": "user",
+            "content": [
+                {
+                    "type": "text",
+                    "text": "which one is bigger, 9.11 or 9.9? think carefully.",
+                }
+            ],
+        },
+    ]
+    model_name = client.models.list().data[0].id
+    response = client.chat.completions.create(
+        model=model_name, messages=messages, stream=False, max_tokens=4096
+    )
+    print(f"response: {response.choices[0].message.content}")
+if __name__ == "__main__":
+    simple_chat(client)
+```
+### Tool call Completion
+This is a simple toll call completion script which shows how to call JoyAI-Flash API.
+```python
+import json
+from openai import OpenAI
+client = OpenAI(base_url="http://IP:PORT/v1", api_key="EMPTY")
+def my_calculator(expression: str) -> str:
+    return str(eval(expression))
+def rewrite(expression: str) -> str:
+    return str(expression)
+def simple_tool_call(client: OpenAI):
+    messages = [
+        {
+            "role": "user",
+            "content": [
+                {
+                    "type": "text",
+                    "text": "use my functions to compute the results for the equations: 6+1",
+                },
+            ],
+        },
+    ]
+    tools = [
+        {
+            "type": "function",
+            "function": {
+                "name": "my_calculator",
+                "description": "A calculator that can evaluate a mathematical equation and compute its results.",
+                "parameters": {
+                    "type": "object",
+                    "properties": {
+                        "expression": {
+                            "type": "string",
+                            "description": "The mathematical expression to evaluate.",
+                        },
+                    },
+                    "required": ["expression"],
+                },
+            },
+        },
+        {
+            "type": "function",
+            "function": {
+                "name": "rewrite",
+                "description": "Rewrite a given text for improved clarity",
+                "parameters": {
+                    "type": "object",
+                    "properties": {
+                        "text": {
+                            "type": "string",
+                            "description": "The input text to rewrite",
+                        }
+                    },
+                },
+            },
+        },
+    ]
+    model_name = client.models.list().data[0].id
+    response = client.chat.completions.create(
+        model=model_name,
+        messages=messages,
+        temperature=1.0,
+        max_tokens=1024,
+        tools=tools,
+        tool_choice="auto",
+    )
+    tool_calls = response.choices[0].message.tool_calls
+    results = []
+    for tool_call in tool_calls:
+        function_name = tool_call.function.name
+        function_args = tool_call.function.arguments
+        if function_name == "my_calculator":
+            result = my_calculator(**json.loads(function_args))
+            results.append(result)
+    messages.append({"role": "assistant", "tool_calls": tool_calls})
+    for tool_call, result in zip(tool_calls, results):
+        messages.append(
+            {
+                "role": "tool",
+                "tool_call_id": tool_call.id,
+                "name": tool_call.function.name,
+                "content": result,
+            }
+        )
+    response = client.chat.completions.create(
+        model=model_name,
+        messages=messages,
+        temperature=1.0,
+        max_tokens=1024,
+    )
+    print(response.choices[0].message.content)
+if __name__ == "__main__":
+    simple_tool_call(client)
+```
+---
+## 6. License
+Both the code repository and the model weights are released under the [Modified MIT License](LICENSE).

figures/joyai-logo.png ADDED Viewed

Git LFS Details

SHA256: 4ea9d6a20a7707ca8dc427d6dcb5db6e2489f7730d5bffea26d8db20b1c54365
Pointer size: 131 Bytes
Size of remote file: 250 kB